Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

235 results

PaperCategoryAuthorsDate
Where Do CoT Training Gains Land in LLM based Agents? ↗ cs.AI 5 Jun 25, 2026
Diagnosing Task Insensitivity in Language Agents ↗ cs.AI 5 Jun 25, 2026
Learning to Recover Task Experts from a Multi-Task Merged Model ↗ cs.AI 4 Jun 25, 2026
Generative Retrieval via Diffusion Transformer with Metric-Order... ↗ cs.AI 11 Jun 25, 2026
Toward Calibrated Mixture-of-Experts Under Distribution Shift ↗ cs.AI 5 Jun 18, 2026
How Do Instructions Shape Speech? Cross-Attention Attribution fo... ↗ cs.AI 7 Jun 18, 2026
LedgerAgent: Structured State for Policy-Adherent Tool-Calling A... ↗ cs.AI 4 Jun 18, 2026
DeepSWIP: Quotient-WMC Counterfactuals for Neural Probabilistic... ↗ cs.AI 3 Jun 18, 2026
FlowEdit: Associative Memory for Lifelong Pronunciation Adaptati... ↗ cs.AI 3 Jun 18, 2026
Multi-LCB: Extending LiveCodeBench to Multiple Programming Langu... ↗ cs.AI 8 Jun 18, 2026
What Do Safety-Aligned LLMs Learn From Mixed Compliance Demonstr... ↗ cs.AI 2 Jun 18, 2026
Context-Aware Hierarchical Bayesian Modeling of IVF Laboratory E... ↗ cs.AI 5 Jun 18, 2026
Interpretable Sperm Morphology Classification via Attention-Guid... ↗ cs.AI 4 Jun 18, 2026
Rethinking Shrinkage Bias in LLM FP4 Pretraining: Geometric Orig... ↗ cs.AI 12 Jun 18, 2026
Automating SKILL.md Generation for Computer-Using Agents via Int... ↗ cs.AI 2 Jun 18, 2026
SoftSkill: Behavioral Compression for Contextual Adaptation ↗ cs.AI 9 Jun 18, 2026
Leveraging systems' non-linearity to tackle the scarcity of data... ↗ cs.AI 4 Jun 18, 2026
Lagrange: An Open-Vocabulary, Energy-Based Sparse Framework for... ↗ cs.AI 4 Jun 18, 2026
Confidence-Aware Automated Assessment of Student-Drawn Scientifi... ↗ cs.AI 6 Jun 18, 2026
Navigating Unreliable Parametric and Contextual Knowledge: Expli... ↗ cs.AI 5 Jun 18, 2026
A Multi-Agent system for Multi-Objective constrained optimizatio... ↗ cs.AI 1 Jun 18, 2026
Thermodynamic Measure of Intelligence ↗ cs.AI 1 Jun 18, 2026
QMFOL: Benchmarking Large Language Model Reasoning via Quantifia... ↗ cs.AI 6 Jun 18, 2026
Augmenting Game AI with Deep Reinforcement Learning ↗ cs.AI 6 Jun 18, 2026
Beyond Accuracy: Measuring Logical Compliance of Predictive Mode... ↗ cs.AI 4 Jun 18, 2026
Apparent Psychological Profiles of Large Language Models are Lar... ↗ cs.AI 3 Jun 18, 2026
Implicit Semantic-Aware Communication Based on Hypergraph Reason... ↗ cs.AI 5 Jun 18, 2026
Modularity-Free Conflict-Averse Training for Generalized PINNs ↗ cs.AI 4 Jun 18, 2026
BIM-Edit: Benchmarking Large Language Models for IFC-Based Build... ↗ cs.AI 7 Jun 18, 2026
RACL: Reasoning-Agent Control Layers for Continuous Metaheuristi... ↗ cs.AI 1 Jun 18, 2026
Learning to Prompt: Improving Student Engagement with Adaptive L... ↗ cs.AI 4 Jun 18, 2026
Automated reproducibility assessments in the social and behavior... ↗ cs.AI 10 Jun 11, 2026
Agents-K1: Towards Agent-native Knowledge Orchestration ↗ cs.AI 25 Jun 11, 2026
EurekAgent: Agent Environment Engineering is All You Need For Au... ↗ cs.AI 8 Jun 11, 2026
Before You Think: System 0, AI-Mediated Cognition and Cognitive... ↗ cs.AI 4 Jun 11, 2026
Beyond Runtime Enforcement: Shield Synthesis as Defensibility An... ↗ cs.AI 2 Jun 11, 2026
AgentBeats: Agentifying Agent Assessment for Openness, Standardi... ↗ cs.AI 29 Jun 11, 2026
Reasoning as Pattern Matching: Shared Mechanisms in Human and LL... ↗ cs.AI 2 Jun 11, 2026
Multi-Agent Reinforcement Learning from Delayed Marketplace Feed... ↗ cs.AI 3 Jun 11, 2026
EpiBench: Verifiable Evaluation of AI Agents on Epigenomics Anal... ↗ cs.AI 5 Jun 11, 2026
Reward Modeling for Multi-Agent Orchestration ↗ cs.AI 8 Jun 11, 2026
Multiagent Protocols with Aggregated Confidence Signals ↗ cs.AI 2 Jun 11, 2026
A Three-Layer Framework for AI in Scientific Discovery ↗ cs.AI 1 Jun 11, 2026
Is It You or Your Environment? A Bayesian Inference Framework fo... ↗ cs.AI 2 Jun 11, 2026
Uncertainty-Aware Hybrid Retrieval for Long-Document RAG ↗ cs.AI 2 Jun 11, 2026
CloudCons: A Comprehensive End-to-End Benchmark for Cloud Resour... ↗ cs.AI 9 Jun 11, 2026
Why Sampling Is Not Choosing: Intentionality, Agency, and Moral... ↗ cs.AI 1 Jun 11, 2026
Evaluation Sovereignty in Metadata-Driven Classification: A Mult... ↗ cs.AI 1 Jun 11, 2026
Optimizing Appliance Scheduling for Solar Energy Management Usin... ↗ cs.AI 4 Jun 11, 2026
Neuro-Symbolic Agents for Regulated Process Automation: Challeng... ↗ cs.AI 3 Jun 11, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.