Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

383 results

PaperCategoryAuthorsDate
vToken: Token-Level Virtualization for Reclaimable KV Caches ↗ cs.AI 7 Aug 13, 2026
Capability Sheaves for Compositional Agent-Harness Repair: Contr... ↗ cs.AI 1 Aug 13, 2026
TsuGO: Probing Search Efficiency in LLM Reasoning via Go Life-an... ↗ cs.AI 7 Aug 13, 2026
Teach the Magnitude, Not the Direction: Verifier-Bounded Credit... ↗ cs.AI 6 Aug 13, 2026
SkillShapley: Boundary-Adaptive Shapley Valuation for Skill Step... ↗ cs.AI 6 Aug 13, 2026
Rethinking Normalization Placement for LLMs: Post-Norm under Cur... ↗ cs.AI 10 Aug 13, 2026
Numeracy in Large Language Models: Fundamental Limitations and P... ↗ cs.AI 1 Aug 13, 2026
SkillEvo: Self-Renewing Evolution Gradients from Multi-Turn Inte... ↗ cs.AI 6 Aug 13, 2026
Robust Dempster-Shafer Evidence Fusion with Chaos-Conflict Measu... ↗ cs.AI 6 Aug 13, 2026
Tracing the Heart: An Evidence-Linked Pipeline for Heart-Failure... ↗ cs.AI 13 Aug 6, 2026
The Low Frequency Trap: Video Language Models Fail at Simple Eve... ↗ cs.AI 8 Aug 6, 2026
Challenges in Evaluating Explanation Methods for Static and Evol... ↗ cs.AI 1 Aug 6, 2026
TRAJDEBUG: Tracing Error Lifecycle to Identify Critical Failures... ↗ cs.AI 13 Aug 6, 2026
Beyond Top-K: Replacing Black-Box Retrieval with Interpretable A... ↗ cs.AI 3 Aug 6, 2026
HarnessOpt-Bench: Evaluating LLMs at Harness Optimization ↗ cs.AI 7 Aug 6, 2026
Bias Analysis of L2 Speaking Assessment Systems Using Concept Ac... ↗ cs.AI 3 Aug 6, 2026
QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agenti... ↗ cs.AI 6 Aug 6, 2026
The Illusion of Visual Tool-Use: A Causal Audit of Thinking with... ↗ cs.AI 4 Aug 6, 2026
Improving the Realism of Synthetic Clinical Benchmarks Under Uti... ↗ cs.AI 9 Aug 6, 2026
DASH: Divergence-Adaptive Supervision Horizons for On-Policy Sel... ↗ cs.AI 12 Aug 6, 2026
TS-RAG: Retrieval Augmented Generation for Time Series Forecasti... ↗ cs.AI 3 Aug 6, 2026
EnvACE: Internalizing Environment Dynamics via World Rehearsal f... ↗ cs.AI 12 Aug 6, 2026
Comparative Approaches to Agent Retrieval over Large Skill Libra... ↗ cs.AI 2 Aug 6, 2026
MicroEvo: Knowledge-Guided LLM Sampling for Efficient Microarchi... ↗ cs.AI 14 Aug 6, 2026
Schema-Guided Hierarchical Information Extraction and Semantic E... ↗ cs.AI 6 Aug 6, 2026
iARCS: Iterative Agentic RL for Controllable 3D Scene Generation ↗ cs.AI 5 Aug 6, 2026
CogVis: Must Open-Vocabulary Change Detection Perceive the Scene... ↗ cs.AI 3 Aug 6, 2026
PaDoc: Layout-Grounded Parallel Decoding for Document Parsing ↗ cs.AI 11 Aug 6, 2026
FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents... ↗ cs.AI 9 Aug 6, 2026
Contextual Information Policy Optimization for Search Agents ↗ cs.AI 4 Aug 6, 2026
Poli-Bias: Understanding and Measuring Large Language Model Bias... ↗ cs.AI 4 Aug 6, 2026
Mind the Gaps: Mixture-of-Minds for Human Simulation ↗ cs.AI 1 Aug 6, 2026
From Siloed Algorithms to Compliance-First Agentic Platforms: A... ↗ cs.AI 3 Aug 6, 2026
ECHO: A Locally-Deployable Agentic Health Assistant with Tempora... ↗ cs.AI 7 Aug 6, 2026
Evaluating Investment Logic in Large Language Models: A Real-Wor... ↗ cs.AI 7 Aug 6, 2026
Signal or Spurious Cue? A Randomized Audit of Survey-Country Met... ↗ cs.AI 4 Aug 6, 2026
When History Lies: Evaluating and Improving Tool Use under Misle... ↗ cs.AI 4 Aug 6, 2026
Integrating Implicit and Explicit Relational Biases through Grap... ↗ cs.AI 4 Aug 6, 2026
From Economic Agents to Agentic Economies: A Systems Blueprint f... ↗ cs.AI 10 Aug 6, 2026
HERALD: Counterfactual Audits and Minimal Repairs for Proof-of-R... ↗ cs.AI 5 Aug 6, 2026
Hybrid Machine Learning Framework for Herd-Level Cattle Growth P... ↗ cs.AI 5 Aug 6, 2026
OPERA: Operator-residual feedback for reliable autonomous optica... ↗ cs.AI 7 Aug 6, 2026
AgentOPSD: Recursive Self-Distillation for Agentic Reinforcement... ↗ cs.AI 13 Aug 6, 2026
Temporal Bridges for Spatial Resolution: Enhancing Climate Data... ↗ cs.AI 4 Aug 6, 2026
Stability of Ranking-dependent Pair-wise Comparison Patterns in... ↗ cs.AI 3 Aug 6, 2026
Training a Conditioned Video Game Agent on a VLM Annotated Datas... ↗ cs.AI 2 Aug 6, 2026
VLMs for Videogame Data Annotation ↗ cs.AI 2 Aug 6, 2026
GAUGE: A Measurement-Grounded Benchmark for Physical Fidelity in... ↗ cs.AI 15 Aug 6, 2026
AISPA: User-Centric System Prompt Auditing for Large Language Mo... ↗ cs.AI 26 Jul 30, 2026
OSReward: Instituting Standardized Evaluation for Cross-Platform... ↗ cs.AI 23 Jul 30, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.