Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

383 results

PaperCategoryAuthorsDate
An Empirical Study of Harness Design for Coding Agents ↗ cs.AI 9 Sep 17, 2026
RAFT: A Stateful Retrieval-Augmented Framework for Troubleshooti... ↗ cs.AI 7 Sep 17, 2026
Q&A on Any Spreadsheet Requires Interpreting Its Grid Structure ↗ cs.AI 1 Sep 17, 2026
Deep Noir: Autonomous Steering Discovery via Architectural Chron... ↗ cs.AI 5 Sep 17, 2026
Ownership in AI-Assisted Everyday Tasks ↗ cs.AI 7 Sep 17, 2026
PAA: The Probabilistic Allen Algebra: A Generative and Complete... ↗ cs.AI 1 Sep 17, 2026
Limits of Confidence in Diffusion ↗ cs.AI 4 Sep 17, 2026
Language-model groups overstate consensus when replaying human d... ↗ cs.AI 1 Sep 17, 2026
Refuse, Decompose, Refresh: A Claim-Safe Protocol for Closed-Loo... ↗ cs.AI 2 Sep 17, 2026
FreqCondNorm: Towards Cross-domain Predictive Maintenance throug... ↗ cs.AI 3 Sep 17, 2026
SoL-Pi: Recursively Scaling Auto-Research Loops for Efficient Ag... ↗ cs.AI 14 Sep 17, 2026
How Do Agent Harnesses Create Value? Planning Information and Re... ↗ cs.AI 3 Sep 17, 2026
SkillAA: Attribution-Guided Skill-Graph Updating with Targeted V... ↗ cs.AI 3 Sep 17, 2026
The Organization of Inference: Information, Resource Constraints... ↗ cs.AI 3 Sep 17, 2026
Generating Heterogeneous 3D Geological Microstructures from 2D I... ↗ cs.AI 4 Sep 17, 2026
A Qualitative Model for Reasoning about Path and Support ↗ cs.AI 2 Sep 17, 2026
Structured Four-Stage Legal Translation: From Natural-Language T... ↗ cs.AI 4 Sep 17, 2026
NeuSOGA3D: A Neuro-Symbolic Framework for Explainable 3D Geometr... ↗ cs.AI 4 Sep 17, 2026
MTVA-Bench: Evaluating the Language Model Inside Cascaded Voice... ↗ cs.AI 4 Sep 17, 2026
WikiSkill: Compiling Agent Experience into Persistent Knowledge... ↗ cs.AI 6 Aug 27, 2026
Solving reactions as movements of electrons — making the mechani... Explained cs.AI 4 Aug 27, 2026
Learning a Continuous Sepsis Severity Score Without Hour-by-Hour... ↗ cs.AI 21 Aug 27, 2026
CorporateBench: Large-Scale Q&A Benchmarking with Temporal Knowl... ↗ cs.AI 7 Aug 27, 2026
Sophistication in GenAI Use: Field Evidence from a Large Firm ↗ cs.AI 4 Aug 27, 2026
Not All Eval-Awareness Is Equal: Capabilities Framing Predicts C... ↗ cs.AI 2 Aug 27, 2026
Verify Smarter, Evolve Further: Efficient Harness Evolution thro... ↗ cs.AI 6 Aug 27, 2026
LLMs Can Design Near-Optimal OR Algorithms ↗ cs.AI 1 Aug 27, 2026
BrailleBench: Investigating Multi-Criteria Braille Comprehension... ↗ cs.AI 6 Aug 27, 2026
Naive Prompt Optimization: Rethinking the Need for Complex Promp... ↗ cs.AI 2 Aug 27, 2026
What Makes Good Agentic Data? An ACE Lens on Data Generation for... ↗ cs.AI 14 Aug 27, 2026
Calibrated Enough to Know, Not Calibrated to Act: Fabricated Evi... ↗ cs.AI 1 Aug 27, 2026
BPMN4CAI: A BPMN Extension for Modeling Dynamic Conversational A... ↗ cs.AI 3 Aug 27, 2026
Thomson: Continual Learning of Frontier Models for SovereignAI ↗ cs.AI 26 Aug 27, 2026
When Tool Outputs Become Commands: Separating Action Induction f... ↗ cs.AI 8 Aug 27, 2026
Feature Transformation Enhanced Jacobi Polynomial Graph Filterin... ↗ cs.AI 3 Aug 27, 2026
GRAIN: Bridging Name and Narrative Shifts in Real-World Graph Re... ↗ cs.AI 13 Aug 27, 2026
TransMeme: A Multi-Agent Framework for Cross-Cultural Meme Trans... ↗ cs.AI 7 Aug 27, 2026
LAAF: A Layered Accountability Architecture Framework for LLM Ap... ↗ cs.AI 8 Aug 27, 2026
pro-team at LLMs4OL 2026 Tasks Flagship and Reuse: Retrieval-Aug... ↗ cs.AI 5 Aug 27, 2026
A Contract-Centered Architecture for Scalable and Manageable Age... ↗ cs.AI 6 Aug 27, 2026
Omni-Interactive Universal Embedder ↗ cs.AI 7 Aug 27, 2026
A Multi-Modal AI Framework for Real-Time Queue Prediction, Manag... ↗ cs.AI 5 Aug 27, 2026
ASIL: Replacing Screenshot-and-Click with Structured State and S... ↗ cs.AI 2 Aug 27, 2026
An Agentic Approach for Active Data Collection, Travel Behavior... ↗ cs.AI 5 Aug 20, 2026
AI4AI-Bench: Benchmarking LLM Agents in Algorithmic Design for R... ↗ cs.AI 10 Aug 20, 2026
Pandora's AI Model Routing Box: Efficient Allocation with Costly... ↗ cs.AI 8 Aug 20, 2026
MidTool: Mid-training Data Synthesis for Agentic Tool Use ↗ cs.AI 8 Aug 20, 2026
Phantom Gains: Auditing Self-Improvement Against a Measured Null ↗ cs.AI 4 Aug 20, 2026
Break It Down, Pass It On: Cross-Task Skill Transfer in LLM Agen... ↗ cs.AI 4 Aug 20, 2026
Catching the Rug: Early Prediction of Fraudulent Memecoins on So... ↗ cs.AI 5 Aug 20, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.