Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

235 results

PaperCategoryAuthorsDate
AISPA: User-Centric System Prompt Auditing for Large Language Mo... ↗ cs.AI 26 Jul 30, 2026
OSReward: Instituting Standardized Evaluation for Cross-Platform... ↗ cs.AI 23 Jul 30, 2026
DualG-MRAG: Decoupling Macro-Reasoning and Micro-Matching for Mu... ↗ cs.AI 5 Jul 30, 2026
Rethinking Inference-Time Scaling in Local Computer-Use Agents:... ↗ cs.AI 2 Jul 30, 2026
MANTA: Multi-Agent Network Topology Adaptation for Self-Evolving... ↗ cs.AI 6 Jul 30, 2026
Selective Credibility-Limited Belief Update ↗ cs.AI 2 Jul 30, 2026
InfoOps Bench: A live information operations safety benchmark ↗ cs.AI 4 Jul 30, 2026
SCOPE: Supply-Chain Operations through Coupled Policies for End-... ↗ cs.AI 7 Jul 30, 2026
A Fuzzy Rule-based Neuro-Symbolic Approach for Pipe Severity Pre... ↗ cs.AI 3 Jul 30, 2026
Towards Autonomous Aircraft Surveillance from Nanosatellites thr... ↗ cs.AI 4 Jul 30, 2026
A report-grounded vision-language foundation model for colonosco... ↗ cs.AI 15 Jul 30, 2026
LeanCSP: A Framework for Certifying Constraint Reformulation and... ↗ cs.AI 2 Jul 30, 2026
SVR: Self-Verifying Refinement via Joint Verdict-Confidence Rein... ↗ cs.AI 3 Jul 30, 2026
Metaphor Tracer: A Theory-Informed Analysis of Hidden States ↗ cs.AI 5 Jul 30, 2026
A foundation model of numerical intelligence with cross-discipli... ↗ cs.AI 3 Jul 30, 2026
When Derived Measurements Mislead: Quantifying and Mitigating LL... ↗ cs.AI 9 Jul 30, 2026
WIDE: Boosting Adaptive LLM Inference via Token-level Dynamic Wi... ↗ cs.AI 6 Jul 30, 2026
QuantWAMs: Calibrating at the Right Granularity for World Action... ↗ cs.AI 7 Jul 30, 2026
GLM-RAG: Graph Language Models for Graph-Based Retrieval-Augment... ↗ cs.AI 5 Jul 30, 2026
When Specifications Conflict: A Symmetry-Based Framework for Mea... ↗ cs.AI 4 Jul 30, 2026
HyperClaim: Fine-Grained Cross-Modal Hypergraph Reasoning for Vi... ↗ cs.AI 5 Jul 30, 2026
How Benchmarks Mis-Score Computer-Use Agents ↗ cs.AI 8 Jul 30, 2026
Correcting What You Cannot See: Credit Assignment for Perception... ↗ cs.AI 3 Jul 30, 2026
Paying for Honesty Without Knowing the Truth: Reputation-Penalty... ↗ cs.AI 8 Jul 30, 2026
PathView-Bench: Can Multimodal Large Language Models Achieve Fin... ↗ cs.AI 4 Jul 30, 2026
One Human, $N$ Agents: Audit-Budget Allocation for LLM Agent Fle... ↗ cs.AI 3 Jul 30, 2026
Tycho: Active Abstraction with Programmatic World Models for ARC... ↗ cs.AI 3 Jul 30, 2026
MemHarness: Memory Is Reconstructed, Not Replayed ↗ cs.AI 13 Jul 30, 2026
LLM-Guided Evolutionary Search for Constraint Model Reformulatio... ↗ cs.AI 4 Jul 30, 2026
Operationally Guided Placement-Aware Learning for Industrial Onl... ↗ cs.AI 3 Jul 30, 2026
AI and Authenticity in Islamic Research: A Critical Evaluation o... ↗ cs.AI 1 Jul 30, 2026
CDAE: Enhancing Perturbation Robustness in Pretrained Language M... ↗ cs.AI 4 Jul 30, 2026
Qwen-UI-Agent Technical Report: Toward Next-Generation Real-Worl... ↗ cs.AI 16 Jul 30, 2026
Old Tricks, New Models: How Simple Image Transformations Break M... ↗ cs.AI 5 Jul 30, 2026
Unsupervised Consensus-Based Anomaly Detection for Spatiotempora... ↗ cs.AI 2 Jul 23, 2026
Beyond Sycophancy: Structured Resistance and Compliance in LLM M... ↗ cs.AI 2 Jul 23, 2026
OpenForgeRL: Train Harness-native Agents in Any Environment ↗ cs.AI 10 Jul 23, 2026
MIRROR: Learning from the Other View for Multi-Modal Reasoning ↗ cs.AI 4 Jul 23, 2026
The Boundaries of Automation: A Theory of Persistent Human Parti... ↗ cs.AI 4 Jul 23, 2026
Same Dangerous Objective, Opposite Advice: Direct Exposure versu... ↗ cs.AI 1 Jul 23, 2026
Agentic Context Management: Solving Agent Memory and Cost by Tre... ↗ cs.AI 1 Jul 23, 2026
Toward Continuous Assurance for the Democratization of AI Agent... ↗ cs.AI 2 Jul 23, 2026
Agentic coding without the cloud: evaluating open-weight large l... ↗ cs.AI 7 Jul 23, 2026
AREX: Towards a Recursively Self-Improving Agent for Deep Resear... ↗ cs.AI 24 Jul 23, 2026
Detecting LLM-Generated Tokens in Human--LLM Coauthored Text ↗ cs.AI 6 Jul 23, 2026
Agent-Guided Relational Concept Discovery: Toward Interpretable... ↗ cs.AI 14 Jul 23, 2026
Bridging the Gap Between Plausibility and Admissibility: Constra... ↗ cs.AI 3 Jul 23, 2026
PATS: Policy-Aware Training Scaffolding for Agentic Reinforcemen... ↗ cs.AI 7 Jul 23, 2026
Logical Regression for Planning with Axioms ↗ cs.AI 2 Jul 23, 2026
Euclid-MCP: A Model Context Protocol Server for Deterministic Lo... ↗ cs.AI 1 Jul 23, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.