Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

276 results

PaperCategoryAuthorsDate
Representing and Parsing Korean Constituency Structure at Differ... ↗ cs.CL 6 Aug 27, 2026
ITL: Interpretable Document Alignment with Structured Reference... ↗ cs.CL 3 Aug 27, 2026
ConceptGuard: Benchmarking Context-Sensitive Unlearning in Large... ↗ cs.CL 2 Aug 20, 2026
G-CARL: Grounded Checklist-Aligned Reward Learning for Patient-O... ↗ cs.CL 6 Aug 20, 2026
Inducing Task Models from Computer-Use Traces ↗ cs.CL 4 Aug 20, 2026
Inject, Align, Recover: Staged Post-Training for Retrieval-Free... ↗ cs.CL 4 Aug 20, 2026
Task-CoEvolve: Efficient Harness Optimization via Adaptive Valid... ↗ cs.CL 3 Aug 20, 2026
FormalTCS: Benchmarking End-to-End Frontier Formal Theoretical C... ↗ cs.CL 5 Aug 20, 2026
When Text and Numbers Disagree: Evidence Arbitration in Large La... ↗ cs.CL 8 Aug 20, 2026
OenoBench: A Wine-Domain Benchmark for Knowledge-Grounded Evalua... ↗ cs.CL 1 Aug 20, 2026
SABET-QA: Temporal Knowledge Graph Question Answering ↗ cs.CL 3 Aug 20, 2026
Auditing Cross-Lingual Fairness in Language Model Watermarking ↗ cs.CL 6 Aug 20, 2026
HealMed: Multilingual Evaluation of Large Language Models in Med... ↗ cs.CL 45 Aug 20, 2026
Robust Incomplete Multimodal Sentiment Analysis via Iterative Pr... ↗ cs.CL 6 Aug 20, 2026
Natural Language Code Retrieval for 1C:Enterprise: An Open Bench... ↗ cs.CL 2 Aug 20, 2026
Dynamic Gated Cross-Modal Fusion with Sarcastic-aware Contrastiv... ↗ cs.CL 6 Aug 20, 2026
Learning how to Forget: Fine-tuning for Long-Context Sparse Atte... ↗ cs.CL 5 Aug 20, 2026
Interrupting the Loop: Periodic Subject Changes Raise Judged Sur... ↗ cs.CL 1 Aug 20, 2026
A knowledge-guided agentic framework for mitigating patient-cont... ↗ cs.CL 6 Aug 20, 2026
LittleLearner: Language Models Under Pedagogically Controlled Kn... ↗ cs.CL 7 Aug 13, 2026
SAEVerbalizer: Generating Explanations for Sparse Autoencoder Fe... ↗ cs.CL 8 Aug 13, 2026
DFM Mimir v1: An Open HRM Delivering Frontier Performance at 1B... ↗ cs.CL 5 Aug 13, 2026
Measuring Task-Agnostic Training Data Influence Across Language... ↗ cs.CL 9 Aug 13, 2026
Toward a Gricean Retreat: Probing LLMs for Knowledge Boundaries... ↗ cs.CL 3 Aug 13, 2026
Are You Sure You're Sure? On the Impact of Instruction Tuning on... ↗ cs.CL 3 Aug 13, 2026
Motor, Cognitive, or Corpus? What Survives Cross-Lingual Transfe... ↗ cs.CL 5 Aug 13, 2026
CROP: Task Relevance via Counterfactuals for Selective On-Policy... ↗ cs.CL 3 Aug 13, 2026
RippleMem: From Isolated Retrieval to Associative Recollection f... ↗ cs.CL 7 Aug 13, 2026
It's How You Ask: Gender-Associated Linguistic Bias in LLMs ↗ cs.CL 2 Aug 13, 2026
Beyond Local Accuracy: A Protocol-Level Identifiability Audit fo... ↗ cs.CL 5 Aug 13, 2026
Refusing Intent, Not Form: Wrapper-Based Intent-Group Supervisio... ↗ cs.CL 10 Aug 13, 2026
Mixture of Training: Recombining Small-Scale Scaffolded Pretrain... ↗ cs.CL 4 Aug 13, 2026
How Do VLMs Behave When Blind or Misled? Behavioral Evaluation o... ↗ cs.CL 7 Aug 13, 2026
Self-Referential Induction Increases Response Instability Relati... ↗ cs.CL 2 Aug 13, 2026
Localize, Then Reason: Visual Latent Structural Reasoning for Mo... ↗ cs.CL 3 Aug 13, 2026
GEM: A Generative Embedding Model Bridging Reasoning and Retriev... ↗ cs.CL 2 Aug 13, 2026
Which LLM Is Your Ideal Companion? Evaluating Emotional Companio... ↗ cs.CL 3 Aug 13, 2026
Better Decomposition, Free Aggregation: A Synthesizer-Folding Fr... ↗ cs.CL 8 Aug 13, 2026
LigBench: A Unified and Human-Aligned Benchmark for LLM-based Re... ↗ cs.CL 9 Aug 13, 2026
Learning When to Trust via Selective Context Preference Optimiza... ↗ cs.CL 7 Aug 6, 2026
The Bitter Lesson of Tool Calling ↗ cs.CL 4 Aug 6, 2026
RP-OPSD: Reasoning-Pivot-Guided On-Policy Self-Distillation for... ↗ cs.CL 3 Aug 6, 2026
Benchmarking the Benchmarks: Evaluating Benchmarks for Conversat... ↗ cs.CL 3 Aug 6, 2026
Benchmarking and Enhancing LLMs for Rule-Intensive Review of Nat... ↗ cs.CL 7 Aug 6, 2026
NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering ↗ cs.CL 2 Aug 6, 2026
Routing Is Least Learnable Where It Is Most Valuable: Bounds on... ↗ cs.CL 4 Aug 6, 2026
Decolonizing Linguistic Policies in Automated Speech Recognition... ↗ cs.CL 5 Aug 6, 2026
Beyond Sequence Order: Syntax-Informed Positional Embeddings for... ↗ cs.CL 3 Aug 6, 2026
Training-Free Token-Level Steering for LLM Personalized Co-Writi... ↗ cs.CL 7 Aug 6, 2026
FormBharo: Designing and Evaluating a Voice Agent for Conversati... ↗ cs.CL 3 Aug 6, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.