Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

276 results

PaperCategoryAuthorsDate
Embedding Models Measure in Peculiar Ways ↗ cs.CL 2 Sep 17, 2026
Unifying Models of Intergroup Hostility in Online Discourse ↗ cs.CL 3 Sep 17, 2026
JEPA-Anything: Learning Predictive Models across Different World... ↗ cs.CL 13 Sep 17, 2026
RetireOPD: Self-Retiring On-Policy Distillation for Agentic Rein... ↗ cs.CL 11 Sep 17, 2026
Harm Laundering in GPT Models: Evidence That Gender Discriminati... ↗ cs.CL 3 Sep 17, 2026
dQwen3.5: Hybrid-Attention Diffusion Language Models ↗ cs.CL 6 Sep 17, 2026
On-Demand Attention: Language Models Know When to Recall ↗ cs.CL 4 Sep 17, 2026
Summarization Bias: The Directional Collapse of Objective Projec... ↗ cs.CL 1 Sep 17, 2026
HerHealthEval: Evaluating Multilingual and Register-Sensitive Un... ↗ cs.CL 5 Sep 17, 2026
UniPolicy: Unified Objective-Specific Policies for Generative Se... ↗ cs.CL 9 Sep 17, 2026
Chronicle: Cut-Point Replay for Regression Testing of LLM Agents ↗ cs.CL 2 Sep 17, 2026
What Does Privileged Information Add to On-Policy Self-Distillat... ↗ cs.CL 5 Sep 17, 2026
WiC is Not WSD: A Study on LLMs and Lexical Ambiguity Resolution ↗ cs.CL 5 Sep 17, 2026
SAFARI: An Industrial Benchmark for LLM-Assisted Hazard Analysis... ↗ cs.CL 5 Sep 17, 2026
Steering the Compass: Aligning Dynamic Psychological Counseling... ↗ cs.CL 13 Sep 17, 2026
An Analysis of Training-Free Self-Reported Confidence in Languag... ↗ cs.CL 5 Sep 17, 2026
Relational Attention for Data-Efficient Language Modeling ↗ cs.CL 3 Sep 17, 2026
Edustories: A Collection of Real-world Case Studies from Classro... ↗ cs.CL 7 Sep 17, 2026
Stress-testing Alignment Midtraining ↗ cs.CL 6 Sep 17, 2026
Xeno-Interpretability: Investigating the Alien Minds of LLMs ↗ cs.CL 6 Sep 17, 2026
Schema-Anchored Latent Reasoning for Semantic Parsing-Based Know... ↗ cs.CL 8 Sep 17, 2026
To Copy or Not to Copy: Controlling Speculative Decoding via Int... ↗ cs.CL 5 Sep 17, 2026
CritICL: Inference-Time Weak-to-Strong Generalization from Small... ↗ cs.CL 7 Aug 27, 2026
TTPO: Test-Time Policy Optimization ↗ cs.CL 11 Aug 27, 2026
Stochastic Estimation of Transduced Language Models ↗ cs.CL 5 Aug 27, 2026
Boosting LLM Exploration via Weak-Model Guidance in RLVR ↗ cs.CL 5 Aug 27, 2026
Consolidating RLVR Capabilities Across Domains: A Deep Dive into... ↗ cs.CL 11 Aug 27, 2026
How Language Models Organize and Structure Moral Knowledge ↗ cs.CL 1 Aug 27, 2026
Making Clinical Language Models Auditable: Concept-Guided Fine-T... ↗ cs.CL 2 Aug 27, 2026
RATIO: A Benchmark for Retrieval Across Typed Ideation Operation... ↗ cs.CL 2 Aug 27, 2026
D2C-Routing: Dimension-to-Composition Evidence Routing for Mixed... ↗ cs.CL 6 Aug 27, 2026
Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 ↗ cs.CL 11 Aug 27, 2026
Your Voice Cloning System is Secretly a Voice Anonymizer ↗ cs.CL 4 Aug 27, 2026
RCMN: Understanding Misleadingness in Influential Public Discour... ↗ cs.CL 1 Aug 27, 2026
INTENT-AS-A-TOOL Makes it Easy to Track Agentic Misalignment ↗ cs.CL 8 Aug 27, 2026
Pair-Level Essay-Scale Republication and Reuse from Fragmented H... ↗ cs.CL 4 Aug 27, 2026
BTS-AgentBench: A Deterministic, Replayable Pipeline from Read-O... ↗ cs.CL 1 Aug 27, 2026
Difference-in-Differences on a Censored Rating Scale Can Manufac... ↗ cs.CL 6 Aug 27, 2026
SCIT: Testing Causal Cache Carriers in Latent Chain-of-Thought M... ↗ cs.CL 3 Aug 27, 2026
BALMS: Benchmarking Agentic LLMs for Longitudinal Mental Health... ↗ cs.CL 10 Aug 27, 2026
When Text Misleads: Inconsistent-Aware Reasoning for Audio-Groun... ↗ cs.CL 11 Aug 27, 2026
Prediction of Prediction (PoP): Inter-Layer Activation Fusion fo... ↗ cs.CL 1 Aug 27, 2026
STAR : Sentence Translation Alignment Rate for Document-to-Docum... ↗ cs.CL 6 Aug 27, 2026
Said Aloud, Read Different: Cross-Modal Instability in Multimoda... ↗ cs.CL 5 Aug 27, 2026
TwinKV: A Composable Repair Pass for KV Cache Eviction via Pairw... ↗ cs.CL 6 Aug 27, 2026
Cross-Lingual Alignment Without Joint Training: Do Monolingual L... ↗ cs.CL 4 Aug 27, 2026
DocTalkBN: A Novel Dataset of Expert Telemedicine Conversations... ↗ cs.CL 6 Aug 27, 2026
Research Design Tracking and Assessment for the Social Sciences ↗ cs.CL 8 Aug 27, 2026
Cascaded Batch Prompting ↗ cs.CL 2 Aug 27, 2026
Reasoning about In-Context Samples for Machine-Translation ↗ cs.CL 3 Aug 27, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.