Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
MASS: Multiplayer World Models with Authoritative Shared State ↗ cs.CV 9 Aug 6, 2026
Toward Deployable Bangla Sign Language Recognition with Expert-V... ↗ cs.CV 2 Aug 6, 2026
PRISM: Distribution-Gated Flow Matching for Controllable Unpaire... ↗ cs.CV 2 Aug 6, 2026
Depth-Guided Video Object Counting in Crowded Scenes ↗ cs.CV 8 Aug 6, 2026
EmoWorld: A Decoupled Affective Field for Controllable Emotional... ↗ cs.CV 4 Aug 6, 2026
CFGPNet: Cross-Attention-Based Fused Gradient Programmed Network... ↗ cs.CV 4 Aug 6, 2026
HOPE: Hand-Object Pressure Estimation from Monocular Videos ↗ cs.CV 3 Aug 6, 2026
EvReflection: Event-Driven Micro-Dynamics for Reflection Removal ↗ cs.CV 6 Aug 6, 2026
Support Operation Factorization: Compositional Readout of Frozen... ↗ cs.CV 4 Aug 6, 2026
BendTwin: Robust Dense-to-Sparse Physical Reconstruction with Be... ↗ cs.CV 9 Aug 6, 2026
Learning visual representations for compositional analysis of ar... ↗ cs.CV 3 Aug 6, 2026
Patient Pose Assessment Using a CT-Based Framework for Synthetic... ↗ cs.CV 10 Aug 6, 2026
Sample-Adaptive Latent Rewards for Uncertainty-Guided Diffusion... ↗ cs.CV 7 Aug 6, 2026
Confidence matters: Leveraging Multi-view Geometric Priors for G... ↗ cs.CV 2 Aug 6, 2026
Dense-Cast: A lightweight ensemble of deep learning architecture... ↗ cs.CV 2 Aug 6, 2026
Domain-Grounded Candidate Selection for Agentic Image Editing: A... ↗ cs.CV 4 Aug 6, 2026
The Next Screenshot Knows: Gated Hindsight Distillation for Mobi... ↗ cs.CV 5 Aug 6, 2026
Bar-JEPA: Extracting Values from Bar Chart with Joint-Embedding... ↗ cs.CV 3 Aug 6, 2026
Learning from Failures: Retrieval-Centric CoT via Hard Negatives... ↗ cs.CV 6 Aug 6, 2026
Keeping up with a growing satellite archive without wrecking the... Explained cs.CV 6 Aug 6, 2026
PaCoNet: Deep Data Extraction for Parallel Coordinates ↗ cs.CV 4 Aug 6, 2026
Iterate or Widen? When Test-Time Refinement Helps LiDAR Scene Co... ↗ cs.CV 2 Aug 6, 2026
Wan-Animate-2: Pushing the Application Boundaries of Character A... ↗ cs.CV 12 Aug 6, 2026
Universal Concept Disruption for SAM3 Image Segmentation ↗ cs.CV 3 Aug 6, 2026
Multi-Year Geospatial Reasoning using Interannually-Consistent H... ↗ cs.CV 5 Aug 6, 2026
Diff-VF: Training-free High-quality Long Video Generation via Di... ↗ cs.CV 4 Aug 6, 2026
Topology-Aware Neighborhood Learning for Source-Free Cross-Scene... ↗ cs.CV 5 Aug 6, 2026
Big, Bright, or Invisible: A Frozen-Feature Benchmark of 3D CT F... ↗ cs.CV 5 Aug 6, 2026
Respect Your Zero-Shot Uncertainty: Conservative Calibration for... ↗ cs.CV 8 Aug 6, 2026
MirrorNet: Can Medical Image Anonymization Really Protect Patien... ↗ cs.CV 1 Aug 6, 2026
Floating Radiance Networks ↗ cs.CV 8 Aug 6, 2026
ReToken: One Token to Improve Vision-Language Models for Visual... ↗ cs.CV 6 Jul 30, 2026
ACE-Data-0: Human-Centric Ambient Capture as Embodied Data Engin... ↗ cs.CV 16 Jul 30, 2026
PhiZero: A World Model Built Around Physical Language ↗ cs.CV 7 Jul 30, 2026
Chimera: Designing and Chinchilla-Scaling Hybrid Visual Diffusio... ↗ cs.CV 12 Jul 30, 2026
Beacon: Knowing When and How to Perform Agentic Visual Reasoning ↗ cs.CV 14 Jul 30, 2026
VAD: Attributing Visual Evidence for Target Reconstruction in Mu... ↗ cs.CV 12 Jul 30, 2026
MixFrag: Fragility-Guided Mixed-Precision Post-Training Quantiza... ↗ cs.CV 3 Jul 30, 2026
ROAD: Reciprocal-Objective Alignment of Discriminative Semantics... ↗ cs.CV 8 Jul 30, 2026
Finding Change in Satellite Archives from Text: How to Combine B... ↗ cs.CV 3 Jul 30, 2026
MIND: Multimodal Intent-Driven Network via Diffusion Transformer... ↗ cs.CV 6 Jul 30, 2026
ScaFE: Data-Efficient Scar Classification with LLM-Generated Cli... ↗ cs.CV 2 Jul 30, 2026
MarkushGlyph and OCSRGlyph: Improved Chemical Structure Recognit... ↗ cs.CV 4 Jul 30, 2026
What to Remove, What to Preserve: Dual-Ambiguity Rectification f... ↗ cs.CV 9 Jul 30, 2026
Beyond Frame Selection: Generative Latent Evidence Aggregation f... ↗ cs.CV 6 Jul 30, 2026
RefCaptioner: Multi-Reference Image-Grounded Video Captioning ↗ cs.CV 19 Jul 30, 2026
AuricularWorld: Hierarchical Action-Guided World Modeling for Fi... ↗ cs.CV 9 Jul 30, 2026
Towards Real-Time PixOOD: Efficient Anomaly Segmentation for Aut... ↗ cs.CV 4 Jul 30, 2026
Can Vision-Language Models Reason about AI Edits in Images? ↗ cs.CV 4 Jul 30, 2026
VisualRouter: Query-Grounded Visual Sampling for Long Video Unde... ↗ cs.CV 8 Jul 30, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.