Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
ViewMind3D: Modular View-Aware Inference for Training-Free 3D-QA ↗ cs.CV 6 Jul 30, 2026
Negative controls reveal volume-driven confounding in radiomics... ↗ cs.CV 9 Jul 30, 2026
Large scale cross-regional remote sensing flood monitoring frame... ↗ cs.CV 9 Jul 30, 2026
Hand-Object Interaction in the Age of Large Foundation Models:Re... ↗ cs.CV 9 Jul 30, 2026
Explaining Image Similarity with Automatically Extracted Concept... ↗ cs.CV 4 Jul 30, 2026
ShadowDancer: Teaching Video World Models Any Action by Learning... ↗ cs.CV 3 Jul 30, 2026
Capturing Token Tendencies for Training-Free Token Pruning in Mu... ↗ cs.CV 7 Jul 30, 2026
Same Branches, Different Trees: A Bifurcation Connectedness Metr... ↗ cs.CV 6 Jul 30, 2026
AdaAnchor4D: Anchor-Conditioned Spatiotemporal Feature Aggregati... ↗ cs.CV 9 Jul 30, 2026
ObjectStream: Latent Objects as Memory Anchors for Streaming Vid... ↗ cs.CV 11 Jul 30, 2026
MonoVoc: Decoupling Geometry and Semantics for Lightweight Monoc... ↗ cs.CV 4 Jul 30, 2026
Filling the Pareto-Optimal Front for Affordance Segmentation on... ↗ cs.CV 5 Jul 30, 2026
Beyond Visual Ambiguity: Guiding Robust Monocular Depth Estimati... ↗ cs.CV 5 Jul 30, 2026
MSCM-net: A hyperspectral image classiffcation method based on m... ↗ cs.CV 7 Jul 30, 2026
Theia: Large-Scale Multimodal Captioning and Automated Validatio... ↗ cs.CV 4 Jul 30, 2026
TARS: Timestep-Aware Data Scaling for 3D-Free Video Re-Shooting ↗ cs.CV 8 Jul 30, 2026
Space2Ground 2.0: A Multi-Source Dataset and Framework for Agric... ↗ cs.CV 5 Jul 30, 2026
EgoGenesis: Egocentric World-Action Modeling with Online Anchore... ↗ cs.CV 12 Jul 30, 2026
FaithEyes: Towards Faithful Tool Use via Multi-Agent Process-Ima... ↗ cs.CV 5 Jul 30, 2026
Scaling Vision-Language Models Is Not Enough to Mitigate Bias ↗ cs.CV 3 Jul 30, 2026
Think with Extra-Image: A Farmland Segmentation Agent Driven by... ↗ cs.CV 7 Jul 30, 2026
3D-Aware VLMs with Implicit and Explicit Geometries ↗ cs.CV 7 Jul 23, 2026
Streaming Multi-Agent Autoregressive Diffusion Model with World... ↗ cs.CV 5 Jul 23, 2026
Unified Video Dense Prediction from Disjoint Data ↗ cs.CV 5 Jul 23, 2026
Inference-Time Scaling of Diffusion Models via Progressive Seed... ↗ cs.CV 2 Jul 23, 2026
GraphVid: Interactive Graph-Controllable Video Generation ↗ cs.CV 8 Jul 23, 2026
Synthetic data generation framework for quality control automati... ↗ cs.CV 4 Jul 23, 2026
Self-Supervised Learning of Structured Dynamics from Videos ↗ cs.CV 3 Jul 23, 2026
Scene Parameter Saliency via Differentiable Light Transport ↗ cs.CV 2 Jul 23, 2026
Visual Contrastive Self-Distillation ↗ cs.CV 7 Jul 23, 2026
SANA-Video 2.0: Hybrid Linear Attention with Attention Residuals... ↗ cs.CV 14 Jul 23, 2026
UnDA: Unpaired Domain Alignment for Cross-Modal Knowledge Transf... ↗ cs.CV 6 Jul 23, 2026
Towards Robust Iris Recognition Through Occlusion Identification... ↗ cs.CV 3 Jul 23, 2026
ElasticTTT: Prior-Preserving Test-Time Tuning for Video Editing ↗ cs.CV 4 Jul 23, 2026
Boosting Robustness for All-Weather Self-Supervised Depth Estima... ↗ cs.CV 4 Jul 23, 2026
Texture++: Elevating 3D Asset Texture Resolution with a Region-A... ↗ cs.CV 7 Jul 23, 2026
Recurrent Sinusoidal INRs for Efficient High-Fidelity Representa... ↗ cs.CV 3 Jul 23, 2026
It can look right and still have the wrong shape — a benchmark f... Explained cs.CV 2 Jul 23, 2026
CLUIE: Clustering-Aware Recurrent Propagation with Local Structu... ↗ cs.CV 6 Jul 23, 2026
SPDCN: Strip-based Deformable Convolutional Network for Steel Su... ↗ cs.CV 4 Jul 23, 2026
GrainGS: Gradient-Decoupled Gaussian Splatting for Efficient Dyn... ↗ cs.CV 9 Jul 23, 2026
DAPM: UAV Monocular Depth Estimation from Any Height, Pitch, Rol... ↗ cs.CV 6 Jul 23, 2026
Adaptive Identity Anchoring: Closed-Loop Keyframe Placement for... ↗ cs.CV 1 Jul 23, 2026
Towards Privacy-Preserving Federated Prompt Tuning under Data He... ↗ cs.CV 9 Jul 23, 2026
When Are Reasoning-Based Guardrails Not Efficient? ResponseGuard... ↗ cs.CV 1 Jul 23, 2026
DINOde: Continuous Vision-Text Alignment for Open-Vocabulary Sem... ↗ cs.CV 4 Jul 23, 2026
ASTRA-Net: Anatomy-Specific Transfer and Representation Alignmen... ↗ cs.CV 9 Jul 23, 2026
Incremental Optimal Assignment for Real-Time Crowd Tracking ↗ cs.CV 1 Jul 23, 2026
Quality-Aware Multimodal Fusion Reveals Implicit Identity in Val... ↗ cs.CV 2 Jul 23, 2026
SlerpFlow: Spherical Trajectory Correction for Rectified Flow In... ↗ cs.CV 7 Jul 23, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.