Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
LongE2V: Long-Horizon Event-based Video Reconstruction, Predicti... ↗ cs.CV 7 Jul 9, 2026
Geometry and Gradient-based Partitioning for Panoramic Outdoor R... ↗ cs.CV 10 Jul 9, 2026
OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step A... ↗ cs.CV 9 Jul 9, 2026
Enhancing In-context Panoramic Generation via Geometric-aware Pr... ↗ cs.CV 5 Jul 9, 2026
OpenCoF: Learning to Reason Through Video Generation ↗ cs.CV 5 Jul 9, 2026
WaspMOT: A Benchmark for Long-Term Multi-Object Tracking of Tric... ↗ cs.CV 7 Jul 9, 2026
Pose-to-Biomechanics: Bridging 3D Human Pose Estimation and Biom... ↗ cs.CV 2 Jul 9, 2026
LTM: Large-scale Terrain Model for Wildfire-prone Landscapes ↗ cs.CV 5 Jul 9, 2026
HumanForge: A Human-Centric Deepfake Video Benchmark with Multi-... ↗ cs.CV 5 Jul 9, 2026
SAM-MT: Real-Time Interactive Multi-Target Video Segmentation ↗ cs.CV 3 Jul 9, 2026
Multi-Resolution Feature Stem for Diabetic Retinopathy lesion se... ↗ cs.CV 2 Jul 9, 2026
Do Transformations Reveal the Truth? Generative Residual Learnin... ↗ cs.CV 5 Jul 9, 2026
When Structured Sparse Autoencoders Learn Consistent Concepts Ac... ↗ cs.CV 3 Jul 9, 2026
Switch-Reasoner: Learn When to Think in Multitask Mixtures via R... ↗ cs.CV 10 Jul 9, 2026
VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segm... ↗ cs.CV 1 Jul 9, 2026
Whareformer: Learning to Track What is Where in Long Egocentric... ↗ cs.CV 5 Jul 9, 2026
Beyond wheelchairs and blindfolds: Investigating disability ster... ↗ cs.CV 3 Jul 9, 2026
Do Egocentric Video-Language Models Capture Both Hand- and Objec... ↗ cs.CV 5 Jul 9, 2026
CT-CLIP Representations for Multimodal Lung Cancer Survival Pred... ↗ cs.CV 7 Jul 9, 2026
Cognitive-structured Multimodal Agent for Multimodal Understandi... ↗ cs.CV 6 Jul 9, 2026
VEGAS: Human-Aligned Video Caption Evaluation via Gaze ↗ cs.CV 8 Jul 9, 2026
Predicting Viticulture Potential through an Ensemble of U-Net an... ↗ cs.CV 3 Jul 9, 2026
DeltaV: Thinking with Visual State Updates in Unified Large Mult... ↗ cs.CV 9 Jul 9, 2026
Track2Map: Online Deformable SLAM with Motion-Aware Pose Optimiz... ↗ cs.CV 9 Jul 9, 2026
Swapping Faces, Saving Features: A Dual-Purpose Pipeline for Ped... ↗ cs.CV 2 Jul 9, 2026
Attribute Retrieving for Open-Vocabulary Endoscopic Compositiona... ↗ cs.CV 6 Jul 9, 2026
Classical versus Deep Mirror-Symmetry Scoring: A Benchmark of Th... ↗ cs.CV 1 Jul 9, 2026
WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Mo... ↗ cs.CV 8 Jul 9, 2026
Texture Representations in Deep Vision Models: Comparing CNNs, V... ↗ cs.CV 4 Jul 9, 2026
ARGUS: Accelerated, Robust, General, and Unsupervised Cell Track... ↗ cs.CV 4 Jul 9, 2026
Enhancing the KidSat Model: Integrating Geographical Encoding an... ↗ cs.CV 7 Jul 9, 2026
Progression as Latent Drift: Generative Forecasting of Slow-Evol... ↗ cs.CV 10 Jul 9, 2026
UniRef-UAV: A Multimodal Benchmark for Universal Referring in UA... ↗ cs.CV 6 Jul 9, 2026
WorldDirector: Building Controllable World Simulators with Persi... ↗ cs.CV 13 Jul 2, 2026
Alignment Is All You Need For X-to-4D Generation ↗ cs.CV 4 Jul 2, 2026
PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimatio... ↗ cs.CV 10 Jul 2, 2026
From SRA to Self-Flow: Data Augmentation or Self-Supervision? ↗ cs.CV 4 Jul 2, 2026
Seek to Segment: Active Perception for Panoramic Referring Segme... ↗ cs.CV 5 Jul 2, 2026
Towards Robustness against Typographic Attack with Training-free... ↗ cs.CV 6 Jul 2, 2026
GeoMix: Descriptor-Free Visual Localization via Global Context a... ↗ cs.CV 5 Jul 2, 2026
Combating Textual Noise and Redundancy: Entropy-Aware Dense Visu... ↗ cs.CV 3 Jul 2, 2026
EAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\... ↗ cs.CV 5 Jul 2, 2026
Interpretation-Oriented Cloud Removal via Observation-Anchored R... ↗ cs.CV 8 Jul 2, 2026
OrbitQuant: Data-Agnostic Quantization for Image and Video Diffu... ↗ cs.CV 8 Jul 2, 2026
MARVEL: Margin-Aware Robust von Mises-Fischer Expert Learning fo... ↗ cs.CV 2 Jul 2, 2026
Learning to Evolve Scenes: Reasoning about Human Activities with... ↗ cs.CV 3 Jul 2, 2026
Wavelet-Guided Semantic Signal Compensation for Inversion-Free I... ↗ cs.CV 3 Jul 2, 2026
Object-centric LeJEPA ↗ cs.CV 2 Jul 2, 2026
Show Me Examples: Inferring Visual Concepts from Image Sets ↗ cs.CV 6 Jul 2, 2026
Transformer Geometry Observatory TGO-II: Representational Simila... ↗ cs.CV 2 Jul 2, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.