Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

334 results

PaperCategoryAuthorsDate
Progression as Latent Drift: Generative Forecasting of Slow-Evol... ↗ cs.CV 10 Jul 9, 2026
UniRef-UAV: A Multimodal Benchmark for Universal Referring in UA... ↗ cs.CV 6 Jul 9, 2026
WorldDirector: Building Controllable World Simulators with Persi... ↗ cs.CV 13 Jul 2, 2026
Alignment Is All You Need For X-to-4D Generation ↗ cs.CV 4 Jul 2, 2026
PointDiT: Pixel-Space Diffusion for Monocular Geometry Estimatio... ↗ cs.CV 10 Jul 2, 2026
From SRA to Self-Flow: Data Augmentation or Self-Supervision? ↗ cs.CV 4 Jul 2, 2026
Seek to Segment: Active Perception for Panoramic Referring Segme... ↗ cs.CV 5 Jul 2, 2026
Towards Robustness against Typographic Attack with Training-free... ↗ cs.CV 6 Jul 2, 2026
GeoMix: Descriptor-Free Visual Localization via Global Context a... ↗ cs.CV 5 Jul 2, 2026
Combating Textual Noise and Redundancy: Entropy-Aware Dense Visu... ↗ cs.CV 3 Jul 2, 2026
EAGLE-360: Embodied Active Global-to-Local Exploration in 360$^\... ↗ cs.CV 5 Jul 2, 2026
Interpretation-Oriented Cloud Removal via Observation-Anchored R... ↗ cs.CV 8 Jul 2, 2026
OrbitQuant: Data-Agnostic Quantization for Image and Video Diffu... ↗ cs.CV 8 Jul 2, 2026
MARVEL: Margin-Aware Robust von Mises-Fischer Expert Learning fo... ↗ cs.CV 2 Jul 2, 2026
Learning to Evolve Scenes: Reasoning about Human Activities with... ↗ cs.CV 3 Jul 2, 2026
Wavelet-Guided Semantic Signal Compensation for Inversion-Free I... ↗ cs.CV 3 Jul 2, 2026
Object-centric LeJEPA ↗ cs.CV 2 Jul 2, 2026
Show Me Examples: Inferring Visual Concepts from Image Sets ↗ cs.CV 6 Jul 2, 2026
Transformer Geometry Observatory TGO-II: Representational Simila... ↗ cs.CV 2 Jul 2, 2026
Representation Distribution Matching for One-Step Visual Generat... ↗ cs.CV 5 Jul 2, 2026
Learning Spectral and Polarimetric Clues for One-to-Multimodal N... ↗ cs.CV 5 Jul 2, 2026
VisionAId: An Offline-First Multimodal Android Assistant for Peo... ↗ cs.CV 2 Jul 2, 2026
GAP-GDRNet: Geometry-Aware Monocular Visual Pose Sensing on a Si... ↗ cs.CV 2 Jul 2, 2026
NEvo: Neural-Guided Evolutionary Video Synthesis for Dynamic Vis... ↗ cs.CV 6 Jul 2, 2026
InvSplat: Inverse Feed-Forward Scene Splatting ↗ cs.CV 5 Jul 2, 2026
Search-based Testing of Vision Language Models for In-Car Scene... ↗ cs.CV 4 Jul 2, 2026
Dual-Selective Network for Domain-Incremental Change Detection ↗ cs.CV 4 Jul 2, 2026
DisciplineGen-1M: A Large-Scale Dataset for Multidisciplinary Vi... ↗ cs.CV 14 Jul 2, 2026
FlowCIR: Semantic Transport via Flow Matching for Zero-Shot Comp... ↗ cs.CV 6 Jul 2, 2026
AGVBench: A Reliability-Oriented Benchmark of Data Augmentation... ↗ cs.CV 7 Jul 2, 2026
AnyGroundBench: A Specialized-Domain Benchmark for Video Groundi... ↗ cs.CV 9 Jul 2, 2026
ArcAD: Anomaly-Rectified Calibration for Cold-Start Supervised A... ↗ cs.CV 8 Jul 2, 2026
When Token Compression Breaks: Structural Pruning vs. Token Redu... ↗ cs.CV 2 Jul 2, 2026
Efficient Waste Sorting for Circular Economy: A Confidence-guide... ↗ cs.CV 5 Jul 2, 2026
DetailAnywhere: Fashion Detail Generation via Cross-Modal Featur... ↗ cs.CV 15 Jul 2, 2026
MedSaab-US: A Backpropagation-Free Multi-Scale Wavelet-Saab Fram... ↗ cs.CV 1 Jul 2, 2026
RadiomicNet: A Hybrid Radiomics-Guided Lightweight Architecture... ↗ cs.CV 1 Jul 2, 2026
Efficient PEFT Methods with Adaptive Checkpointing for Vision Mo... ↗ cs.CV 2 Jul 2, 2026
Patient-Specific Articulated Digital Twins from a Single Full-Bo... ↗ cs.CV 3 Jul 2, 2026
SAMoR: Motion Modelling for Articulated Objects of Any Skeleton... ↗ cs.CV 3 Jul 2, 2026
AdaCount: Training-Free Similarity-Guided Spatial and Feature Ad... ↗ cs.CV 2 Jul 2, 2026
AbsoluteDegradation: A Physics-Inspired Synthetic Film-Degradati... ↗ cs.CV 6 Jul 2, 2026
X-Splat: Gaussian Splatting for 3D CBCT Generation from Single P... ↗ cs.CV 5 Jul 2, 2026
WBMM: Windowed Batch Matrix Multiplication for Efficient Large R... ↗ cs.CV 7 Jul 2, 2026
LongEgoRefer: A Benchmark for Long-Form Egocentric Video Referri... ↗ cs.CV 6 Jul 2, 2026
Multimodal Fusion for Fine-Grained Classification of Breast Fibr... ↗ cs.CV 7 Jul 2, 2026
DanceOPD: On-Policy Generative Field Distillation ↗ cs.CV 11 Jun 25, 2026
Ask, Solve, Generate: Self-Evolving Unified Multimodal Understan... ↗ cs.CV 8 Jun 25, 2026
Paying More Attention to Visual Tokens in Self-Evolving Large Mu... ↗ cs.CV 7 Jun 25, 2026
DnA: Denoising Attention for Visual Tasks ↗ cs.CV 5 Jun 25, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.