Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
Can 4D Foundation Models Remember? ↗ cs.CV 3 Sep 17, 2026
SplashSplat: Reconstructing Splashing Liquids from Real-World Mu... ↗ cs.CV 6 Sep 17, 2026
FAMOS: Feed-Forward 3D Articulation Modeling from Sparse Observa... ↗ cs.CV 8 Sep 17, 2026
Paint-Anything: Unified Any-Color Control for Image Generation a... ↗ cs.CV 5 Sep 17, 2026
ERCPMP-Gx: Endoscopic Image and Video Dataset for Morphological,... ↗ cs.CV 5 Sep 17, 2026
FlowSGS: Improving Flow Matching Priors for Inverse Imaging with... ↗ cs.CV 3 Sep 17, 2026
Should This Case Be Adapted? Prediction Fragmentation Controls T... ↗ cs.CV 8 Sep 17, 2026
FunArt: Decoding Functional Structure and Articulation from Gene... ↗ cs.CV 3 Sep 17, 2026
Earth Surface Immune System for Rapid Monitoring of Unknown Anom... ↗ cs.CV 6 Sep 17, 2026
PROVIA: Procedure State Tracking for Online Mistake Detection in... ↗ cs.CV 10 Sep 17, 2026
Refinement Is Inherently Editable: Training-Free Prompt-to-Promp... ↗ cs.CV 6 Sep 17, 2026
PhGS: Post-Hoc Pruning and Refinement of Single-View Feed-Forwar... ↗ cs.CV 5 Sep 17, 2026
RawSLAM: Online HDR Gaussian SLAM from Linear Radiance ↗ cs.CV 2 Sep 17, 2026
DocAttriBench: Benchmarking Answer Grounding in Document Visual... ↗ cs.CV 6 Sep 17, 2026
A Dual-Stream Regulated Reconstruction and Segmentation Network... ↗ cs.CV 3 Sep 17, 2026
Automated Goldsmith's Mark Retrieval in Silverware ↗ cs.CV 8 Sep 17, 2026
Grounded Product Understanding in Livestream Videos ↗ cs.CV 7 Sep 17, 2026
SenseFuse: Label-Free Fusion of Image and Shape Encoders for Ope... ↗ cs.CV 5 Sep 17, 2026
Cross-Architecture Foundation-Model Distillation for Edge Flood... ↗ cs.CV 3 Sep 17, 2026
When Do Language-Grounded Explanations Help? A Graph-Bottleneck... ↗ cs.CV 2 Sep 17, 2026
WeVisDoc: From Coverage to Capability for Robust End-to-End Docu... ↗ cs.CV 7 Sep 17, 2026
TouchSight: Bare-Handed Tactile Prediction from Egocentric Video... ↗ cs.CV 6 Sep 17, 2026
Compact Vision Models for Iris Presentation Attack Detection und... ↗ cs.CV 2 Sep 17, 2026
MM-Future: Multi-Mode Joint World-Action Modeling for Autonomous... ↗ cs.CV 8 Sep 17, 2026
EliGSiR: Continual RGB-D Mapping with Gaussian Splatting under B... ↗ cs.CV 3 Sep 17, 2026
Fast Cross-Strength Multi-Contrast Brain MRI Translation using L... ↗ cs.CV 2 Sep 17, 2026
FreqDINO++: A Frequency-Guided Multi-Task Routing Vision Foundat... ↗ cs.CV 10 Sep 17, 2026
AgriScope: Pixel-Grounded Multimodal Understanding for Agricultu... ↗ cs.CV 4 Sep 17, 2026
Needles in a Raystack: Ultra-Sparse LiDAR Occupancy Detection fo... ↗ cs.CV 5 Sep 17, 2026
Ischemic Stroke Segmentation and Net Water Uptake Quantification... ↗ cs.CV 8 Sep 17, 2026
Task-Oriented Semantic Feature Transmission for Multi-Task Satel... ↗ cs.CV 4 Sep 17, 2026
Bridging Modalities on the Cortex: Surface-based MRI to PET Tran... ↗ cs.CV 7 Sep 17, 2026
Cross-Modal Attention Acts as a Frequency Filter: Why Verbose Pr... ↗ cs.CV 8 Sep 17, 2026
UrbanGround: From Local Perception to Spatial Agency in a Real-S... ↗ cs.CV 18 Aug 27, 2026
The work of visual grounding sits in under 2% of attention heads... Explained cs.CV 4 Aug 27, 2026
Reconstructing Humans and Objects in Interaction using Large Rec... ↗ cs.CV 2 Aug 27, 2026
Removing the scaffolding built to prevent collapse — simplifying... Explained cs.CV 7 Aug 27, 2026
Successive Capacity Growth: Task-Complexity-Driven Width and Dep... ↗ cs.CV 1 Aug 27, 2026
KnockGS:interaction-Grounded Calibrationof Physical Gaussian Rep... ↗ cs.CV 9 Aug 27, 2026
PAWBench: How Far Are We from Probabilistically Aligned World Mo... ↗ cs.CV 14 Aug 27, 2026
R2M-Bench: Evaluating Revisit Memory via Relative Consistency in... ↗ cs.CV 10 Aug 27, 2026
Detection of Christmas tree plantations from high-resolution aer... ↗ cs.CV 6 Aug 27, 2026
TADP: Task-Aware Deformable Prediction for Single-Stage 3D Objec... ↗ cs.CV 6 Aug 27, 2026
Sidecar: Training-Free Semantic Reuse for Character-Consistent F... ↗ cs.CV 2 Aug 27, 2026
UniFLM: United Segmentation and Measurement on Fetal Limb Ultras... ↗ cs.CV 6 Aug 27, 2026
DINOcular: Self-Supervised Visuospatial Representations ↗ cs.CV 3 Aug 27, 2026
CODE: Cross-Modal Calibration and Dynamic Suppression for Open W... ↗ cs.CV 4 Aug 27, 2026
PACE: A Unified Condense-and-Extract Paradigm for Fast VLM Infer... ↗ cs.CV 3 Aug 27, 2026
Vision-centric generative AI models: A software-hardware perspec... ↗ cs.CV 4 Aug 27, 2026
Unsupervised Adaptation of 3D CT Foundation Models for 3D CBCT S... ↗ cs.CV 4 Aug 27, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.