Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
Pseudo-Text-Conditioned 3D Grounding DINO for Organ Localization... ↗ cs.CV 6 Jun 25, 2026
PanoImager: Geometry-Guided Novel View Synthesis and Reconstruct... ↗ cs.CV 2 Jun 25, 2026
Stop building a bespoke model and put it on a general base — spo... Explained cs.CV 1 Jun 25, 2026
Event-Aware Instructed Assistant for Referring Video Segmentatio... ↗ cs.CV 4 Jun 25, 2026
Unison: Benchmarking Unified Multimodal Models via Synergistic U... ↗ cs.CV 4 Jun 25, 2026
Geometric Gradient Rectification for Safe Open-Set Semi-Supervis... ↗ cs.CV 7 Jun 25, 2026
Computer Vision for MOBA Analytics: A Dataset and Baseline for V... ↗ cs.CV 5 Jun 25, 2026
Scaling Multi-Reference Image Generation with Dynamic Reward Opt... ↗ cs.CV 9 Jun 25, 2026
TraMP-LLaMA: Generative Interpretability with Decoupled Instruct... ↗ cs.CV 5 Jun 25, 2026
Focusing on What Matters: Saliency-Harnessing Accurate Routing f... ↗ cs.CV 7 Jun 25, 2026
PortraitGen: Exemplar-Driven GRPO with Dual-Reward Guidance for... ↗ cs.CV 8 Jun 25, 2026
PhysRAG: Enhancing Physics-Awareness in Video Generation via Ret... ↗ cs.CV 5 Jun 25, 2026
Qwen-Image-Agent: Bridging the Context Gap in Real-World Image G... ↗ cs.CV 21 Jun 25, 2026
Confidence-Aware Tool Orchestration for Robust Video Understandi... ↗ cs.CV 3 Jun 25, 2026
Tractography-Driven Synthetic Data Generation for Fiber Bundle S... ↗ cs.CV 7 Jun 25, 2026
Modeling Local, Global, and Cross-Modal Context in Multimodal 3D... ↗ cs.CV 5 Jun 25, 2026
JanusMesh: Fast and Zero-Shot 3D Visual Illusion Generation via... ↗ cs.CV 4 Jun 18, 2026
TimeProVe: Propose, then Verify for Efficient Long Video Tempora... ↗ cs.CV 5 Jun 18, 2026
UNIEGO: Proxies as Mediators for Unified Egocentric Video Repres... ↗ cs.CV 5 Jun 18, 2026
Thinking in Boxes: 3D Editing in Real Images Made Easy ↗ cs.CV 7 Jun 18, 2026
Current World Models Lack a Persistent State Core ↗ cs.CV 11 Jun 18, 2026
SSD: Spatially Speculative Decoding Accelerates Autoregressive I... ↗ cs.CV 4 Jun 18, 2026
CalTennis: Large Multi-View Tennis Video Dataset and Benchmark o... ↗ cs.CV 5 Jun 18, 2026
The FID Lottery: Quantifying Hidden Randomness in Generative-Mod... ↗ cs.CV 3 Jun 18, 2026
VisDom: Sparse Novel View Synthesis with Visible Domain Constrai... ↗ cs.CV 6 Jun 18, 2026
SARLO-80: Worldwide Slant SAR Language Optic Dataset 80cm ↗ cs.CV 5 Jun 18, 2026
HumanScale: Egocentric Human Video Can Outperform Real-Robot Dat... ↗ cs.CV 22 Jun 18, 2026
S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intellig... ↗ cs.CV 13 Jun 18, 2026
FreeStyle: Free Control of Style-Content Dual-Reference Generati... ↗ cs.CV 13 Jun 18, 2026
How Fragile Are Training-Free AI-Generated Image Detectors? A Co... ↗ cs.CV 2 Jun 18, 2026
Scalable Training of Spatially Grounded 2D Vision-Language Model... ↗ cs.CV 7 Jun 18, 2026
PCFootprint: A Large-Scale Dataset and Benchmark for Vectorized... ↗ cs.CV 4 Jun 18, 2026
InfantFace: Detecting infant faces in neonatal clinical environm... ↗ cs.CV 5 Jun 18, 2026
Spectral Query-Key Product Weight Steering for Training-Free VLM... ↗ cs.CV 3 Jun 18, 2026
FlowBender: Feedback-Aware Training for Self-Correcting Conditio... ↗ cs.CV 4 Jun 18, 2026
Geometry-Aware Superpixel Graph Transformer with Metadata for Sk... ↗ cs.CV 4 Jun 18, 2026
Reliability-Aware Prototype Calibration for Frozen Pose-Flow Vid... ↗ cs.CV 6 Jun 18, 2026
Through the PRISM: Preference Representation in Intermediate Sta... ↗ cs.CV 6 Jun 18, 2026
GEN-Guard: Correcting Generalization Failures for Deployable Fed... ↗ cs.CV 4 Jun 18, 2026
CUPID: Reconstructing UV Texture Maps for Interpretable Person-o... ↗ cs.CV 5 Jun 18, 2026
CMDS-AD: Cross-Modal Dual-Stream Decoupling for Few-Shot Anomaly... ↗ cs.CV 7 Jun 18, 2026
U$^2$Mamba: A Two-level Nested U-structure Mamba for Salient Obj... ↗ cs.CV 3 Jun 18, 2026
Single-Stage Hierarchical Rectification for Weakly Supervised Hi... ↗ cs.CV 4 Jun 18, 2026
SPOT-E: Test-Time Entropy Shaping with Visual Spotlights for Fro... ↗ cs.CV 9 Jun 18, 2026
BAFIS: Dataset + Framework to assess occupational Bias and Human... ↗ cs.CV 3 Jun 18, 2026
DeepForestVisionV2: Ecology-Driven Taxonomy Expansion for Camera... ↗ cs.CV 27 Jun 18, 2026
Evaluation of Image Matching for Art Skills Assessment ↗ cs.CV 4 Jun 18, 2026
Distill Once, Adapt Life-Long: Exploring Dataset Distillation fo... ↗ cs.CV 4 Jun 18, 2026
HilDA: Hierarchical Distillation with Diffusion for Advancing Se... ↗ cs.CV 7 Jun 18, 2026
Evaluating and Enhancing Negation Comprehension in Remote Sensin... ↗ cs.CV 4 Jun 18, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.