Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

515 results

PaperCategoryAuthorsDate
ARTEMIS: Agent-guided Reliability-aware Temporal Mask Evolution... ↗ cs.CV 7 Jun 18, 2026
NAMESAKES: Probing Identity Memorization in Text-to-Image Models ↗ cs.CV 5 Jun 18, 2026
HEad and neCK TumOR (HECKTOR) 2025: Benchmark of Segmentation, D... ↗ cs.CV 30 Jun 18, 2026
SA-VIS: Sparse frame Annotations for training Video Instance Seg... ↗ cs.CV 5 Jun 18, 2026
TriFlow: Generating Artist-Like 3D Mesh Topology via Nearest-Ver... ↗ cs.CV 7 Jun 18, 2026
SAM3 Self-Distillation for Fine-Grained GOOSE 2D Semantic Segmen... ↗ cs.CV 1 Jun 18, 2026
InterleaveThinker: Reinforcing Agentic Interleaved Generation ↗ cs.CV 7 Jun 11, 2026
Modality Forcing for Scalable Spatial Generation ↗ cs.CV 5 Jun 11, 2026
RepWAM: World Action Modeling with Representation Visual-Action... ↗ cs.CV 8 Jun 11, 2026
SpatialClaw: Rethinking Action Interface for Agentic Spatial Rea... ↗ cs.CV 11 Jun 11, 2026
Flex4DHuman: Flexible Multi-view Video Diffusion for 4D Human Re... ↗ cs.CV 5 Jun 11, 2026
World Tracing: Generative Pixel-Aligned Geometry Beyond the Visi... ↗ cs.CV 9 Jun 11, 2026
Surflo: Consistent 3D Surface Flow Model with Global State ↗ cs.CV 6 Jun 11, 2026
Revisiting Vehicle Color Recognition in Long-Tailed Surveillance... ↗ cs.CV 5 Jun 11, 2026
Towards Effective Waste Segmentation for Automated Waste Recycli... ↗ cs.CV 6 Jun 11, 2026
EvTexture++: Event-Driven Texture Enhancement for Video Super-Re... ↗ cs.CV 4 Jun 11, 2026
Contrast-Informed Augmentation and Domain-Adversarial Training f... ↗ cs.CV 4 Jun 11, 2026
Edit the Bits, Diff the Codes: Bitwise Residual Editing for Visu... ↗ cs.CV 5 Jun 11, 2026
What's Old is New Again: Classical Dimensionality Reduction for... ↗ cs.CV 2 Jun 11, 2026
MaskWAM: Unifying Mask Prompting and Prediction for World-Action... ↗ cs.CV 7 Jun 11, 2026
Measurement-Calibrated Multi-Camera Fusion for Vision-Based Indo... ↗ cs.CV 3 Jun 11, 2026
Heterogeneous LiDAR Early Fusion and Learned Re-Ranking Strategy... ↗ cs.CV 5 Jun 11, 2026
Budget-Constrained Step-Level Diffusion Caching ↗ cs.CV 4 Jun 11, 2026
Point-Wise Geometry-Aware Transformer for Partial-to-Full Point... ↗ cs.CV 2 Jun 11, 2026
VISA: VLM-Guided Instance Semantic Auditing for 3D Occupancy Wor... ↗ cs.CV 5 Jun 11, 2026
OmniDirector: General Multi-Shot Camera Cloning without Cross-Pa... ↗ cs.CV 11 Jun 11, 2026
VietFashion: Benchmarking Sketch-Text Composed Image Retrieval f... ↗ cs.CV 5 Jun 11, 2026
Person Identification from Contextual Motion ↗ cs.CV 3 Jun 11, 2026
SmartFont: Dynamic Condition Allocation for Few-Shot Font Genera... ↗ cs.CV 2 Jun 11, 2026
MoVerse: Real-Time Video World Modeling with Panoramic Gaussian... ↗ cs.CV 7 Jun 11, 2026
Dual-Constrained Diffusion Image Compression for Operational Rat... ↗ cs.CV 3 Jun 11, 2026
JointEdit3D: Feed-Forward 3D Scene Editing in a Unified Latent S... ↗ cs.CV 7 Jun 11, 2026
Dual-Domain Equivariant Generative Adversarial Network for Multi... ↗ cs.CV 3 Jun 11, 2026
OR-Action: Multi-Role Video Understanding with Fine-Grained Acti... ↗ cs.CV 6 Jun 11, 2026
Masked and Predictive Self-Supervised Foundation Models for 3D B... ↗ cs.CV 4 Jun 11, 2026
MagPlus: Bridging Micro-to-Regular Facial Expressions through Le... ↗ cs.CV 2 Jun 11, 2026
ReFree: Towards Realistic Co-Speech Video Generation via Reward-... ↗ cs.CV 4 Jun 11, 2026
DuET: Dual Expert Trajectories for Diffusion Image Editing ↗ cs.CV 3 Jun 11, 2026
HYDRA-X: Native Unified Multimodal Models with Holistic Visual T... ↗ cs.CV 14 Jun 11, 2026
Cross-Modal Masked Compositional Concept Modeling for Enhancing... ↗ cs.CV 3 Jun 11, 2026
Zero-Shot Captioning for Cultural Heritage: Automated Image Anal... ↗ cs.CV 3 Jun 11, 2026
TimeLens: On-Device Artifact Recognition with Retrieval-Augmente... ↗ cs.CV 6 Jun 11, 2026
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scen... ↗ cs.CV 5 Jun 4, 2026
Complexity-Balanced Diffusion Splitting ↗ cs.CV 3 Jun 4, 2026
AI that "imagines" unseen space to reason — "Astra," a spatial-r... Explained cs.CV 7 Jun 4, 2026
A Vision-language Framework for Comparative Reasoning in Radiolo... ↗ cs.CV 8 Jun 4, 2026
HomeWorld: A Unified Floorplan-to-Furnished Framework for Genera... ↗ cs.CV 5 Jun 4, 2026
Leave the model alone and amplify what passes through it — a plu... Explained cs.CV 9 Jun 4, 2026
Visual Commonsense Driven Knowledge Refinements for Scene Graph... ↗ cs.CV 5 Jun 4, 2026
GMBFormer: An NDVI-Guided Global Memory Bank Transformer for Urb... ↗ cs.CV 7 Jun 4, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.