Browse all

arXiv research papers — all

All arXiv papers (AI/ML focus) — filter and search by category and year.

Clear

334 results

PaperCategoryAuthorsDate
JointEdit3D: Feed-Forward 3D Scene Editing in a Unified Latent S... ↗ cs.CV 7 Jun 11, 2026
Dual-Domain Equivariant Generative Adversarial Network for Multi... ↗ cs.CV 3 Jun 11, 2026
OR-Action: Multi-Role Video Understanding with Fine-Grained Acti... ↗ cs.CV 6 Jun 11, 2026
Masked and Predictive Self-Supervised Foundation Models for 3D B... ↗ cs.CV 4 Jun 11, 2026
MagPlus: Bridging Micro-to-Regular Facial Expressions through Le... ↗ cs.CV 2 Jun 11, 2026
ReFree: Towards Realistic Co-Speech Video Generation via Reward-... ↗ cs.CV 4 Jun 11, 2026
DuET: Dual Expert Trajectories for Diffusion Image Editing ↗ cs.CV 3 Jun 11, 2026
HYDRA-X: Native Unified Multimodal Models with Holistic Visual T... ↗ cs.CV 14 Jun 11, 2026
Cross-Modal Masked Compositional Concept Modeling for Enhancing... ↗ cs.CV 3 Jun 11, 2026
Zero-Shot Captioning for Cultural Heritage: Automated Image Anal... ↗ cs.CV 3 Jun 11, 2026
TimeLens: On-Device Artifact Recognition with Retrieval-Augmente... ↗ cs.CV 6 Jun 11, 2026
PAR3D: A Unified 3D-MLLM with Part-Aware Representation for Scen... ↗ cs.CV 5 Jun 4, 2026
Complexity-Balanced Diffusion Splitting ↗ cs.CV 3 Jun 4, 2026
AI that "imagines" unseen space to reason — "Astra," a spatial-r... Explained cs.CV 7 Jun 4, 2026
A Vision-language Framework for Comparative Reasoning in Radiolo... ↗ cs.CV 8 Jun 4, 2026
HomeWorld: A Unified Floorplan-to-Furnished Framework for Genera... ↗ cs.CV 5 Jun 4, 2026
EasyLens: A Training-Free Plug-and-Play Subtle-Lesion Representa... ↗ cs.CV 9 Jun 4, 2026
Visual Commonsense Driven Knowledge Refinements for Scene Graph... ↗ cs.CV 5 Jun 4, 2026
GMBFormer: An NDVI-Guided Global Memory Bank Transformer for Urb... ↗ cs.CV 7 Jun 4, 2026
Physics in 2-Steps: Locking Motion Priors Before Visual Refineme... ↗ cs.CV 6 Jun 4, 2026
Comparison of Deep Learning Frameworks For Rice Disease Mapping... ↗ cs.CV 6 Jun 4, 2026
StoryVideoQA: Scaling Deep Video Understanding with a Large-Scal... ↗ cs.CV 9 Jun 4, 2026
RhymeFlow: Training-Free Acceleration for Video Generation with... ↗ cs.CV 6 Jun 4, 2026
Towards One-to-Many Temporal Grounding ↗ cs.CV 8 Jun 4, 2026
Synthetic Data Generation and Vision-based Wrinkle and Keypoint... ↗ cs.CV 3 Jun 4, 2026
Geodesic Flow Matching on a Riemannian Degradation Manifold for... ↗ cs.CV 6 Jun 4, 2026
GRAMformer: Any-Order Modality Interactions via Volumetric Multi... ↗ cs.CV 3 Jun 4, 2026
SAM-Flow: Source-Anchored Masked Flow for Training-Free Image Ed... ↗ cs.CV 6 Jun 4, 2026
Symb-xMIL: Symbolic Explanations for Multiple Instance Learning... ↗ cs.CV 7 Jun 4, 2026
DisasterBench: A Multimodal Benchmark for UAV-Based Disaster Res... ↗ cs.CV 6 Jun 4, 2026
SC-MFJ: A Simple Haptic Quality Metric for Medical Image Segment... ↗ cs.CV 3 Jun 4, 2026
Adversarial Attacks Already Tell the Answer: Directional Bias-Gu... ↗ cs.CV 8 Jun 4, 2026
RQUL-UIE: Revitalizing Quality-Unstable Labels for Underwater Im... ↗ cs.CV 4 Jun 4, 2026
Adaptive Tokenisation Via Temporal Redundancy Masking And Latent... ↗ cs.CV 6 Jun 4, 2026

Source: official U.S. government open data. This is an organized index, not an official U.S. government site. "Explained" links to our summary page; otherwise links go to the official primary source.

Disclaimer: This site independently summarizes and classifies information based on official data sources. Always verify the latest and accurate information with the official sources. Content on finance, health, legal, and security is information, not advice. This site is not an official website of the U.S. government.