Collections
Discover the best community collections!
Collections including paper arxiv:2604.06870
-
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44 -
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
Paper • 2604.04934 • Published • 43 -
VOID: Video Object and Interaction Deletion
Paper • 2604.02296 • Published • 54 -
FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On
Paper • 2604.08526 • Published • 18
-
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 337 -
Fantastic Reasoning Behaviors and Where to Find Them: Unsupervised Discovery of the Reasoning Process
Paper • 2512.23988 • Published • 19 -
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
Paper • 2512.25075 • Published • 16 -
Guiding a Diffusion Transformer with the Internal Dynamics of Itself
Paper • 2512.24176 • Published • 8
-
WildDet3D: Scaling Promptable 3D Detection in the Wild
Paper • 2604.08626 • Published • 95 -
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44 -
ClawBench: Can AI Agents Complete Everyday Online Tasks?
Paper • 2604.08523 • Published • 377
-
Prompt-Free Universal Region Proposal Network
Paper • 2603.17554 • Published • 4 -
VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images
Paper • 2604.09531 • Published • 10 -
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44
-
WildDet3D: Scaling Promptable 3D Detection in the Wild
Paper • 2604.08626 • Published • 95 -
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44 -
ClawBench: Can AI Agents Complete Everyday Online Tasks?
Paper • 2604.08523 • Published • 377
-
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44 -
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision
Paper • 2604.04934 • Published • 43 -
VOID: Video Object and Interaction Deletion
Paper • 2604.02296 • Published • 54 -
FIT: A Large-Scale Dataset for Fit-Aware Virtual Try-On
Paper • 2604.08526 • Published • 18
-
Prompt-Free Universal Region Proposal Network
Paper • 2603.17554 • Published • 4 -
VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images
Paper • 2604.09531 • Published • 10 -
RefineAnything: Multimodal Region-Specific Refinement for Perfect Local Details
Paper • 2604.06870 • Published • 44
-
mHC: Manifold-Constrained Hyper-Connections
Paper • 2512.24880 • Published • 337 -
Fantastic Reasoning Behaviors and Where to Find Them: Unsupervised Discovery of the Reasoning Process
Paper • 2512.23988 • Published • 19 -
SpaceTimePilot: Generative Rendering of Dynamic Scenes Across Space and Time
Paper • 2512.25075 • Published • 16 -
Guiding a Diffusion Transformer with the Internal Dynamics of Itself
Paper • 2512.24176 • Published • 8