MC-Sparse: Deconstructing and Closing the Dense-Sparse Attention Gap in Diffusion Transformers Paper • 2610.06801 • Published 6 days ago • 33
Does Native 3D Texture Generation Necessarily Require 3D Assets for Training? Paper • 2609.34621 • Published 13 days ago • 10
Does Native 3D Texture Generation Necessarily Require 3D Assets for Training? Paper • 2609.34621 • Published 13 days ago • 10
FastVAR: Linear Visual Autoregressive Modeling via Cached Token Pruning Paper • 2503.23367 • Published Mar 30, 2025 • 1
COVE: Unleashing the Diffusion Feature Correspondence for Consistent Video Editing Paper • 2406.08850 • Published Jun 13, 2024
GrootVL: Tree Topology is All You Need in State Space Model Paper • 2406.02395 • Published Jun 4, 2024 • 1
Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models Paper • 2604.25636 • Published Apr 28 • 25
Efficient Autoregressive Video Diffusion with Dummy Head Paper • 2601.20499 • Published Jan 28 • 8
Does Native 3D Texture Generation Necessarily Require 3D Assets for Training? Paper • 2609.34621 • Published 13 days ago • 10
SpatialEdit: Benchmarking Fine-Grained Image Spatial Editing Paper • 2604.04911 • Published Apr 6 • 33
ViGoR-Bench: How Far Are Visual Generative Models From Zero-Shot Visual Reasoners? Paper • 2603.25823 • Published Mar 26 • 43
Efficient Autoregressive Video Diffusion with Dummy Head Paper • 2601.20499 • Published Jan 28 • 8
Running on Zero Agents 67 RF-Solver-Edit 🪄 67 High Quality Inversion and Editing of FLUX and OpenSora.
Running on Zero Agents 67 RF-Solver-Edit 🪄 67 High Quality Inversion and Editing of FLUX and OpenSora.