Collections
Discover the best community collections!
Collections including paper arxiv:2604.14268
-
LLM Pruning and Distillation in Practice: The Minitron Approach
Paper • 2408.11796 • Published • 62 -
Less Gaussians, Texture More: 4K Feed-Forward Textured Splatting
Paper • 2603.25745 • Published • 16 -
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
Paper • 2604.14268 • Published • 126 -
Easy3E: Feed-Forward 3D Asset Editing via Rectified Voxel Flow
Paper • 2602.21499 • Published • 1
-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 117 -
Cambrian-S: Towards Spatial Supersensing in Video
Paper • 2511.04670 • Published • 40 -
MagicWorld: Interactive Geometry-driven Video World Exploration
Paper • 2511.18886 • Published • 18 -
WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling
Paper • 2512.14614 • Published • 72
-
infly/OpenCoder-8B-Instruct
Text Generation • 8B • Updated • 1.01k • • 208 -
infly/OpenCoder-8B-Base
Text Generation • 8B • Updated • 104 • 32 -
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
Paper • 2406.11931 • Published • 71 -
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Paper • 2401.14196 • Published • 74
-
Mobile-O: Unified Multimodal Understanding and Generation on Mobile Device
Paper • 2602.20161 • Published • 23 -
A Very Big Video Reasoning Suite
Paper • 2602.20159 • Published • 201 -
Speed by Simplicity: A Single-Stream Architecture for Fast Audio-Video Generative Foundation Model
Paper • 2603.21986 • Published • 125 -
AURA: Always-On Understanding and Real-Time Assistance via Video Streams
Paper • 2604.04184 • Published • 50
-
Towards Scalable Pre-training of Visual Tokenizers for Generation
Paper • 2512.13687 • Published • 108 -
MMGR: Multi-Modal Generative Reasoning
Paper • 2512.14691 • Published • 121 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100 -
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
Paper • 2512.23576 • Published • 66
-
LLM Pruning and Distillation in Practice: The Minitron Approach
Paper • 2408.11796 • Published • 62 -
Less Gaussians, Texture More: 4K Feed-Forward Textured Splatting
Paper • 2603.25745 • Published • 16 -
HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds
Paper • 2604.14268 • Published • 126 -
Easy3E: Feed-Forward 3D Asset Editing via Rectified Voxel Flow
Paper • 2602.21499 • Published • 1
-
Mobile-O: Unified Multimodal Understanding and Generation on Mobile Device
Paper • 2602.20161 • Published • 23 -
A Very Big Video Reasoning Suite
Paper • 2602.20159 • Published • 201 -
Speed by Simplicity: A Single-Stream Architecture for Fast Audio-Video Generative Foundation Model
Paper • 2603.21986 • Published • 125 -
AURA: Always-On Understanding and Real-Time Assistance via Video Streams
Paper • 2604.04184 • Published • 50
-
Emu3.5: Native Multimodal Models are World Learners
Paper • 2510.26583 • Published • 117 -
Cambrian-S: Towards Spatial Supersensing in Video
Paper • 2511.04670 • Published • 40 -
MagicWorld: Interactive Geometry-driven Video World Exploration
Paper • 2511.18886 • Published • 18 -
WorldPlay: Towards Long-Term Geometric Consistency for Real-Time Interactive World Modeling
Paper • 2512.14614 • Published • 72
-
Towards Scalable Pre-training of Visual Tokenizers for Generation
Paper • 2512.13687 • Published • 108 -
MMGR: Multi-Modal Generative Reasoning
Paper • 2512.14691 • Published • 121 -
Coupling Experts and Routers in Mixture-of-Experts via an Auxiliary Loss
Paper • 2512.23447 • Published • 100 -
LiveTalk: Real-Time Multimodal Interactive Video Diffusion via Improved On-Policy Distillation
Paper • 2512.23576 • Published • 66
-
infly/OpenCoder-8B-Instruct
Text Generation • 8B • Updated • 1.01k • • 208 -
infly/OpenCoder-8B-Base
Text Generation • 8B • Updated • 104 • 32 -
DeepSeek-Coder-V2: Breaking the Barrier of Closed-Source Models in Code Intelligence
Paper • 2406.11931 • Published • 71 -
DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Paper • 2401.14196 • Published • 74