-
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 191 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 166 -
Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation
Paper • 2605.03849 • Published • 125 -
ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration
Paper • 2605.03042 • Published • 124
Collections
Discover the best community collections!
Collections including paper arxiv:2605.03849
-
Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation
Paper • 2605.03849 • Published • 125 -
HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation
Paper • 2604.28196 • Published • 72 -
SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies
Paper • 2605.04637 • Published • 3
-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Paper • 2402.04252 • Published • 31 -
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
Paper • 2402.03749 • Published • 15 -
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 24
-
SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation
Paper • 2503.09641 • Published • 42 -
Decoupled DMD: CFG Augmentation as the Spear, Distribution Matching as the Shield
Paper • 2511.22677 • Published • 35 -
One-step Diffusion with Distribution Matching Distillation
Paper • 2311.18828 • Published • 3 -
Improved Distribution Matching Distillation for Fast Image Synthesis
Paper • 2405.14867 • Published • 15
-
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 191 -
From Context to Skills: Can Language Models Learn from Context Skillfully?
Paper • 2604.27660 • Published • 166 -
Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation
Paper • 2605.03849 • Published • 125 -
ARIS: Autonomous Research via Adversarial Multi-Agent Collaboration
Paper • 2605.03042 • Published • 124
-
Stream-R1: Reliability-Perplexity Aware Reward Distillation for Streaming Video Generation
Paper • 2605.03849 • Published • 125 -
HERMES++: Toward a Unified Driving World Model for 3D Scene Understanding and Generation
Paper • 2604.28196 • Published • 72 -
SWE-WebDevBench: Evaluating Coding Agent Application Platforms as Virtual Software Agencies
Paper • 2605.04637 • Published • 3
-
SANA-Sprint: One-Step Diffusion with Continuous-Time Consistency Distillation
Paper • 2503.09641 • Published • 42 -
Decoupled DMD: CFG Augmentation as the Spear, Distribution Matching as the Shield
Paper • 2511.22677 • Published • 35 -
One-step Diffusion with Distribution Matching Distillation
Paper • 2311.18828 • Published • 3 -
Improved Distribution Matching Distillation for Fast Image Synthesis
Paper • 2405.14867 • Published • 15
-
EVA-CLIP-18B: Scaling CLIP to 18 Billion Parameters
Paper • 2402.04252 • Published • 31 -
Vision Superalignment: Weak-to-Strong Generalization for Vision Foundation Models
Paper • 2402.03749 • Published • 15 -
ScreenAI: A Vision-Language Model for UI and Infographics Understanding
Paper • 2402.04615 • Published • 45 -
EfficientViT-SAM: Accelerated Segment Anything Model Without Performance Loss
Paper • 2402.05008 • Published • 24