MaLiang-Harness: A Programmable Path to Image and Video Generation Paper • 2609.34309 • Published 13 days ago • 394
VisionHOPE: Visual Backbones as Self-Modifying Learning Systems Paper • 2609.33325 • Published 14 days ago • 318
Beyond Dyadic Memory: Interaction-Aware Multimodal Memory with Adaptive Agentic Retrieval for Multi-Party Spoken Conversations Paper • 2609.32522 • Published 15 days ago • 71
YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality Paper • 2609.33757 • Published 14 days ago • 234
Vidu S2: Real-Time Interactive, Editable, and Spatial Video Generation Paper • 2609.11638 • Published Sep 10 • 657
NCP-ArchPreview Technical Report: Moving towards Latent Space Language Models through Next Concept Prediction Paper • 2609.10715 • Published Sep 9 • 269