Evidence-RL: Towards Evidence-intensive Visual Reasoning Paper • 2608.08021 • Published 4 days ago • 8
InfiniSplat: Implicit Gaussian Decoding for Large-Baseline Monocular View Synthesis Paper • 2608.02437 • Published 9 days ago • 66
MatrAIx: Simulating the World with 8.3 Billion Persona Agents Paper • 2608.04205 • Published 8 days ago • 22
Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution Paper • 2608.08311 • Published 4 days ago • 52
Agent Memory Distillation: Empowering Small LLM Agents with Hierarchical Teacher Memory Paper • 2608.07169 • Published 5 days ago • 26
SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring Paper • 2608.09802 • Published 1 day ago • 104
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published 1 day ago • 91
SFT Conflicts, RL Coexists: A Theoretical and Empirical Analysis of Multi-Task Learning for LLMs Paper • 2608.03573 • Published 6 days ago • 44
Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning Paper • 2608.03571 • Published 6 days ago • 39
SimWAM: A Simple World Action Model for End-to-End Autonomous Driving Paper • 2608.07468 • Published 5 days ago • 22
YOLO-PEFT: Parameter-Efficient Fine-Tuning on YOLO Family Paper • 2608.07051 • Published 5 days ago • 15
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published 6 days ago • 43
Towards Physics of Multimodal Pretraining: Knowledge Flow, Modality Synergy, Early Unification, and Recipes Paper • 2608.05000 • Published 6 days ago • 58
Invisible Shortcuts: Why Vision Encoders Know Your Camera Paper • 2608.05424 • Published 7 days ago • 16
SmartMage: Dynamic Modality Orchestration for 3D Scene Understanding Paper • 2608.05137 • Published 1 day ago • 25