Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts Paper • 2607.00666 • Published Jul 1 • 25
Are We Measuring Strategy or Phrasing? The Gap Between Surface- and Approach-Level Diversity in LLM Math Reasoning Paper • 2606.29985 • Published Jun 29 • 20
PoLAR: Factorizing Extent and Mode in Latent Actions for Robot Policy Learning Paper • 2606.21139 • Published Jun 19 • 13
CIPER: A Unified Framework for Cross-view Image-retrieval and Pose-estimation Paper • 2606.05011 • Published Jun 3
Human Psychometric Questionnaires Mischaracterize LLM Behavior Paper • 2509.10078 • Published May 29 • 36
RobotValues: Evaluating Household Robots When Human Values Conflict Paper • 2606.03312 • Published Jun 2 • 26
ArcANE: Do Role-Playing Language Agents Stay in Character at the Right Time? Paper • 2606.05553 • Published Jun 4 • 50
One Click per Cell Type Suffices: Training-free Group Interaction for Cell Instance Segmentation Paper • 2605.29429 • Published May 28 • 8
ResearchMath-14K: Scaling Research-Level Mathematics via Agents Paper • 2605.28003 • Published May 27 • 50
Self-Improving CAD Generation Agents with Finite Element Analysis as Feedback Paper • 2605.17448 • Published May 17 • 20
Rule2DRC: Benchmarking LLM Agents for DRC Script Synthesis with Execution-Guided Test Generation Paper • 2605.15669 • Published May 15 • 8
KL for a KL: On-Policy Distillation with Control Variate Baseline Paper • 2605.07865 • Published May 8 • 22
Your Language Model is Its Own Critic: Reinforcement Learning with Value Estimation from Actor's Internal States Paper • 2605.07579 • Published May 8 • 18
Training-Free Dense Hand Contact Estimation with Multi-Modal Large Language Models Paper • 2605.05886 • Published May 7 • 3
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization Paper • 2605.09996 • Published May 11 • 8
Vanast: Virtual Try-On with Human Image Animation via Synthetic Triplet Supervision Paper • 2604.04934 • Published Apr 6 • 48
4DGS360: 360° Gaussian Reconstruction of Dynamic Objects from a Single Video Paper • 2603.21618 • Published Mar 23 • 15
Uncertainty-guided Compositional Alignment with Part-to-Whole Semantic Representativeness in Hyperbolic Vision-Language Models Paper • 2603.22042 • Published Mar 23 • 3
Coherent Human-Scene Reconstruction from Multi-Person Multi-View Video in a Single Pass Paper • 2603.12789 • Published Mar 13 • 4