-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 87 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 234 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 161 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
Collections
Discover the best community collections!
Collections including paper arxiv:2608.07110
-
LongCat-Flash-Thinking-2601 Technical Report
Paper • 2601.16725 • Published • 184 -
DeepSeek-OCR 2: Visual Causal Flow
Paper • 2601.20552 • Published • 74 -
Linear representations in language models can change dramatically over a conversation
Paper • 2601.20834 • Published • 21 -
BMAM: Brain-inspired Multi-Agent Memory Framework
Paper • 2601.20465 • Published • 6
-
Contrastive Decoding Improves Reasoning in Large Language Models
Paper • 2309.09117 • Published • 40 -
Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
Paper • 2310.08491 • Published • 57 -
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
Paper • 2411.04282 • Published • 37 -
Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models
Paper • 2411.14432 • Published • 26
-
MMDiff: Extending Diffusion Transformers for Multi-Modal Generation
Paper • 2606.16673 • Published • 5 -
Tent: Fully Test-time Adaptation by Entropy Minimization
Paper • 2006.10726 • Published -
Test-Time Training with Self-Supervision for Generalization under Distribution Shifts
Paper • 1909.13231 • Published • 1 -
Parallel Rollout Approximation for Pixel-Space Autoregressive Image Generation
Paper • 2606.27978 • Published • 6
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.76k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
Continuous Latent Diffusion Language Model
Paper • 2605.06548 • Published • 87 -
Scaling Latent Reasoning via Looped Language Models
Paper • 2510.25741 • Published • 234 -
Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Paper • 2502.05171 • Published • 161 -
Pretraining Language Models to Ponder in Continuous Space
Paper • 2505.20674 • Published • 3
-
MMDiff: Extending Diffusion Transformers for Multi-Modal Generation
Paper • 2606.16673 • Published • 5 -
Tent: Fully Test-time Adaptation by Entropy Minimization
Paper • 2006.10726 • Published -
Test-Time Training with Self-Supervision for Generalization under Distribution Shifts
Paper • 1909.13231 • Published • 1 -
Parallel Rollout Approximation for Pixel-Space Autoregressive Image Generation
Paper • 2606.27978 • Published • 6
-
LongCat-Flash-Thinking-2601 Technical Report
Paper • 2601.16725 • Published • 184 -
DeepSeek-OCR 2: Visual Causal Flow
Paper • 2601.20552 • Published • 74 -
Linear representations in language models can change dramatically over a conversation
Paper • 2601.20834 • Published • 21 -
BMAM: Brain-inspired Multi-Agent Memory Framework
Paper • 2601.20465 • Published • 6
-
lusxvr/nanoVLM-222M
Image-Text-to-Text • 0.2B • Updated • 1.76k • 103 -
Search-R1: Training LLMs to Reason and Leverage Search Engines with Reinforcement Learning
Paper • 2503.09516 • Published • 41 -
AlphaOne: Reasoning Models Thinking Slow and Fast at Test Time
Paper • 2505.24863 • Published • 98 -
QwenLong-L1: Towards Long-Context Large Reasoning Models with Reinforcement Learning
Paper • 2505.17667 • Published • 89
-
Contrastive Decoding Improves Reasoning in Large Language Models
Paper • 2309.09117 • Published • 40 -
Prometheus: Inducing Fine-grained Evaluation Capability in Language Models
Paper • 2310.08491 • Published • 57 -
Language Models are Hidden Reasoners: Unlocking Latent Reasoning Capabilities via Self-Rewarding
Paper • 2411.04282 • Published • 37 -
Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models
Paper • 2411.14432 • Published • 26