-
A Picture is Worth More Than 77 Text Tokens: Evaluating CLIP-Style Models on Dense Captions
Paper • 2312.08578 • Published • 20 -
ZeroQuant(4+2): Redefining LLMs Quantization with a New FP6-Centric Strategy for Diverse Generative Tasks
Paper • 2312.08583 • Published • 11 -
Vision-Language Models as a Source of Rewards
Paper • 2312.09187 • Published • 12 -
StemGen: A music generation model that listens
Paper • 2312.08723 • Published • 49
Chuanming Liu
Chuanming
AI & ML interests
Artificial Intelligence, AGI, NLP, LLMs, Multimodality, MLSys. Python/Golang/C/C++/Shell/awk&sed
Recent Activity
liked
a model
about 15 hours ago
mlx-community/Qwen3-TTS-12Hz-1.7B-Base-bf16
upvoted
an
article
3 days ago
Scaling Real-Time Voice Agents with Cache-Aware Streaming ASR
liked
a model
4 days ago
PaddlePaddle/PaddleOCR-VL-1.5