Liangyu Wang
ly4096
AI & ML interests
Efficient reinforcement learning (RL) for LLMs reasoning
Distributed training and inference of LLMs
Efficient algorithm and infrastructure design for LLMs
Recent Activity
liked a model about 2 hours ago
Qwen/Qwen3.8-2.4T-A95B-FP8 upvoted a paper 3 months ago
SlimQwen: Exploring the Pruning and Distillation in Large MoE Model Pre-training authored a paper 6 months ago
Canzona: A Unified, Asynchronous, and Load-Balanced Framework for Distributed Matrix-based Optimizers