🌌 To an Unceasing Future. 致永无止境的明天
@yuchenwang3 | LLM post training
🌌 To an Unceasing Future. 致永无止境的明天
@yuchenwang3 | LLM post training
A high-throughput and memory-efficient inference and serving engine for LLMs
Ongoing research training transformer models at scale
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …
An LLM post-training framework with vLLM for RL Scaling
Scalable toolkit for efficient model reinforcement
Developer Asset Hub for NVIDIA Nemotron — A one-stop resource for training recipes, usage cookbooks, datasets, and full end-to-end reference examples to build with Nemotron models