🚀NEW LABGetting Started with Claude AgentsStart lab
Training · Reasoning

Post Training of LLMs

First page
Post Training of LLMs
Paper summary

PoLMs like OpenAI-o1/o3 and DeepSeek-R1 tackle LLM shortcomings in reasoning, ethics, and specialized tasks. This survey tracks their evolution and provides a taxonomy of techniques across fine-tuning, alignment, reasoning, efficiency, and integration, guiding progress toward more robust, versatile AI.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack