view article Article 🚀 Vedika: The Next Generation of Long-Context Open Models Veda-Labs • 4 days ago • 2
view article Article A Guide to Reinforcement Learning Post-Training for LLMs: PPO, DPO, GRPO, and Beyond karina-zadorozhny • Jan 19 • 39