Back to AI intel
重点
Training open models with RL is now possible across multiple coding environments like Claude Code
AI intel briefing
Core summary
One sentence to understand this update
Ben Burtenshaw announces that open models can now be trained with Reinforcement Learning (RL) within various coding environments such as Claude Code, codex, opencode, or Pi, using unmodified harnesses.
Impact & opportunity
What this could mean
This advancement simplifies the RL training workflow for developers, allowing for greater flexibility and efficiency when experimenting with and fine-tuning open AI models across different platforms.
Source
View original