Harness-native RL for coding agents: the trained policy and the training task index.
-
Lego-X/qwen3_5_35b_a3b_base_ohsdk_200k_rl
Text Generation • 36B • Updated • 11 • 2 -
Lego-X/Lego-RL-2699
Viewer • Updated • 2.7k • 42 • 2 -
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents
Paper • 2608.17393 • Published • 24 -
Lego-X/qwen3_5_35b_a3b_base_cc_200k_rl
Text Generation • 36B • Updated • 1