Harness-native RL for coding agents: the trained policy and the training task index.
-
LEGO-RL: Harness-Native Reinforcement Learning for Coding Agents
Paper • 2608.17393 • Published • 25 -
Lego-X/qwen3_5_35b_a3b_ohsdk_200k_rl
Text Generation • 36B • Updated • 683 • 3 -
Lego-X/qwen3_5_35b_a3b_cc_200k_rl
Text Generation • 36B • Updated • 495 • 1 -
Lego-X/qwen3_5_35b_a3b_oc_200k_rl
Text Generation • 36B • Updated • 492 • 1