Dear author,
Thanks for releasing Harness-R1—the code, patch protocol, and both engineer checkpoints have been very helpful.
I am trying to reproduce the Agent SFT + Harness-R1 row in Table 1 (59.2 → 64.2). The README notes that agent-sft-harness-r1 is only meaningful against the frozen agent-SFT Qwen3.5-9B target it was trained on, but that target is not in the release. Without the same weights, we cannot pair the public engineer with a matching actor, so the number is not reproducible even though the engineer itself is available.
Would you consider releasing the agent-SFT target checkpoint (the Qwen3.5-9B after the 2,515-trajectory SFT in Appendix B.3)?
Thank you for considering this.
Dear author,
Thanks for releasing Harness-R1—the code, patch protocol, and both engineer checkpoints have been very helpful.
I am trying to reproduce the Agent SFT + Harness-R1 row in Table 1 (59.2 → 64.2). The README notes that agent-sft-harness-r1 is only meaningful against the frozen agent-SFT Qwen3.5-9B target it was trained on, but that target is not in the release. Without the same weights, we cannot pair the public engineer with a matching actor, so the number is not reproducible even though the engineer itself is available.
Would you consider releasing the agent-SFT target checkpoint (the Qwen3.5-9B after the 2,515-trajectory SFT in Appendix B.3)?
Thank you for considering this.