Topic

Reinforcement Learning

2 pieces in this thread.

  1. A World Worth Learning From

    Before an agent can learn from experience, its environment has to produce experience worth learning from. A sixteen-run engineering study tested what survives after the agent commits.

    harness-engineeringagent-evaluationaec-benchtask-worlds
  2. Making aec-bench Trainable with Prime Lab

    How aec-bench and Prime Intellect's Lab turn engineering benchmarks into verifier-backed RL environments, adapter training runs, and inspectable traces.

    aec-benchprime-labreinforcement-learningagent-evaluation