Paper club
One paper a fortnight on RL environments, benchmarks, simulators and the things that make them hold up. We read it, someone leads, everyone argues. Sessions are recorded and the notes land here afterwards.
It is open. You do not need to work on this full time, and you do not need to have finished the paper.
Next session
To be announced. Sign-up link goes here.
How it runs
- Format. Forty-five minutes. Ten on what the paper claims, the rest on whether it holds and what it means for anyone building environments.
- Where. A video call; sign up beforehand and the link is emailed to you.
- Recording. Sessions are recorded and embedded on the session page. Say so and we keep you out of the recording.
- Notes. Each session gets a page here linking the paper’s page in the wiki.
Reading so far
Sessions appear here as they happen.
(none yet)
Suggesting a paper
Bring it to the Discord. We are most interested in work on environment generation, verifier soundness, user simulators, and benchmark validity. Recent things on our list are in the ITSMBench and WorldSmith reading lists.