Call for Contribution
Last update: Aug 6, 2026
RSI Bench is built by and for the AI research community. If you're working at the frontier of ML research, we want your hardest, most recent problems, the ones that would genuinely test whether an AI agent can do your job.
What you get
- Co-authorship
- on the RSI Bench research paper. The author position reflects the scope and difficulty of what you contribute.
- $2,000 per accepted task
- (for the initial cohort of 50 tasks).
- Modal compute credits,
- to cover the GPU and compute costs of building and iterating on your task.
- Access to the RSI Bench research community:
- once your initial proposal is selected, we'll invite you to the RSI Bench Community Slack, connecting you with other selected contributors and researchers across our networks.
Task structure
Every task is made up of four components:
- Instructions:
- the agent-facing description of what to attempt.
- Baseline:
- a strong reference solution that sets the bar to beat.
- Environment:
- a container with the agent's workspace, and the read-only baseline artifacts and reference assets.
- Verifier:
- the script that scores a submission after the agent's session ends.
More details coming soon.
How it works
- Submit a short proposal describing the task you'd contribute, your background, the sub-domain, and what the task involves.
- If approved, you'll get computer access to build it (or use your own setup).
- Build and submit the task.
- Every submitted task goes through review, by the community and the Scale team, before being accepted into the benchmark.
Full contribution guidelines, including task format and submission details will be released soon.