Results
Per task
How Kapso runs it
Three stages, with the campaign clock starting at brief-in.- Preflight. One agent session ingests the official task brief, downloads the data, and writes the task statement.
- Campaign. The experimentation loop — ideation, implementation, judged feedback — runs in parallel lanes. Each lane cycles submit-and-learn rounds through the official submission system: predict the score, submit, bank the result, study the gap, go again.
- Shared learning. Lanes learn from every sibling submission on the board, and ideas are grounded in a lesson bank distilled from past olympiad tasks.
Usage
Layout
Full integration notes are in
benchmarks/ioai2026/.
Related
MLE-Bench
Kaggle machine-learning competitions
RelBench
Predictive tasks over relational databases
pip install leeroo-kapso · Every page as plain text: llms.txt.