AI that builds
better AI.
Autonomously.
Each day it comes up with an idea, implements it, evaluates it in the environment, and learns from every experiment, all on its own. Or drop in your own idea, and it biases the search toward that.
Autonomous AI research loop.
Better AI system, built by an AI researcher that runs the whole loop: ideate, implement, experiment, evaluate, learn, repeat. Every attempt is preserved and every result feeds a growing archive of knowledge that steers the next idea, so the loop runs open-ended toward a measurably better system.
POWERED BY KAPSO ↗SCROLL TO EXPLORE THE FULL LOOP →
Verified in the open.
Machine-Learning Engineering
Given a new problem, it builds the entire solution on its own, from raw data to a trained, tuned model, and outperforms every other open agent.
MEDAL RATE
MEDAL RATE BY TASK COMPLEXITY
BEST REPORTED PER TIER: R&D-AGENT (LOW) · AIRA-DOJO (MEDIUM) · ML-MASTER (HIGH)
Long-Horizon Algorithm Design
The problems have no perfect answer. It designs an algorithm, then rewrites it for hours to push the score higher, head to head with human experts.
Open-Ended Optimization Problem.
A sample of the loop on the hardest class of problems: open-ended optimization with no known best answer. This one is drawn from a live AtCoder contest. Encode up to a hundred messages as networks; the channel randomly rewires each one and erases its labels; the receiver must still decode it. There is no formula for the best design. The loop invented its own, rated on the contest's human ladder.
VIEW THE PROBLEM ON ATCODER ↗AVERAGE HUMAN 1260 · ALE-BENCH
A better support agent, for every domain.
DOCUMENT RETRIEVAL
SCORE▲ 24 PTS
BOOKINGS & CHANGES
SCORE▲ 25 PTS
ORDERS & RETURNS
SCORE▲ 4 PTS
Banking knowledge.
LIVE RUN · DOCUMENT RETRIEVALBASELINE 15% · TAU-BENCH · THE BASE MODEL ON THIS BENCHMARK
Give the researcher an idea to try.
IT BIASES THE NEXT EXPERIMENT TOWARD YOUR IDEA