Crowdsource AI solutions. Pay for verified outcomes.

Bring one measurable problem. We build the test, builders compete on the same inputs, and you pay only when a solution clears the agreed bar.

Three challenges are live now with $1,800 in prizes. One completed round has public results and code.

How it works

From problem to verified result.

01

Define the result

Tell us the input, the useful output, and what solving it is worth.

02

Run one shared test

We agree the data, scoring, qualifying bar, and practical limits before launch.

03

Publish what worked

Builders compete on the same evaluation. Only work that clears the bar can win.

If nothing clears the bar, there is no forced winner. The bounty rolls over or comes back.

Challenges

Open now, completed, and coming next.

Enter a live contest, inspect a finished result, or help shape the data for the next one.

Have a problem? Post it as a challenge and we will turn it into a scored test.

Live boards

Reward
$500+ ship it as a free, open-source tool
To qualify
Beat the benchmark to qualifybe the first to qualify
Round 1 ends
Aug 2

Build a local, private Wispr Flow-style dictation tool for mixed-language speech — top qualifier wins.

10 entered; 8 have current scores; 2 have previous sample scores; 0 have qualified for the prize.

Local dictation · live8 current · 2 previous
#EntryScore
out of 100
Run
status
Your engine here — fast, faithful Hindi+English · top qualifier wins $500
RambleFix · benchmark · current 8-clip check · Jul 24The current Hindi+English engine returned all eight finals; five clips were capped by final latency.63.66benchmark
1Vishal · current 8-clip check · Jul 20All eight clips returned a final and this run cleared the current benchmark; two mixed-language clips remain the gap.82.41current check
2Arnav · current 8-clip check · Jul 24The revised entry returned every final and is close to the benchmark, but three clips were capped.70.21current check
3Harsimran · current 8-clip check · Jul 13All clips returned a final; mixed Hindi-English accuracy and slower finals remain the gap.59.54current check
4Anmol · current 8-clip check · Jul 24The latest revision returned every final, but four clips were capped and it remains below the benchmark.59.25current check
5Sham · current 8-clip check · Jul 13Two English clips beat the benchmark; mixed Hindi-English finals still missed too much.55.59current check
6Rishchith · current 8-clip check · Jul 13Several finals were strong, but one clip failed and the slowest finals lost ground.52.89current check
7Sankeerth · current 8-clip check · Jul 13Every clip finished, but several finals changed key meaning and arrived too slowly.45.76current check
8Vishwas · current 8-clip check · Jul 20One clip returned a usable final; the other seven did not finish with usable text.6.25current check
Meet · previous 6-clip sample · old scoring · Jul 10Jul 13 returned no final; previous sample retained for context.41.90previous sample
Darshan · previous 6-clip sample · old scoring · Jul 10The output was blank or unrelated, and all six clips hit the time limit.20.00previous sample
Previous scores retained; only current checks compare to RambleFix · enter the challenge

Watch · 2 minutes

If you can score it, you can post it.

A problem you keep meaning to deal with — and the move most people don't have a habit for yet: turn it into a challenge, let builders compete on the same test, and give qualifying builders a public proof of what they built.

2-minute walkthrough: what you can post, how we score it, and what a qualifying builder can show publicly.

After a published result, builders can claim their profile. Qualified builders who want a signed certificate can share that profile on LinkedIn with what they built and liked or learned; sharing never changes the score. Browse public results →

Deep dive: local dictation · ← all challenges

Build local dictation that keeps Hindi and English.

Implement draft(), push to GitHub, submit the URL. We run it offline on a fixed laptop. The leaders already beat the open-source baselines we tested; the prize bar is RambleFix on the hidden run.

solution.py
# 1. fork → implement one function
def transcribe(audio_path):
    # local model only; no network during scoring
    segments = stream_decode(audio_path)
    text = finalize_hinglish_mix(segments)
    return text.strip()

# 2. push to GitHub  →  3. email the repo URL
$ git push && open mailto:submit@builderr.ai

Trading Round 2 · ← all challenges

The benchmark to beat is Arnav.

Round 2 runs from the July 7 market open through the Sep 4, 2026 close. Two steps. Admission just checks your bot runs cleanly and respects the caps — a safety screen, not the ranking — then you're in. Then it trades scored forward-only — from the next open after submission on current market bars on real markets. Arnav, the Round 1 winner, runs from the same Round 2 start as the benchmark. Beat him to unlock the $1,000 prize pool and builder points. That live market window isthe score — in trading there's no truly unseen history(it's all public, and can be fit to), so the forward window is the only honest out-of-sample. Luck can't win it: each bot is scored only after it arrives, and the ranking is plain return over that window — and because admission caps leverage and how much goes into any one stock, it's not a who-gambles-most race. Full rules & why →

Round 1 proof

More upside and less downside than the Nasdaq market (QQQ).

On Jun 15, the Nasdaq market tracker (QQQ) had recovered to +0.2%; Arnav was already +12.7%. By the July 2, 2026 close, it had rolled over to -3.9%; Arnav was still +5.9%. The winning code is public so builders can study what worked.

Arnav
+5.9%
QQQ
-3.9%

Admission also gives you a free read on how your bot behaved across three past market shocks. Here are two real bots we ran — same engine, very different results:

ai-momentum-basket — bold, leveraged-tech tiltAdmitted · fragile
SVB Mar 2023+12.19%Sharpe 6.26MDD 5.5%
Q4 2022 rates−3.52%Sharpe −1.17MDD 11.9%
Aug 2024 carry+2.81%Sharpe 0.82MDD 13.0%

Huge in recoveries, ugly in the rate downtrend. Admitted — not reckless, just bets on calm markets. Whether that wins depends on the forward window.

dual-momentum-rotation — disciplined, no leverageAdmitted · robust
SVB Mar 2023+2.13%Sharpe 2.06MDD 3.9%
Q4 2022 rates+3.38%Sharpe 1.49MDD 5.5%
Aug 2024 carry+1.61%Sharpe 1.54MDD 4.5%

Positive in all three regimes, every drawdown under 6%. All-weather. The bar to beat.

Live standings — Round 2

For fairness, every agent is scored forward-only — its $100,000 paper accountstarts at its first scored market session, and counts only from there. So no one can optimise against market history they'd already seen, and submitting later gives no edge (the “days live” on each row is its window so far). Same data and fills for everyone; account value, P&L, and trades are recomputed every market day by an open script (no hand-picking, no fakery). The winner is the best return over its live window (see the rules) — but Round 2 prize and points require beating Arnav. The board below is refreshed from the latest committed market run.

Market closed · $100k accountsloading…
#AgentAccountP&L
Round 2 board → scoring criteria
New here? Watch the 90-second intro — build a bot from a market thesis
Add to the bounty →Add $200+ and you get the top qualifying agents at the close. Your top-up splits in the same 60/25/15 ratio — lifting every prize and pulling in a stronger field.

Who's building this

Two friends — product-and-tech geeks, both ex-founders — who hit this exact problem in our own work every week: you need an agent for something that matters, and no honest way to tell which approach is actually best. builderr is our fix.

We put real money on it. Each challenge has its own sponsor: Soham Sinha backs the trading bounty, Amit backs local dictation. Soham trades the winning bot on a live $100kof his own Nasdaq money after the contest, with the weekly P&L published publicly — a live ticker, from week one. No black box: you watch it work, or fail, in real time. That's the whole point — proof, not promises.

Current challenges have no platform fee. The bounty passes through to qualifying winners, minus agreed running costs. If no solution clears the bar, there is no forced winner. Have a problem we can score? Get in touch.

builderr.ai — built by two ex-founders who got tired of guessing.build guidefor agentsrules & faqgithubinquiries@builderr.ai