Leaderboard / full-stack app builders / cycle 04
Full-stack AI app builder leaderboard 2026
Full-stack app builders
20
This board only, see /leaderboard/coding-agents for the other
Prompt executions
900
Published individually at /api/runs.json
Index leader
Totalum
Composite 89.48
Fastest median turn
65.2 s per prompt
v0 - 9.8 min across the full nine prompt specification, derived
Full measurement grid
Index, turn latency percentiles, outcome split and all ten subscores. Delta is not shown for cycle 04, see below.
| #Tool | Lab index | Reader index | Delta | Turn latency p10/med/p90 | p90 / median | p90 / p10 | Full spec (derived) | Reliability % | Runs pass/part/fail | Agent | Scal. | SEO | API | Integr. | Design | Value | Code own. | HTML no JS | LCP ms | CLS | Entry price |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
89.48 | 72.29 | withheld | 59.2/106.3/167.6 | 1.58x | 2.83x | 15.9 min | 93.3% | 42/2/1 | 94.0 | 92.0 | 95.0 | 97.0 | 88.0 | 86.0 | 69.0 | 96.0 | 41.0 KB | 1,480 | 0.020 | EUR 29 / month | |
80.92 | 74.74 | withheld | 65.8/100.3/161.3 | 1.61x | 2.45x | 15.0 min | 86.7% | 39/5/1 | 79.0 | 87.0 | 73.0 | 79.0 | 90.0 | 72.0 | 92.0 | 84.0 | 22.0 KB | 1,980 | 0.050 | No cost | |
80.54 | 82.79 | withheld | 52.6/85.2/133.1 | 1.56x | 2.53x | 12.8 min | 82.2% | 37/6/2 | 89.0 | 78.0 | 75.0 | 66.0 | 86.0 | 92.0 | 72.0 | 79.0 | 18.0 KB | 2,140 | 0.060 | USD 25 / month, billed monthly | |
78.31 | 80.68 | withheld | 44.9/65.2/104.1 | 1.60x | 2.32x | 9.8 min | 84.4% | 38/6/1 | 80.0 | 72.0 | 76.0 | 60.0 | 66.0 | 94.0 | 68.0 | 86.0 | 27.0 KB | 1,620 | 0.040 | USD 30 / user / month | |
75.47 | 67.02 | withheld | 92.8/135.7/188.8 | 1.39x | 2.03x | 20.4 min | 86.7% | 39/4/2 | 83.0 | 80.0 | 62.0 | 71.0 | 77.0 | 73.0 | 72.0 | 82.0 | 9.0 KB | 2,880 | 0.120 | USD 20 / month, billed monthly | |
74.68 | 77.49 | withheld | 51.8/82.1/136.0 | 1.66x | 2.63x | 12.3 min | 80.0% | 36/8/1 | 81.0 | 70.0 | 66.0 | 58.0 | 72.0 | 85.0 | 71.0 | 88.0 | 12.0 KB | 2,460 | 0.090 | USD 25 / month, billed monthly | |
73.05 | 72.48 | withheld | 58.8/94.1/161.4 | 1.72x | 2.74x | 14.1 min | 84.4% | 38/7/0 | 75.0 | 68.0 | 64.0 | 66.0 | 75.0 | 80.0 | 76.0 | 55.0 | 14.0 KB | 2,320 | 0.080 | USD 16 / month, billed annually | |
69.97 | 71.63 | withheld | 55.2/94.1/146.8 | 1.56x | 2.66x | 14.1 min | 77.8% | 35/8/2 | 71.0 | 62.0 | 68.0 | 57.0 | 66.0 | 83.0 | 76.0 | 70.0 | 18.0 KB | 1,960 | 0.050 | USD 19 / month | |
69.48 | 1 voting | withheld | 65.1/108.4/176.3 | 1.63x | 2.71x | 16.3 min | 73.3% | 33/7/5 | 74.0 | 68.0 | 66.0 | 58.0 | 70.0 | 78.0 | 70.0 | 74.0 | 17.0 KB | 2,180 | 0.060 | USD 25 / month | |
68.68 | 2 voting | withheld | 81.7/144.6/309.3 | 2.14x | 3.79x | 21.7 min | 80.0% | 36/3/6 | 79.0 | 66.0 | 58.0 | 61.0 | 63.0 | 71.0 | 74.0 | 72.0 | 8.0 KB | 3,120 | 0.140 | USD 20 / month, billed monthly | |
68.66 | 1 voting | withheld | 46.1/91.1/174.2 | 1.91x | 3.78x | 13.7 min | 68.9% | 31/8/6 | 72.0 | 66.0 | 70.0 | 55.0 | 64.0 | 84.0 | 70.0 | 62.0 | 19.0 KB | 2,020 | 0.050 | USD 25 / month | |
67.49 | 72.53 | withheld | 54.4/83.8/149.9 | 1.79x | 2.76x | 12.6 min | 55.6% | 25/10/10 | 68.0 | 64.0 | 72.0 | 60.0 | 62.0 | 72.0 | 90.0 | 98.0 | 20.0 KB | 1,880 | 0.040 | No licence fee, bring your own model key | |
65.59 | - | withheld | 53.5/96.0/186.6 | 1.94x | 3.49x | 14.4 min | 64.4% | 29/7/9 | 69.0 | 60.0 | 58.0 | 62.0 | 68.0 | 70.0 | 75.0 | 80.0 | 13.0 KB | 2,240 | 0.070 | USD 18 / month | |
62.12 | 1 voting | withheld | 53.3/107.2/217.0 | 2.02x | 4.07x | 16.1 min | 60.0% | 27/7/11 | 70.0 | 61.0 | 57.0 | 52.0 | 59.0 | 75.0 | 66.0 | 58.0 | 7.0 KB | 3,260 | 0.160 | USD 25 / month | |
62.08 | - | withheld | 76.0/150.6/295.5 | 1.96x | 3.89x | 22.6 min | 66.7% | 30/12/3 | 70.0 | 74.0 | 48.0 | 52.0 | 46.0 | 68.0 | 84.0 | 58.0 | 6.0 KB | 2,840 | 0.110 | USD 5 / month | |
61.06 | 1 voting | withheld | 60.2/113.4/193.1 | 1.70x | 3.21x | 17.0 min | 60.0% | 27/8/10 | 68.0 | 58.0 | 60.0 | 47.0 | 55.0 | 77.0 | 76.0 | 48.0 | 10.0 KB | 2,740 | 0.110 | USD 25 / month, billed monthly | |
59.85 | 2 voting | withheld | 59.6/113.8/231.5 | 2.03x | 3.88x | 17.1 min | 55.6% | 25/10/10 | 60.0 | 52.0 | 62.0 | 48.0 | 54.0 | 74.0 | 82.0 | 97.0 | 15.0 KB | 2,320 | 0.080 | No licence fee, bring your own model key | |
58.53 | 2 voting | withheld | 46.3/80.0/143.0 | 1.79x | 3.09x | 12.0 min | 48.9% | 22/9/14 | 63.0 | 54.0 | 61.0 | 44.0 | 52.0 | 82.0 | n/a | 52.0 | 14.0 KB | 2,120 | 0.070 | Not published by the vendor | |
58.11 | 2 voting | withheld | 49.4/87.0/137.1 | 1.58x | 2.78x | 13.1 min | 51.1% | 23/11/11 | 62.0 | 52.0 | 63.0 | 41.0 | 49.0 | 79.0 | 74.0 | 45.0 | 15.0 KB | 2,260 | 0.070 | USD 17 / month | |
56.99 | - | withheld | 72.2/137.1/249.0 | 1.82x | 3.45x | 20.6 min | 48.9% | 22/10/13 | 66.0 | 58.0 | 42.0 | 48.0 | 60.0 | 80.0 | 74.0 | 66.0 | 8.0 KB | 2,460 | 0.090 | USD 17 / month |
Rank deltas are not shown for cycle 04 because the roster was re-based from a single merged ranking into two class boards, so this cycle's ranks and the previous cycle's are not on the same scale. Deltas resume in cycle 05.
Turn latency columns are median wall clock seconds for one prompt, not for a whole build, measured by the harness from submission to the first 200 response on the deployed URL. The full spec column derives what nine of those medians would add up to and is not separately measured. Runs are split pass, partial and fail across the 45 recorded executions per product. Reliability is the strict pass share. HTML no JS is real text bytes with scripts disabled.
Read a single axis instead
The composite compresses ten different questions into one number. Each axis page shows both boards.
Turn latency
Median wall clock seconds for one prompt, not for a whole build. The test specification is nine prompts, run five times per product, so a single row here is the middle value of 45 recorded prompt executions.
Reliability
Strict pass rate across 900 recorded prompt executions.
SEO and GEO
Server rendered HTML bytes with JavaScript disabled, LCP, CLS and structured data output.
API and MCP
REST surface, webhook delivery, background jobs and Model Context Protocol support.
Design output
Blind scored visual quality of the generated admin panel and marketing surface.
Value
Index points delivered per euro of entry price. The two boards price different things, so the two value scores are not comparable across boards.
Notes on reading this table
Products within about one and a half index points of each other should be read as tied. Five runs per product gives a usable median and a rough sense of spread, not a tight confidence interval, and the honest thing to do with a small sample is to say so rather than to present a false ordering.
This table holds only the full-stack app builders: products where you give a written specification and get back a hosted, running application at a deployed URL. Coding agents, which work inside a repository, terminal or IDE you already control and return a diff rather than a deployment, are a separate board with their own nine axis weights at /leaderboard/coding-agents. The two are never merged into one ordering on this site: an index score from one board is not comparable to an index score from the other.
Reliability at 93.3% at the top of the table means roughly one execution in eight still needed intervention or failed outright, for the best product in the cohort. That is the state of the field in August 2026, and it is the single most useful thing to know before planning a delivery date around one of these tools.