B

Author

BuilderProof editorial team

Benchmarks

Bolt vs v0 (2026): Benchmarked Across 6 Axes

Bolt and v0 on BuilderProof's documented axes. The six-axis totals and the output-quality cells are withdrawn, and the two split the five surviving axis wins. Here is the per-axis scorecard, with every score cited.

9 min read173
Benchmarks

Lovable vs Bolt vs Replit (2026): 6-Axis Leaderboard

Lovable, Bolt and Replit read against BuilderProof's documented axes in 2026. The six-axis totals and the output-quality cells are withdrawn. The three pairwise results form a near-cycle, not a clean 1-2-3. See who wins each surviving row.

11 min read198
Benchmarks

Replit vs Lovable (2026): Benchmarked Across 6 Axes

On BuilderProof's axis set for 2026, the six-axis totals are withdrawn. Replit wins first-build stability, deployment breadth and default auth posture; Lovable wins code portability. A reproducible, documentation-sourced head-to-head.

10 min read348
Benchmarks

Replit vs Bolt (2026): Benchmarked Across 6 Axes

On BuilderProof's axis set for 2026, the six-axis totals are withdrawn. Replit wins first-build stability and deployment breadth; Bolt wins default-on credential hygiene. A reproducible, documentation-sourced head-to-head.

9 min read175
lab-notes

Deploy quality: why two Lighthouse runs disagree (2026)

Two Lighthouse runs on the same deployed AI-builder output rarely return the same score. This note documents the variance phenomenon, names its seven sources from Google's own documentation, and publishes the median-of-five protocol a hands-on deploy-quality harness would have to meet. Corrected August 21, 2026: the June 2026 result table it accompanied has been withdrawn, and the account of how that table was produced is withdrawn with it.

16 min read280
Methodology

First-build stability: a v2 axis proposal (June 2026)

Proposing first-build stability as the fifth BuilderProof axis: the fraction of OQ-7 prompts that complete without manual intervention. Failure-mode taxonomy, measurement protocol, scoring rubric and open questions, dated June 20, 2026.

10 min read210
Methodology

How We Benchmark AI App Builders: The BuilderProof Methodology v1

BuilderProof methodology v1.1: the published rubric, brief OQ-7, environment standards and weights used to score AI app builders on output quality, speed, deploy quality and agency suitability. The four June 2026 result sets were withdrawn on August 21, 2026 as placeholder data, so the lab currently publishes method, not scores.

11 min read222