What is a Fuse Score?
A 0–1000 ranking assigned after reviewing each test run — code quality, correctness, polish, and how well the output matches the prompt. It is not an automated metric; it reflects real harness evaluation.
Join Fuse Intelligence — manage subscriptions & downloads
Already registered?
By registering you agree to our Terms and Privacy Policy.
Signing in…
Fuse Intelligence · Community
Real capability tests on models and providers — not vibes. Each run stores the HTML, CSS, and JavaScript the model produced, plus a hand-assigned Fuse Score from 0–1000. Models with multiple tests show an average score.
Sorted by average Fuse Score across all recorded showcase tests.
composer-2.5-fast
Agentic coding model benchmarked on the same widget-native card-game prompt.
A 0–1000 ranking assigned after reviewing each test run — code quality, correctness, polish, and how well the output matches the prompt. It is not an automated metric; it reflects real harness evaluation.
Every benchmark keeps the model’s HTML, CSS, and JavaScript so you can inspect what it actually built — not just a number on a leaderboard.
Card-game UI is the first harness. Widget layout, agent tool-use, and marketplace widget generation tests will appear here as they are run.