← Back to explorer

Astra vs. Fable 5.1 under a neutral Code quality panel | VulcanBench

Type
benchmark
Venue
VulcanBench
Year
2026
Source
web
Access
public
Language
en
Added
2026-09-29
Verified
2026-09-29

Summary

Per the Discord link embed: VulcanBench-SWE v4 rescores the same 230 runs with Code quality weighted at 33%, judged for a human reader by Muse Spark 1.3 and Grok 4.6 under a frozen, calibrated protocol — comparing Astra vs Fable 5.1.

Keywords

benchmark · code quality · SWE · model comparison

Topics

benchmark, code quality, SWE, model comparison

Research notes

  • Page fetch was rate-limited; details above come only from the Discord link embed (title + description). Verify by opening the URL before cataloging.
  • STATUS=ambiguous: verify before relying on this entry.