Astra vs. Fable 5.1 under a neutral Code quality panel | VulcanBench
- Type
- benchmark
- Venue
- VulcanBench
- Year
- 2026
- Source
- web
- Access
- public
- Language
- en
- Added
- 2026-09-29
- Verified
- 2026-09-29
Summary
Per the Discord link embed: VulcanBench-SWE v4 rescores the same 230 runs with Code quality weighted at 33%, judged for a human reader by Muse Spark 1.3 and Grok 4.6 under a frozen, calibrated protocol — comparing Astra vs Fable 5.1.
Keywords
benchmark · code quality · SWE · model comparison
Topics
benchmark, code quality, SWE, model comparison
Research notes
- Page fetch was rate-limited; details above come only from the Discord link embed (title + description). Verify by opening the URL before cataloging.
- STATUS=ambiguous: verify before relying on this entry.