Previewing Locus
- Type
- blog
- Venue
- Intology blog
- Year
- 2026
- Source
- web
- Access
- public
- Language
- en
- Added
- 2026-09-29
- Verified
- 2026-09-29
Summary
Locus is Intology's automated AI research/post-training system. Per the Discord embed, it sustains improvement over days, exceeds human experts on RE-Bench at equal time and compute, and sets SOTA on KernelBench and MLE-Bench Lite. Related Intology material describes Locus as SOTA on PostTrainBench (agents post-training models given 10 H100 hours) and PostTrainBench+, with Locus post-trained Qwen3 models surpassing human post-trained Qwen3 and running in production.
Keywords
AI research agents · post-training · RE-Bench · KernelBench · MLE-Bench
Topics
AI research agents, post-training, RE-Bench, KernelBench, MLE-Bench
Research notes
- Method: Automated post-training agent that runs for days; evaluated on PostTrainBench/PostTrainBench+ (post-training Qwen3 1.7B-Base with thousands of H100 hours).
- Key findings: Exceeds human experts on RE-Bench at equal time and compute (per Intology); SOTA on KernelBench and MLE-Bench Lite (per Intology); SOTA on PostTrainBench and PostTrainBench+; Locus post-trained Qwen3 beats human post-trained Qwen3
- Limitations: All claims are from Intology's own materials; independent verification not checked.
- Blog page fetch was rate-limited; summary combines the Discord embed with Intology's announcement summary found via web search. Confirm details against the blog post before final cataloging.