2026-10-11 17:20 UTC

Independent evaluation will determine whether FrontisAI’s open 35B Frontis-MA1 demonstrates reproducible recursive self-improvement beyond ordinary fine-tuning or benchmark optimization.

state: expiredheat: lowuncertainty: highconvergesscott: mediumopen-models recursive-self-improvement frontier-capabilitiesFrontisAI

What is this?

FrontisAI presents Frontis-MA1 as an open 35B-parameter model post-trained on OpenMLE, a stack combining executable machine-learning tasks, reinforcement learning, and long-horizon evolutionary search. Its repository also lists model derivatives, datasets, and evaluation infrastructure intended to make AI-for-AI improvement measurable and reproducible. The supplied snippets report a 71% average on MLE-Bench Lite, but provide no clearly independent evaluation establishing recursive self-improvement beyond post-training, search, or benchmark optimization; the web summary’s claim of independent confirmation is unsupported by the listed results.

Why it matters to Scott

FrontisAI’s claimed evaluated, substrate-changing improvement cycles align with Scott’s definition of Self-Improving Loops. Independent evaluation would directly test his load-bearing distinction between genuine learning across cycles and agentic search or benchmark gaming, although the supplied evidence does not yet validate the claim.
ip:concept.self-improving-loopsip:concept.search-not-learningip:concept.evaluation-driven-developmentip:concept.specification-gamingradar:concept.open-modelsradar:concept.model-evaluationradar:concept.benchmark-integrity
queries asked of Scott's wikis
  • recursive self-improvement versus agentic search
  • AI agents automating machine-learning engineering
  • reproducible evaluation of self-improving systems
  • open-weight models for AI research automation
  • benchmark optimization versus capability improvement
  • coding-agent harnesses with execution feedback

Measured heat

no measured readings yet — the hourly heat pass fills this in

How the heat travelled

no chain yet — the hourly chain pass fills this in

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟧 hnFrontis-MA1: Open 35B model toward recursive self-improvementDSemba10
🟧 echo.blog ⭐Introduces Frontis-MA1 as an open 35B model aimed toward recursive self-improvement.FrontisAI——

Interpretation history

Decision trace