2026-10-11 16:38 UTC

Base Labs, Hugging Face, and Goodfire claim their announced partnership will produce transparent safety evaluation and monitoring methods integrated into open-model training and serving, potentially establishing reusable deployment controls without implying restrictions on abliterated model hosting.

state: watchingheat: highuncertainty: highnovelscott: lowopen-model-safety model-hostingBase LabsBasetenHugging FaceGoodfire
Surfaced 2026-09-19T10:24:11Z — Base Labs says openness provides greater visibility into model behavior and greater means of turning safety research into actionable, transp — The loud spread reading raises attention urgency, but the discussion is amplifying fears of hosting restrictions rather than supplying evidence of delivered safety controls or a policy change. The announcement warrants watching; its reporting and reconstructed company statement are not independent validation of the proposed methods.

What is this?

The case describes a claimed partnership between Base Labs, Hugging Face, and Goodfire to develop transparent safety evaluation and monitoring integrated into open-model training and serving, but none of the supplied search snippets verifies that announcement. Goodfire’s model cards on Hugging Face do establish that it publishes sparse autoencoders for inspecting Llama models’ internal representations and supporting interpretability and model steering. The supplied material does not establish Base Labs’ identity or role, any relationship to Baseten, or delivered partnership controls; it also provides no evidence of a change to Hugging Face’s hosting policy for abliterated models.

Why it matters to Scott

The claimed partnership has only broad overlap with Scott’s 12-Factor Agents emphasis on managed evaluation and observability; the supplied material establishes neither reusable controls affecting his deployments nor adoption of his stronger containment-and-scoped-authority position. No radar hit tracks this partnership, and the unverified announcement provides no demonstrated hosting-policy change or actionable integration that would change what Scott builds or argues.
queries asked of Scott's wikis
  • open-weight sovereignty hosting restrictions platform dependence
  • transparent safety evaluations reproducible deployment controls
  • runtime model monitoring agent harness safety enforcement
  • mechanistic interpretability sparse autoencoders model steering
  • local model deployment refusal removal safety tradeoffs

Measured heat

now 0 pts/hpeak 0 pts/hcomments 0/hpeers p0momentum: steady2 platformsage 549h
points/hour across evidence · reading as of 2026-10-12 02:59:37.977291+11:00 · deterministic, not a model opinion

How the heat travelled

09-18 19:23 (minted)⭐ origin echo-reconstructedBase Labs says openness provides greater visibility into model behavior and greater means of turning safety research into actionable, transp
Base Labs on x (echo) · attributed from reddit.post.1wjyn95 · published time unknown
—
09-18 18:43first on r/LocalLLaMA · published · lag ?Is HF starting to move against abliterated models?
returnity
—
09-18 18:43amplified on r/LocalLLaMA 👑reddit.post.1wjyn95
returnity
peak 458 · 214 comments · 100% of case engagement
09-18 19:20our radar first saw it · lag ?discovery anchor: reddit.post.1wjyn95—
09-19 10:24reached heat=high · lag ? · via ledger——
pace: p86 vs 1032 stories at the 336h mark (now 549h old) — ahead of desert-ant-on-device-models (1.0x), behind codex-bundled-libreoffice-runtime (1.0x)

Evidence (2) — ⭐ canonical anchor

sourceobjectauthorscorecomments
🟠 redditIs HF starting to move against abliterated models?
LocalLLaMA
Retrieved article excerpt

Open article · Retrieved 2026-09-18T19:22:27.599565+00:00

[Baseten](https://www.baseten.co) launched a new safety infrastructure standard alongside its Base Labs research arm on Wednesday, partnering with Hugging Face and Goodfire AI to build safety evaluation and monitoring infrastructure for open-weight models.

The announcement lands amid debate for the safety of open-weight models — which can be made dangerous by removing their safeguards through a rising technique known as [abliteration](https://techcrunch.com/2026/09/03/abliteration-ai-is-making-a-business-out-of-removing-ai-guardrails/). The scale of the problem is massive: Hugging Face, which hosts open source AI models, currently lists over 6,000 abliterated models.

Base Labs, the research group Baseten spun up earlier this year, will develop and publish methods for training and monitoring open models. The company is framing their future work as a “standard” for open models that is transparent and built into how models are trained and deployed, rather than bolted on afterward.

“We believe openness to be an advantage for AI safety,” the company said on [X](https://x.com/baselabs/status/2100286099121705396). “Openness provides more visibility into the behavior of models and, most importantly, greater means of turning safety research into actionable and transparent controls than closed-source.”

The companies haven’t disclosed how the partnership will work technically, though Goodfire framed the goal in a reply to Baseten’s post: “Safety must be built into open models and provided by those who serve them.” Goodfire, which specializes in opening AI’s “black box” to explain how models make decisions, is the likeliest candidate for the “built into” part.

Baseten, an AI inference provider, [raised](https://techcrunch.com/2026/06/18/ai-inference-startup-baseten-reportedly-raising-1-5b-months-after-its-last-mega-round/) a $1.5 billion Series F in June, vaulting its valuation to $13 billion. Partner Goodfire AI is similarly well-capitalized, having raised a $150 million Series B led by B Capital earlier this year to advance its model interpretability platform.

Looking ahead, Baseten is putting out an open call to the broader developer ecosystem to contribute to the framework. “Together, we are building an ecosystem of open models that are safe and accessible to all,” the company noted.

Topics

[AI](https://techcrunch.com/category/artificial-intelligence/), [Hugging Face](https://techcrunch.com/tag/hugging-face/), [responsible ai](https://techcrunch.com/tag/responsible-ai/), [TC](https://techcrunch.com/category/tc/)

*When you purchase through links in our articles, [we may earn a small commission](https://techcrunch.com/techcrunch-affiliate-monetization-standards/). This doesn’t affect our editorial independence.*

Aditya Mehta

[View Bio](https://techcrunch.com/author/aditya-mehta/)

Event Logo

October 13 – 15

San Francisco

Last day to book an exhibit table is September 18. Don’t miss out on high-impact leads, investor access, and a brand spotlight in Disrupt’s Expo Hall.

[**BOOK NOW**](https://techcrunch.com/events/techcrunch-disrupt/exhibit/?promo=rightrail_exhibit&utm_campaign=disrupt2026&utm_content=exhibit&utm_medium=ad&utm_source=tc)

## Most Popular

- ### [OpenAI caught its models leaving notes to successors to hide bad behavior](https://techcrunch.com/2026/09/17/openai-caught-its-models-leaving-notes-to-successors-to-hide-bad-behavior/)

  - [Rebecca Bellan](https://techcrunch.com/author/rebecca-bellan/)
- ### [Clean tech startup Fluxnium found a way to tap 50,000 years’ worth of nuclear fuel](https://techcrunch.com/2026/09/16/clean-tech-startup-fluxnium-found-a-way-to-tap-50000-years-worth-of-nuclear-fuel/)

  - [Tim De Chant](https://techcrunch.com/author/tim-de-chant/)
- ### [Salesforce and Nvidia’s new reasoning model is everything the AI labs should fear](https://techcrunch.com/2026/09/15/salesforce-and-nvidias-new-reasoning-model-is-everything-the-ai-labs-should-fear/)

  - [Julie Bort](https://techcrunch.com/author/julie-bort/)
- ### [Jensen Huang took a call from Trump, and showed off something else, too](https://techcrunch.com/2026/09/14/jensen-huang-took-a-call-from-trump-and-showed-off-something-else-too/)

  - [Connie Loizos](https://techcrunch.com/author/connie-loizos/)
- ### [The 9 buzziest startups from Y Combinator’s latest Demo Day, according to VCs](https://techcrunch.com/2026/09/13/the-9-buzziest-startups-from-y-combinators-latest-demo-day-according-to-vcs/)

  - [Marina Temkin](https://techcrunch.com/author/marina-temkin/)
  - [Dominic-Madori Davis](https://techcrunch.com/author/dominic-madori-davis/)
- ### [Tesla says it will finally unveil the second-generation Roadster on October 1](https://techcrunch.com/2026/09/12/tesla-says-it-will-finally-unveil-the-second-generation-roadster-on-october-1/)

  - [Anthony Ha](https://techcrunch.com/author/anthony-ha/)
- ### [Revolut confirms customer data breach through fake government requests](https://techcrunch.com/2026/09/12/revolut-confirms-customer-data-breach-through-fake-government-requests/)

  - [Jagmeet Singh](https://techcrunch.com/author/jagmeet-singh/)
returnity458214
🟧 echo.x ⭐Base Labs says openness provides greater visibility into model behavior and greater means of turning safety research into actionable, transpBase Labs——

Interpretation history

Decision trace