Back
LLMsNEW

Base Labs, Hugging Face and Goodfire Are Building a Safety Standard for Open-Weight Models

Base Labs, the safety research arm spun out of Baseten, partnered with Hugging Face and Goodfire on September 17, 2026 to build safety evaluation and monitoring infrastructure for open-weight models. The announcement is a stated intention backed by real funding, not a shipped tool — there is no technical detail yet on what the infrastructure actually does.

3 min read
Updated Sep 18, 2026
QUICK ANSWER

Base Labs, the safety research arm spun out of Baseten, partnered with Hugging Face and Goodfire on September 17, 2026 to build safety evaluation and monitoring infrastructure for open-weight models.

Key Takeaways
  • This guide provides comprehensive, actionable information
  • Consider your specific workflow needs when evaluating options
  • Explore our curated LLMs tools for specific recommendations

What actually shipped

On September 17, 2026, Base Labs — the AI safety research arm spun out of Baseten earlier this year — announced a partnership with Hugging Face and Goodfire to build safety evaluation and monitoring infrastructure for open-weight models, reported first by TechCrunch. Goodfire is contributing its model-interpretability work; the stated goal is to develop and publish methods for training and monitoring open models so safety is built in during development rather than patched on afterward. Base Labs put it directly: "We believe openness to be an advantage for AI safety. Openness provides more visibility into the behavior of models."

6,000+
abliterated (guardrail-removed) models currently hosted on Hugging Face
$1.5B
Baseten's June 2026 Series F, at a $13B valuation
$150M
Goodfire's 2026 Series B, led by B Capital

The number that explains why this exists

Hugging Face's own count, cited in the announcement, is that the platform currently hosts more than 6,000 abliterated models — models with their refusal behavior surgically removed. That is the scale problem this partnership is aimed at: not a hypothetical future risk, but a running total of already-published finetunes with guardrails stripped out, on the largest open-weight hosting platform there is. For what abliteration actually does to a model and what it costs in accuracy, see uncensored AI models: what abliteration actually does — this partnership is effectively an industry response to the exact gap that guide measures.

What's real here, and what isn't yet

The funding and the parties are real and verifiable: Baseten closed a $1.5B Series F in June 2026 at a $13B valuation, and Goodfire raised a $150M Series B earlier in 2026. Goodfire's own framing is a direct claim of responsibility: "Safety must be built into open models and provided by those who serve them." What's not yet real is any shipped artifact. The reporting is explicit that no specific technical implementation has been disclosed, and Baseten has only issued an open call for developer-ecosystem contributions. This is a funded, credible group of three companies stating an intention — worth tracking, not yet something to evaluate against a spec, because there isn't one.

Who should actually care right now

You publish or fine-tune open-weight models
Nothing to adopt yet — there's no released tooling or standard. Worth watching for what Base Labs actually publishes, since it's explicitly meant to be used at training time.
You're evaluating an abliterated or otherwise "uncensored" finetune today
This partnership doesn't change your risk today. Keep testing math and truthfulness yourself before trusting a downloaded checkpoint — see the abliteration guide above for what to check.
You're deciding whether "open weight" and "safe" are compatible for a product decision
Treat this as a data point that credible, funded labs are actively arguing yes — not as proof the problem is solved.

For the open-source model landscape this sits inside, see open-source AI tools ranked and every LLM in the directory.

FREQUENTLY ASKED QUESTIONS
What did Base Labs, Hugging Face and Goodfire actually announce, and does it change anything about open-weight model safety today?
Base Labs, the safety research arm spun out of Baseten, partnered with Hugging Face and Goodfire on September 17, 2026 to build safety evaluation and monitoring infrastructure for open-weight models.
EXPLORE TOOLS

Ready to try AI tools? Explore our curated directory:

SHARE THIS GUIDE

Base Labs, the safety research arm spun out of Baseten, partnered with Hugging Face and Goodfire on September 17, 2026 to build safety evaluation and monitoring infrastructure for open-weight models.

Share on X LinkedIn Reddit Email
Copied to clipboard