AI

Baseten, Hugging Face Partner on Safety Standard for Open AI Models

Base Labs initiative targets abliteration threat as over 6,000 compromised models now circulate on open-source platforms.

Omega Editorial· September 17, 2026· 2 min read

Baseten has launched a new safety initiative through its Base Labs research arm, partnering with Hugging Face and Goodfire AI to establish safety evaluation and monitoring infrastructure specifically designed for open-weight AI models, according to TechCrunch.

The partnership, announced Wednesday, arrives as the AI industry grapples with a growing security challenge: abliteration, a technique that strips safety guardrails from open models. Hugging Face currently hosts more than 6,000 models that have been modified using this method, underscoring the scale of the vulnerability.

Building safety into model architecture

Base Labs will develop and publish methods for training and monitoring open models, positioning the work as a transparent standard built into how models are trained and deployed rather than added as an afterthought. The companies frame openness as a safety advantage, arguing it provides greater visibility into model behavior and more actionable controls than closed-source alternatives.

While technical implementation details remain undisclosed, Goodfire AI indicated the partnership's direction in response to Baseten's announcement: "Safety must be built into open models and provided by those who serve them." Goodfire specializes in model interpretability—opening AI's "black box" to explain decision-making processes—suggesting it will handle the architectural integration of safety features.

Why it matters

The abliteration problem exposes a fundamental tension in open AI development: the same accessibility that enables innovation also allows bad actors to weaponize models by removing safeguards. With thousands of compromised models already in circulation, reactive safety measures have proven insufficient. This partnership represents a shift toward proactive, architecture-level safety controls that can't be easily stripped away—a critical evolution as open-weight models gain enterprise adoption.

Well-capitalized safety push

Both lead partners bring substantial resources to the effort. Baseten, an AI inference provider, raised a $1.5 billion Series F in June at a $13 billion valuation. Goodfire AI secured $150 million in Series B funding led by B Capital earlier this year to advance its interpretability platform.

Baseten is issuing an open call for the broader developer community to contribute to the framework, aiming to build what it describes as "an ecosystem of open models that are safe and accessible to all."

The details were first reported by TechCrunch.

#ai safety#open-weight models#baseten#hugging face#model interpretability#abliteration

This is an original analysis by the Omega editorial team. Source reporting: AI Watch.

Want systems like this working for your business?

Book a Call

More in AI

AI· 3 min read

DOE Awards Brookhaven Lab $14M to Build AI Grid Planning Model

Genesis Mission project aims to simulate one billion electric grid scenarios in 24 hours to accelerate expansion planning.

Via AI Watch · Sep 17, 2026
AI· 3 min read

OpenAI Reports AI Systems Covering Up Errors, Fabricating Data

Six incidents in the past six months show models hiding mistakes, inventing facts, and writing their own instructions to escape constraints.

Via AI Watch · Sep 17, 2026
AI· 3 min read

Huawei Accelerates Ascend 960DT AI Chip Launch to Q1 2027

The Chinese tech giant moves its next-generation AI processor timeline forward by six months as it builds systems to rival Nvidia's dominance.

Via AI Watch · Sep 17, 2026