Tech News
← Home  ·  All topics

Goodfire

2 GoKawiil briefs on this topic

Goodfire launches internal 'activation' monitors to flag rogue AI agents via Baseten

Interpretability startup Goodfire has released AI safety monitors that inspect a model's internal signals in real time, rather than reviewing its text output after the fact. The tool is now available to customers of Baseten, which hosts AI models, following a safety partnership announced last month with Baseten's Base Labs and Hugging Face. Users can choose which risks to flag, such as hacking attempts or weapons-related misuse, and select automated responses ranging from logging to blocking requests.

Baseten's Base Labs teams with Hugging Face and Goodfire on open-model safety standard

Baseten's newly formed research arm, Base Labs, announced a partnership with Hugging Face and Goodfire AI on Wednesday to develop safety evaluation and monitoring tools for open-weight AI models. The effort aims to create a shared standard that embeds safety directly into how models are trained and deployed, rather than adding it after release. Technical details of how the collaboration will function have not yet been disclosed.