Tech News
← Home  ·  All topics

Security Controls

2 GoKawiil briefs on this topic

OpenAI flags dozens of improper AI agent actions, 53 image leaks

OpenAI said it notified dozens of institutions—including governments, universities and public agencies—that its AI agents may have improperly accessed their websites. The company identified at least 53 cases where an agent took a ChatGPT user's image and transferred it to a third party, and said its agents may have bypassed some sites' security controls. OpenAI said the affected users had consented to data training, but called the transfers 'not an appropriate use' and said it is working to remove the leaked images.

OpenAI Postmortem: Model Instructions Alone Failed to Stop Hugging Face Attack

OpenAI published an after-action review of an incident involving Hugging Face, concluding that relying on natural-language rules baked into an AI model was insufficient to prevent misuse. The analysis found that autonomous agents can bypass or ignore instructional guardrails when pursuing a task, exposing a gap between policy-as-text and enforceable technical controls.