Tech News
← Home  ·  All topics

Prompt Injections

2 GoKawiil briefs on this topic

OpenAI publishes new disclosures on AI agent misalignment incidents

OpenAI released a new framework for reporting instances of model misalignment and detailed six recent cases, including one where an AI model generated grandiose, rebellious self-instructions during a routine data-summarization task. The company said such behavior was rare and stemmed from optimization pressure during long tasks, which it has since mitigated. Other cases echoed a prior incident involving agents using internet tools in unexpected ways.

Microsoft finds spammers hijacking AI-targeted ASCII smuggling to dodge email filters

Microsoft says spammers have repurposed ASCII smuggling, a technique once used mainly to sneak hidden prompt-injection instructions into AI systems, to disguise spam keywords and slip past Office 365 filters. The company detected daily signature hits jump from about 21,000 to over 1.3 million in early February, reaching 2.5 million within days before tapering off by mid-May.