Gemini went rogue, hacked three companies, and Google hid it
(theverge.com)
1.
2.
Gemini Hacked Three Companies in First Known Breakout by Google’s AI
(feeds.content.dowjones.io)
3.
How to end a client relationship on good terms
(feeds.feedburner.com)
4.
5.
OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance
(news.ycombinator.com)
6.
OpenAI details more cases of AI agents taking unauthorized actions
(bleepingcomputer.com)
8.
OpenAI Model Misalignment Report
(news.ycombinator.com)
9.
OpenAI reveals six more safety issues and unveils plan to disclose incidents
(feeds.bbci.co.uk)
11.
An OpenAI Agent Tried to Jailbreak Itself
(wired.com)
12.
13.
Why are AI agents lying, cheating and coordinating?
(news.ycombinator.com)
14.
15.
16.
17.
OpenAI admits to German wiki ‘incident’
(theverge.com)
18.
Painting the sides of railroad rails white to reduce derailment
(news.ycombinator.com)
19.
Alignment pretraining: AI discourse creates self-fulfilling (mis)alignment
(news.ycombinator.com)
20.
21.
Teaching Claude Why
(news.ycombinator.com)
22.
23.
24.
OpenAI can rehabilitate AI models that develop a “bad-boy persona”
(technologyreview.com)
25.
Agentic Misalignment: How LLMs could be insider threats
(news.ycombinator.com)
26.
OpenAI can rehabilitate AI models that develop a “bad boy persona”
(technologyreview.com)