Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
Gemini went rogue, hacked three companies, and Google hid it (theverge.com)
2.
Gemini Hacked Three Companies in First Known Breakout by Google’s AI (feeds.content.dowjones.io)
3.
How to end a client relationship on good terms (feeds.feedburner.com)
4.
OpenAI caught its models leaving notes to successors to hide bad behavior (techcrunch.com)
5.
OpenAI's Misalignment Framework: A Tactical Bid to Preempt Global AI Governance (news.ycombinator.com)
6.
OpenAI details more cases of AI agents taking unauthorized actions (bleepingcomputer.com)
7.
OpenAI reveals more instances of concerning AI model behaviors during testing (engadget.com)
8.
OpenAI Model Misalignment Report (news.ycombinator.com)
9.
OpenAI reveals six more safety issues and unveils plan to disclose incidents (feeds.bbci.co.uk)
10.
OpenAI reports 6 new instances of 'concerning model behavior' since March (cnbc.com)
11.
An OpenAI Agent Tried to Jailbreak Itself (wired.com)
12.
OpenAI Creates a New Framework to Disclose Bad AI Behavior (wired.com)
13.
Why are AI agents lying, cheating and coordinating? (news.ycombinator.com)
14.
OpenAI admits to 'wiki incident' after its agents were discovered using a programming hub to communicate — says more transparency is needed regarding misalignments (tomshardware.com)
15.
OpenAI responds after report exposed another incident in which its AI agents went rogue (engadget.com)
16.
OpenAI confirms ‘wiki incident,’ says it’s ‘working on a framework’ for more disclosure (techcrunch.com)
17.
OpenAI admits to German wiki ‘incident’ (theverge.com)
18.
Painting the sides of railroad rails white to reduce derailment (news.ycombinator.com)
19.
Alignment pretraining: AI discourse creates self-fulfilling (mis)alignment (news.ycombinator.com)
20.
Anthropic says ‘evil’ portrayals of AI were responsible for Claude’s blackmail attempts (techcrunch.com)
21.
Teaching Claude Why (news.ycombinator.com)
22.
Continuously graded-doped SnO<sub>2</sub> for efficient n–i–p perovskite solar cells (feeds.nature.com)
23.
Training large language models on narrow tasks can lead to broad misalignment (feeds.nature.com)
24.
OpenAI can rehabilitate AI models that develop a “bad-boy persona” (technologyreview.com)
25.
Agentic Misalignment: How LLMs could be insider threats (news.ycombinator.com)
26.
OpenAI can rehabilitate AI models that develop a “bad boy persona” (technologyreview.com)
Today's top topics: artificial intelligence chatgpt donald trump android authority openai anthropic jensen huang ai regulation ai agents show hn
View all today's topics →