Skip to content
Tech News
clear
Topics: Today This Week This Month This Year
1.
How and when to use artificial intelligence in your science job application (feeds.nature.com)
2.
The AI safety test is becoming a safety risk (techcrunch.com)
3.
Third-party cyber evaluations involving OpenAI models (news.ycombinator.com)
4.
Stop graphing everything: When GraphRAG actually beats vector RAG (venturebeat.com)
5.
Structured AI data pipelines score 10.9 points below free-form code — DataFlow-Harness closes the gap (venturebeat.com)
6.
Not just OpenAI - Anthropic says Claude's hacking spree 'falls short of ideal behavior' (zdnet.com)
7.
Investigating three real-world incidents in our cybersecurity evaluations (news.ycombinator.com)
8.
Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it (news.ycombinator.com)
9.
At Waymo, an AI project isn't ready until its evals are — not when the model performs well (venturebeat.com)
10.
vBulletin fixes critical pre-auth RCE flaw with public exploit (bleepingcomputer.com)
11.
Angels in Coptic Magic I: Introduction (news.ycombinator.com)
12.
Claude Cookbook (news.ycombinator.com)
13.
From Evaluation to Guardrails: What We Brought to ACM FAccT 2026 (news.ycombinator.com)
14.
AI agents aren't confidently wrong because of bad context — they're wrong because of bad data engineering (venturebeat.com)
15.
AI Can Generate Pictures, but It Can Also Help Locate Your Lost Real Photos (cnet.com)
16.
OpenAI Says Its Unreleased Model Broke Containment and Went Rogue (gizmodo.com)
17.
Hugging Face Said Last Week It Was Attacked. An Unreleased OpenAI Model Did It, OpenAI Now Says (gizmodo.com)
18.
OpenAI and Hugging Face address security incident during model evaluation (news.ycombinator.com)
19.
Evals are the new PRD, Expedia’s AI chief tells VB Transform 2026 (venturebeat.com)
20.
Evidence of inconsistencies in evaluation process and selection of winners (news.ycombinator.com)
21.
The AI context gap: Enterprise AI organizations have a trust problem, not a retrieval problem — and most are still building the fix (venturebeat.com)
22.
The agent evaluation gap: Enterprise AI organizations have a reality-alignment problem, not a coverage problem — and most are shipping to production anyway (venturebeat.com)
23.
Murati's Thinking Machines Releases Open-Weights 975B Parameter LLM (news.ycombinator.com)
24.
GPC3-specific dnTGFβRII-armoured CAR T cells for hepatocellular carcinoma (feeds.nature.com)
25.
Why we're moving off Cloudflare Durable Objects (news.ycombinator.com)
26.
Separating signal from noise in coding evaluations (news.ycombinator.com)
27.
Hazel (YC W24) Is Hiring for Our Largest Government Contract (news.ycombinator.com)
28.
Medieval-style fortifications are back in the Sahel (news.ycombinator.com)
29.
A startup taught humanoid robots to retrieve packages, climb stairs, and unpack boxes – no human steering needed (techspot.com)
30.
Memora: A Harmonic Memory Representation Balancing Abstraction and Specificity (news.ycombinator.com)
Today's top topics: openai google anthropic apple samsung spotify claude amazon artificial intelligence android authority
View all today's topics →