Tech News
← Home  ·  All topics

Rlhf

4 GoKawiil briefs on this topic

Ex-OpenAI Researcher's TypeSafe AI Launches Non-Chat Model 'Jev' With $40M Backing

TypeSafe AI, founded by former OpenAI RLHF researcher Diogo Almeida, has emerged from two years of stealth development with $40 million in funding and its first product, a model called Jev. Rather than generating conversational text like ChatGPT, Jev is built to deliver fast, cheap yes/no or categorical decisions—in about a tenth of a second and for a fraction of a cent—aimed at software-to-software interactions rather than human chat.

TypeSafe AI, founded by ChatGPT co-creator, launches non-text model Jev

Diogo Almeida, a former OpenAI researcher who helped invent RLHF and build ChatGPT, has released a new transformer-based model called Jev through his startup TypeSafe AI. Unlike large language models, Jev doesn't generate text—it produces calibrated probability outputs, making it faster, cheaper and immune to hallucination. Developers have shown strong demand, with the company briefly unable to keep up with API traffic.

New Technical Book Teaches Foundation Model Engineering End-to-End

A new textbook titled Foundation Model Engineering has been released, aimed at AI engineers and research-minded readers who want to understand foundation models beyond basic API use. It ties together topics like attention, mixture-of-experts, RLHF, multimodality, long-context inference, retrieval-augmented generation, and agents into a single engineering narrative, using PyTorch examples, quizzes, and interactive visualizers.

Anthropic's blog prose diverges sharply from Claude's signature writing style

An observer notes that Claude has developed a highly distinctive writing style—short, punchy sentences full of dashes and staccato phrasing—that is now spreading to distilled models like Kimi K3 and even to human writers exposed to heavy Claude output. Yet Anthropic's own public communications, such as its posts on the J-Space or Claude's Constitution, read nothing like this; they resemble conversational, well-edited blog prose instead. Other major models like GPT-5 and Gemma 4 don't share Claude's quirks either, suggesting the voice is a deliberate or emergent Anthropic-specific trait rather than an industry norm.