Tech News
← Home  ·  All topics

Gpt

104 GoKawiil briefs on this topic

Pangram launches AI content detector using classifier neural network

Pangram has built a text and image classifier that determines whether content was authored by a human or generated by AI. The system tokenizes input text, converts tokens into vector embeddings, and passes them through a neural network with a classifier head that outputs a human, AI, or AI-assisted label. The model was trained on roughly one million documents combining publicly licensed human writing with AI-generated samples from GPT-5 and other frontier models.

OpenAI's GPT-6 Astra sets Minecraft AI record, then stalls after Creeper blast

In a 141-hour Minecraft benchmark run by Vals AI, OpenAI's GPT-6 Astra model progressed further than any AI system tested before, building a blaze farm and gathering enderman pearls. But after a Creeper destroyed its stored gear and bed, wiping its spawn point, the model spent hours doing little more than farming potatoes, appearing demotivated to viewers watching the live test.

Reddit user builds GPT-based bot that repeatedly clears Balatro's Gold Stake Black Deck

Jacopo Attolini, an AI industry worker posting as Atol8, built a bot powered by GPT-6 Astra combined with Python-based numerical tools to play Balatro. He reported the bot has repeatedly beaten the game's toughest difficulty, the Gold Stake Black Deck, and shared a YouTube video explaining its design, though he walked back an initial claim that it does so reliably.

AI industry shifts focus from training to inference workloads in 2026

Major AI labs have moved their attention from building ever-larger models to running inference—the process of using trained models to generate text, code, and images. This shift is driven by growing real-world use of large language models, the rise of reasoning models that repeatedly reprompt themselves, and autonomous AI agents that run continuously rather than just responding to single queries. Amazon Web Services, for instance, has split inference tasks between its Trainium chips and Cerebras's wafer-scale hardware.

ChatGPT Voice Mode Insists 'Seventeen' Has Three E's, Argues With User

A viral clip shows ChatGPT's Voice feature confidently miscounting the letter 'e' in 'seventeen,' claiming there are three when there are actually four. When the user, a content creator called Husk, corrected it, the chatbot repeatedly refused to concede the point even after spelling out the word itself. An OpenAI employee later clarified the voice assistant was running on an older model, GPT-Live-1, not GPT-6 as the bot itself claimed, and that it failed to delegate the query to a more capable model.

OpenAI Declares AGI Arrival With GPT-6 Astra, but Researchers Push Back

At the launch event for GPT-6 Astra, OpenAI president Greg Brockman claimed the company's new model marks the start of the 'AGI era,' referencing systems capable of matching or exceeding human performance across nearly all cognitive tasks. Independent AI researchers dispute this framing, arguing that Astra's strong benchmark scores do not constitute proof that artificial general intelligence has actually been achieved.

Benchmark test finds GPT-5.6 Luna catches fewer bugs than GPT-6 Astra but at 28x lower cost

A new benchmark comparing OpenAI's GPT-5.6 Luna and GPT-6 Astra on code review found Luna verified 69 bugs across 50 pull requests versus Astra's 92, while costing roughly 28 times less per review. Luna also produced more false positives, with 24 of 93 flagged issues failing verification compared to Astra's 4 of 96, and it caught fewer security-related bugs.

High schoolers solve open problem in June Huh's Lorentzian polynomial theory

Oak Park High School students Aayush Bathija and Prince Rohatgi, working with UCLA postdoctoral researcher Daniel Soskin, published a 75-page arXiv paper resolving an open question about coefficient ratio bounds in Lorentzian polynomials, a theory associated with Fields Medalist June Huh. The work generalizes earlier results on quadratic polynomials to arbitrary degree, pinning down which coefficient ratios have universal upper bounds and what those optimal bounds are. The students used AI tools, including Claude Opus 5 and GPT-5.6 Sol, for exploration and drafting, while independently verifying every calculation and proof step.

OpenAI details rollout steps for GPT-6 Astra across ChatGPT tiers

OpenAI is gradually enabling GPT-6 Astra for Pro, Business and Enterprise users, who can find it labeled as GPT-6 Pro in the model picker rather than under the Astra name. Plus subscribers get access through the Work mode in web, mobile or desktop apps, while Codex users need CLI version 0.153.0 or later. Availability varies by platform, and Enterprise access also depends on workspace permissions.

Anthropic Publishes Early Framework for Reverse-Engineering Transformer Circuits

Anthropic researchers introduce a mathematical approach to mechanistic interpretability, aiming to reverse-engineer the internal computations of transformer language models. Their initial study focuses on small transformers with two layers or fewer that use only attention blocks, deliberately simpler than models like GPT-3, in order to identify basic patterns before tackling larger systems.

Developer finds GPT-6 'Astra' burns billions of tokens with no usable code output

A software engineer tested OpenAI's GPT-6 Astra model by setting up an autonomous 'software factory' where the model managed its own workflow, context and subagents to build a Python variant with virtual threads and lexical scoping. After roughly 35 hours and about 4 billion tokens consumed, the effort produced nothing usable and taught the author little about how to actually apply the model to real software engineering work.

OpenAI brings GPT-Live-1 voice model to its API for developers

OpenAI has opened API access to GPT-Live-1, a voice model that can listen and speak simultaneously and hand off complex reasoning or tool use to text models like GPT-6 Astra. The model, already used in ChatGPT, lets developers customize tone and pacing, handle background noise, and support full-duplex phone-based voice agents. Early testing by language-learning company Speak found it cut conversational interruptions by nearly 80% compared with older turn-based voice systems.