Skip to content
Tech News
clear
Topics: Today This Week This Month This Year

Benchmark tests 10 model-harness pairs on identical Three.js coding task

A developer ran the same prompt—building a self-contained sci-fi hangar scene with Three.js, including hovering drones, animated lights, and camera paths—across 10 combinations of AI models (including GLM, Luna, SOL, Astra, and Qwen variants) and coding harnesses like Codex, OMP, OpenCode, and DSH. The test tracked metrics such as completion time, token usage, tool calls, error rates, and whether the model verified its own output by opening the file in a browser and checking screenshots.

Arm unveils CSS for Mobile 2 with AI-native Mali G2-Ultra GPU for 2027 phones

Arm has introduced its CSS for Mobile 2 platform for chipmakers, featuring a redesigned GPU called the Mali G2-Ultra NX that it describes as its first AI-native graphics processor. The chip uses three AI techniques—neural super sampling, frame rate upscaling, and denoising—to boost resolution and frame rates directly within the graphics pipeline, while updated CPU clusters target AI agent workloads. The technology is expected to appear in premium Android phones starting in 2027.

Nvidia unveils PAIR tool to cluster home GPUs for AI agent tasks

At IFA 2026, Nvidia introduced the Personal AI Router (PAIR), software that distributes sub-tasks from a local AI agent across idle GPUs on other PCs in the same home network. Instead of overloading one machine with multiple sub-agents competing for resources, PAIR spreads the workload and returns results to the main system, adapting in real time as household GPUs become busy or free again.

NVIDIA launches PAIR, a free tool to pool idle PC GPUs for AI tasks

At IFA 2026, NVIDIA unveiled Personal AI Router (PAIR), a free, open-source beta tool that detects idle PCs on a home network and routes AI subagent tasks to them instead of overloading a single GPU. This lets multi-step AI agent jobs, like sorting an inbox, run faster by spreading the workload across multiple machines while the main PC stays free for gaming or work.

Google details limitations for new Gemini 3.8 Flash model

Google published documentation outlining the intended uses and known limitations of Gemini 3.8 Flash, a model aimed at cost-effective, production-scale agent deployment for developers and enterprises. The company flagged persistent issues including hallucinations, occasional slowness or timeouts, and inconsistent knowledge freshness across domains despite a stated March 2026 cutoff.

Perplexity launches Hybrid Compute to mix local and cloud AI processing

Perplexity has introduced Hybrid Compute, a feature that divides AI tasks between cloud-based models and models running locally on a user's device. It flags files or data containing personal information and lets users choose to process those locally while sending the rest of the task to the cloud, or send everything to the cloud. The feature currently works only in the Perplexity app on Apple Silicon Macs, supporting local models like Gemma 4 E4B and Qwen 3.6 alongside cloud options such as Claude Opus 5 and GPT 5.6 Sol.

Anthropic releases Claude Fable 5.1, cuts agentic task costs by up to 45%

Anthropic has released Claude Fable 5.1 and Mythos 5.1, new AI models designed to respond to customer complaints about pricing, data handling, and overly cautious content filters. Fable 5.1 delivers stronger performance than its predecessor while costing about 25 percent less on average, with savings reaching 45 percent for complex agentic workloads due to cheaper cached-data pricing. Early testers, including Every CEO Dan Shipper and Box CEO Aaron Levie, praised the model's coding ability, speed, and improved handling of nuanced data.

Perplexity launches Hybrid Compute to split AI tasks between cloud and local models

Perplexity has introduced Hybrid Compute, a new feature within its Perplexity Computer platform that divides a single task between a cloud-based frontier model, such as Opus 5 or GPT-5.6 Sol, and a smaller model running locally on a user's Mac. The system automatically flags sensitive files or data and routes them to the local model, while less sensitive parts of the task go to the cloud for stronger reasoning power. Users can review and adjust which files are kept local before the task runs, and choose from local options including Gemma E4B and two Qwen 3.6 variants.

Perplexity brings agentic AI 'Personal Computer' feature to Windows

Perplexity has extended its agentic AI feature, previously exclusive to Mac, to Windows 10 and 11 users on paid Pro, Max, or Enterprise plans. The tool can access local files, control native Windows applications, and connect to cloud services like OneDrive, Google Drive, Box, and Dropbox to complete multi-step tasks with minimal user input. A ZDNET reviewer tested it on five complex tasks to gauge its real-world usefulness.

OpenAI Documents ChatGPT's Automations Tool for Scheduled Tasks

OpenAI has published internal reference details describing an 'automations' capability within ChatGPT, meant to trigger actions when a user requests something to happen later, repeatedly, or upon a future condition. This includes reminders, recurring summaries, scheduled searches, and management of existing scheduled tasks.

Android Authority proposes five upgrades Google Clock badly needs

A review piece examines Google's preinstalled Clock app and argues it lags behind specialized alternatives like Alarmy, Chrono and NFC Alarm Clock despite its wide reach on Android phones. The author's top request is adding task-based alarm challenges—math problems, CAPTCHAs, QR or NFC scans, or sliding puzzles—that force users to physically engage before an alarm can be dismissed.

Today's top topics: artificial intelligence samsung snapdragon 8 elite extreme gen 6 donald trump united nations
View all today's topics →