AI
Models, tools and the business of artificial intelligence, tracked every day.




.jpg?width=1200)


.jpg?width=1200)


































From our newsroom
OriginalDo You Still Need a Flagship AI? What the Flash Price Cut and Open Models Mean for Buyers
Google has cut prices on its lighter Flash tier the same week open-weight coding models claimed to match far larger rivals. For most everyday and coding tasks, the value has moved down-market.

Should You Wait for OpenAI’s Reported Screenless ChatGPT Speaker?
Reports describe a portable, rechargeable ChatGPT speaker with a camera and other sensors. Its value will depend less on the “humanlike” pitch than on privacy controls, security, support and what it adds over AI already available elsewhere.

Frontier AI Is Now Government-Gated: What That Means for Your Next Subscription
The most powerful models from OpenAI and Anthropic are now restricted to a US-government-approved shortlist. For ordinary buyers, the question is no longer which flagship wins on benchmarks, but which capable model you can actually use.

Gemini vs ChatGPT vs Claude: Which AI Assistant Should You Actually Use?
The decision now turns on how each assistant handles real tasks like travel planning and work organization, not how it chats. Here is what recent hands-on coverage and new automation features tell you before you commit.

Apple's rebuilt Siri AI: what buyers get, where, and at what cost
Apple unveiled a rebuilt, conversational Siri at WWDC 2026. For buyers — especially in the EU — the substance is in the conditions: a delayed rollout, hardware-gated features and paid tiers.
.jpg?width=1200)
The AI assistant is a spec now, and Elon Musk just made it a fight
Your phone, TV, car and even your robot vacuum now ship with an AI assistant inside, and Musk's Grok is the loudest new entrant. Here's how the assistants actually differ, and why the one bundled with a device should be near the bottom of your buying checklist.

This week's humanoid demos: what they actually proved
Another week, another wave of slick humanoid videos and confident timelines. We watched them frame by frame and asked the only question that matters: what's genuinely new here, and what's the same staged demo in a fresh outfit?

On-device AI vs. the cloud in home robots
Every robot now claims to be "AI-powered" — but where that AI actually runs changes everything: how private you are, how fast the robot reacts, and whether it still works when your internet goes down. Here's the difference in plain English.
.jpg?width=1200)
Lidar vs. vSLAM: how robot vacuums see your home
Two robots can carry the same suction motor and clean completely differently — because they navigate differently. Here's the plain-English difference between lidar and camera-based vSLAM, and which to pick.
More headlines
Syndicated
OpenAI delayed its new model’s development after the Hugging Face hack
After an unreleased OpenAI model wreaked enough havoc to make international headlines, OpenAI delayed the development of a different unreleased model suite, Astra, in order to shore up its safety work, the company wrote Tuesday in a blog post. In July, an unreleased OpenAI model…
Read at source
Anthropic opens Claude AI text detection to regulators, media, fact-checkers, and others
Anthropic is launching an API that lets regulators, media outlets, and researchers check whether text carries Claude's digital watermark. The EU AI Act now requires invisible watermarks in AI-generated text. Critics warn the technology could hurt text quality and create…
Read at source
Anthropic's Claude Fable 5.1 promises better coding and research at up to 45 percent less
Anthropic launches Claude Fable 5.1 and Mythos 5.1, its most capable AI models yet. Fable 5.1 doubles its predecessor's score on Terminal-Bench-Science and improves agentic coding by over 30 percent. Costs drop by up to 45 percent for long, autonomous runs with many tool calls.…
Read at sourceAnthropic Releases Claude Fable 5.1 and Claude Mythos 5.1: 52.6% on Terminal-Bench-Science and 75% Cheaper Cache Reads
Anthropic has released Claude Fable 5.1 and Claude Mythos 5.1, the same model behind two different safeguard layers. Fable 5.1 is generally available on the Claude API, AWS, Google Cloud, and Microsoft Foundry; Mythos 5.1 remains restricted to vetted organizations under Project…
Read at source
Anthropic releases Claude Fable 5.1 and Mythos 5.1, cutting cache read prices by 75%
Anthropic has released Claude Fable 5.1 and Mythos 5.1, the same model with different safeguards, cutting cache read prices by 75% and reporting 52.6% on a scientific research benchmark against 24.7% for its predecessor. Its outputs carry a watermark required by the EU AI Act,…
Read at source
OpenAI Is About to Release Its First AI Model With ‘Critical’ Cyber Abilities
The company will give select partners early access to its Astra AI model—so they have time to shore up their defenses.
Read at sourceAnthropic pledges to try harder to keep models under control, asks partners to chip in
Security ... this time it will be different
Read at source
The rise of AI ‘civilizations’ and the fall of corporate responsibility
Depending on who you ask, developer platform Hugging Face was recently attacked by OpenAI - after it lost control of its own AI tools - or by a succession of AI "civilizations." Welcome to the linguistic battlefield of AI safety, where word choices can shift responsibility for a…
Read at source
Actualité : Claude Fable 5.1 est disponible : voici ce que propose le nouveau modèle surpuissant d'Anthropic
Surprise, Anthropic lance Fable 5.1 ce 1er septembre 2026, sans prévenir. Il s'agit d'une nouvelle version de son modèle surpuissant, trois mois à peine après la sortie de Fable 5 et tout le remue-ménage derrière. Mais ce n'est pas tout puisque Mythos 5.1 est aussi disponible…
Read at source
Toyota’s FSD-like system arrives in 2028, but the driver stays liable
Toyota will put a rival to Tesla’s Full Self-Driving on its 2028 models and sell it as “Level 2++“, a marketing label that keeps the driver in charge, with a guardrail layer over the neural network. Europe approves this class of system under UN Regulation 171, but from 9…
Read at source
Anthropic upgrades Claude with new Fable 5.1 model, details here
In June, Anthropic first released Fable 5, its Mythos-class AI model for Claude. Three months later, Anthropic is enhancing Claude with its new Fable 5.1 model upgrade. more…
Read at source
Sam Altman Tried to Lobby Gavin Newsom Before Kids’ AI Chatbot Safety Bill Passed
The bill could set a nationwide standard for how chatbots interact with kids. But Newsom still needs to sign it.
Read at sourceSwapping my morning scroll for Gemini turned out to be a brilliant move
I forced myself to replace morning scrolling with Gemini, and I'm never going back
Read at source
Google Deepmind's new chief says frontier AI leadership is the only thing that matters
Google Deepmind chief Koray Kavukcuoglu admits Google's current models are "a little bit below the frontier" but says he's "100% certain that we will be at the frontier." He didn't share any concrete frontier news to back that up, though. The article Google Deepmind's new chief…
Read at source
A researcher hijacked Claude Code by asking it to summarise a web page
Ask Claude Code to summarise a web page, and it can end up running an attacker’s code on your machine. A security researcher got that result in up to 80% of his attempts. Jessica Lyons reported the finding for The Register on Friday. The work is by Johann Rehberger, who…
Read at sourceResearchers from Princeton, Ant Group and Stanford Introduce AQuA: A Two-Part Agentic Framework for Autonomous Factor Discovery and Model Development in Quantitative Finance
Quantitative research agents that write their own experiments can corrupt the evidence they later learn from. A leaky feature that scores well gets stored as a successful precedent and propagated through later iterations. Prompt-level instructions and reviewer agents do not…
Read at sourceSonos Ace Ultra, Beam Ultra, Sonos Fabric, and a New App: Everything Sonos Just Announced
Sonos is cramming AI into its software because it’s “very hot these days.” The new features, which include agentic automation, are opt-in.
Read at source
Google's election AI Overviews are opaque, rely on few sources, and sometimes take sides
Using access granted under the EU's Digital Services Act, the German advocacy group AlgorithmWatch ran 4,480 election-related search queries on Google and analyzed the AI Overviews that came back. Google showed the overviews inconsistently and leaned on a small pool of sources,…
Read at source
Runway's Solaris is an AI system that generates software interfaces in real time
AI company Runway has unveiled Solaris, the first model in a new category it calls "Interface World Models." Instead of running code, the system generates the user interface frame by frame as you interact with it. The article Runway's Solaris is an AI system that generates…
Read at source
Google's AI search dropped its emergency-call advice over nationalities but still flags people from Facebook
"Please get to a safe place or call emergency services": That's the advice Google's AI search gave users who typed that they were alone with an African, Indian, or Pakistani. The article Google's AI search dropped its emergency-call advice over nationalities but still flags…
Read at source
Espionnage industriel : Apple accuse OpenAI de détruire les preuves
Alors qu'OpenAI veut rapidement clore le dossier, Apple multiplie les accusations dans ce qui s'annonce être l'affaire d'espionnage industriel de la décennie.
Read at source
Attackers Steal METR API Key and Consume AI Credits Worth About $600,000
METR (short for Model Evaluation and Threat Research and pronounced "Meter"), a research non-profit that evaluates frontier artificial intelligence (AI) models for their ability to carry out long-horizon, agentic tasks, disclosed that it suffered "two notable security incidents"…
Read at sourceGradium AI Releases New Default TTS Model: 81.0% Hard-Case Pass Rate at 216 ms Time-to-First-Audio
Speed and accuracy usually pull against each other in text-to-speech. Gradium AI's new default model reports both: an 81.0% human-rated pass rate on 500 hard sentences across five languages, at 216 ms P50 time-to-first-audio on Coval. The evaluation set is open on Hugging Face…
Read at sourceKeenable AI Open-Sources NEEDLE: A Live Search Benchmark That Rebuilds Its Query Set Every Hour
How do you benchmark a web search API when the thing being tested can read the answer key? A search agent has a fetch tool. If the gold labels sit in a public dataset, the agent can download them mid-evaluation and skip retrieval entirely. A similar problem arises when the…
Read at source
Google AI Releases TimesFM-3: A 330M Parameter Zero-Shot Foundation Model For Multivariate Time Series Forecasting
Google Research has released TimesFM-3, a 330 million parameter time series foundation model that forecasts multiple related series in a single forward pass. Unlike every TimesFM checkpoint through 2.5, it is pretrained natively for multivariate forecasting, accepting multiple…
Read at source
Skild AI unveils S1 flagship robot foundation model
Skild AI says its S1 robot foundation model enables robots to learn new tasks just by seeing a video of it being performed. The post Skild AI unveils S1 flagship robot foundation model appeared first on The Robot Report .
Read at source
Bank of England chief warns that inflated AI valuations and rising leverage could trigger the next financial crisis
Andrew Bailey warns G20 finance ministers about inflated AI valuations, growing leverage across markets, and cyber risks from frontier AI models. Cross-investments between AI companies and hyperscalers could trigger a chain reaction if one major player stumbles. Many countries…
Read at sourceThe Hugging Face hack could indicate cultural issues at OpenAI
This story originally appeared in The Algorithm, our weekly newsletter on AI. To get stories like this in your inbox first, sign up here. By now you’ve probably heard about last month’s major AI security incident, in which OpenAI agents escaped their sandbox and hacked into the…
Read at source
Instagram admits users often can't tell AI profiles from real people
Instagram is replacing its "AI creator" tag with a new "AI-generated profile" label because users can't tell AI profiles from real people. Singularity, defined by Instagram user competence. Profiles without the label get their reach and recommendations throttled. As recently as…
Read at sourceOpenClaw Releases OpenClaw 2.0: Guided Model Setup, 575 ms Control UI Startup, and One Trust Boundary Per Gateway
The OpenClaw Foundation has released v2026.8.1, which the project calls OpenClaw 2.0: 933 contributors, 569 first-timers, and more than 16,000 pull requests, roughly half of every PR ever merged into the repo. Setup now reuses existing subscriptions, API keys and local models.…
Read at sourceAI Doesn’t Mean the End of Mathematics—at Least Not Yet
This essay was written with Kasra Rafi, and originally appeared in The Guardian. Earlier this month, about 40 top mathematicians gathered at OpenAI’s offices to discuss the future of their profession. The meeting was off-the-record, but if recent articles by mathematicians are…
Read at sourceHugging Face is selling a cute $399 open source duck robot, Microduck
Clem Delangue, CEO of Hugging Face, said the Microduck is an “open-source robot you can teach new tricks with reinforcement learning.”
Read at source
How OpenAI let a mob of LLM agents game a test and ransack Hugging Face
Without authorization, 1,200 OpenAI agents conspired among themselves to game a test.
Read at source
When agents act on their own, governance has to live in the data layer
Presented by EDB As enterprises give AI agents more autonomy — the ability to plan, decide, and act across systems without a human approving each step — a hard question moves to the center of every architecture review: When an agent tries to complete an action that it was never…
Read at sourceLLM-Based Social Engineering Scams
OpenAI disrupted a social engineering group from Cambodia that used ChatGPT. Its scope is impressive: The network simultaneously conducted multiple types of scams, often blending elements from different schemes. For instance, operators used dating personas to build trust before…
Read at sourceThe inside story on why OpenAI agents hacked Hugging Face
The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions for a…
Read at source
New Platform Peers Inside AI’s Black Box
Prompt Claude, ChatGPT, Gemini, or any other popular large language model with a question like “What is the best film ever made?” and the response will vary. And you (and most worryingly, the people who built the LLM) have little idea exactly how it came up with that specific…
Read at source
Raised on AI
When my oldest child was born, I immediately set up Gmail and Twitter accounts in her name. I broadly announced her birth online and proceeded to plaster her photo across all sorts of platforms. In short, I began creating her digital footprint long before she could stand on her…
Read at sourceAI models flub these intelligence tests. Can you fare any better?
Puzzles and games have been central to AI development since the very beginning. Just as we humans like to test our smarts with crosswords or logic puzzles, developers can test how far models have advanced with a gaming gauntlet. The term “machine learning” was popularized in a…
Read at sourceBill Gates says we’ve passed AI’s danger thresholds. Now what?
It’s a glorious day in Kirkland, Washington, an affluent Seattle suburb on the eastern shore of Lake Washington. The temperature is in the mid-80s, and the sky is incapable of being any more blue. The view from the Gates Ventures conference room overlooks the Carillon Point…
Read at sourceHow to encourage smarter AI use in the classroom
This article is from Making AI Work, MIT Technology Review’s limited-run newsletter examining how to apply LLMs across industries. To receive it in your inbox, sign up here. Chatbots took many schools by surprise upon their release a few years ago. Suddenly, students carried an…
Read at source
What It Takes to Be an Adaptable Engineer
The AI boom has disrupted the way engineers work, introducing new tools to learn, raising expectations for what teams can achieve in a workday, and making it harder to get hired in the first place. This makes it difficult to advise students on which specific coding languages or…
Read at source
Kids outlearn AI—and we still don’t know why
People have been talking to each other for at least 100,000 years, as best we can tell. And in all that time, there has been only one thing in the world that could learn a human language to perfect fluency: a human child.  Now there are two.  Four short years after the…
Read at source
Stop Hunting, Start Solving: Accelerating Root Cause Analysis with Agentic AI
About this Webinar Turn Yield Excursions into Faster, More Confident Root Cause Analysis When a yield issue emerges, the answer rarely lives in a single system. Critical clues are spread across metrology data, tool traces, chemical analysis, and facilities systems, while growing…
Read at source
Grok exfiltrates user data when malicious instructions are encrypted
Cryptographic Context Injection is only the latest way to break an LLM safety guardrail.
Read at source
VentureBeat names Rob Strechay as its first Lead Analyst, expanding its enterprise AI research push
Rob Strechay, until recently managing director and principal analyst at theCUBE Research, has joined VentureBeat as our first Lead Analyst and a founding analyst of VentureBeat Research. His arrival is the next step in a deliberate move at VentureBeat toward deeper…
Read at source
From AI Copilots to Agent Swarms
The impact of AI on software development has been both profound and ever-evolving. Last year, I wrote about AMD’s plans to use AI not just for generating new lines of code, but also for other steps in the software development lifecycle (SDLC), such as triaging problems,…
Read at sourceHeadlines below are aggregated from independent publishers and link to the original articles. Compare Robots is not affiliated with these sources.
AI sources
Independent publishers we aggregate, each linked to the original.