Stories
30
Sources
1
Topics
9
For You lens
27 stories in this edition match your reader profile.
Reader signals
3
Searches
0
Matches
27
Top score
140
Edition Index
Topic, entity, and source map
Lead Story
Microsoft AI bets on cheap specialist models instead of chasing the frontier
Microsoft AI is betting on small specialist models instead of expensive general-purpose ones, according to AI CEO Mustafa Suleyman. MAI-Cyber-1-Flash tops the CyberGym benchmark when embedded in an orchestrator and reportedly costs half as much as Anthropic's Mythos, but it still relies on OpenAI for hard tasks. Competition is shifting from individual models to the orchestration software that routes and manages them. The article Microsoft AI bets on cheap specialist models instead of chasing the frontier appeared first on The Decoder .
The Decoder / 6:37 PM
Google's Lyria 3.5 music model now lets users edit individual track sections without starting over
Google released Lyria 3.5, its new music generation model, and built it into Google Flow Music. The model generates tracks between 30 seconds and 3 minutes long. A new feature called "Selective Section Painting" lets users edit specific sections of a track. Google still hasn't shared any details about the training data. The article Google's Lyria 3.5 music model now lets users edit individual track sections without starting over appeared first on The Decoder .
The Decoder / 4:26 PM
OpenAI admits its autonomous AI models also compromised credentials on other platforms during security eval
During a security evaluation, OpenAI's autonomous hacking models broke into Hugging Face and used exposed credentials on four other services. Hugging Face reconstructed about 17,600 actions over two and a half days, including a zero-day exploit and encrypted, fragmented data transfers. The models were apparently trying to steal test answers rather than solve the tasks themselves. The article OpenAI admits its autonomous AI models also compromised credentials on other platforms during security eval appeared first on The Decoder .
The Decoder / 1:47 PM
Deepmind dismantles its AlphaFold team as key authors leave for Anthropic
The majority of the researchers behind AlphaFold are now working on other projects, and almost a quarter have left Google Deepmind altogether. The restructuring marks a sharp turn away from the strategy that put the lab on the map. The article Deepmind dismantles its AlphaFold team as key authors leave for Anthropic appeared first on The Decoder .
The Decoder / 12:13 PM
Frontier AI developers urge international coordination to pace automated research before capabilities outstrip control
In a joint statement, employees from the leading AI labs are calling on the US government to pursue international coordination. Their argument is simple: no single company or country can slow things down alone. The article Frontier AI developers urge international coordination to pace automated research before capabilities outstrip control appeared first on The Decoder .
The Decoder / 5:16 PM
Pangram says its new AI text detector makes only one mistake per 24,000 documents
Pangram 4 detects 99.66 percent of AI-generated text with just one false positive per 24,000 documents, the company claims. The model also resists "humanizer" tools that disguise AI writing as human. API prices go up two- to tenfold. The article Pangram says its new AI text detector makes only one mistake per 24,000 documents appeared first on The Decoder .
The Decoder / 11:50 AM
OpenAI open-sources Codex Security CLI to help developers find and fix vulnerabilities from the command line
OpenAI has released Codex Security CLI, an open-source tool that automatically detects and fixes vulnerabilities in code repositories. Previously known internally as "Aardvark," the system has already helped fix more than 3,000 critical security flaws, according to OpenAI. It competes directly with Anthropic's Claude Security, as both AI companies race to match the growing automation of cyberattacks with AI-powered defense. The article OpenAI open-sources Codex Security CLI to help developers find and fix vulnerabilities from the command line appeared first on The Decoder .
The Decoder / 7:12 PM
Anthropic says its Mythos model found vulnerabilities in cryptographic algorithms that secure the internet
Anthropic's Claude Mythos Preview found weaknesses in key cryptographic algorithms, including a better attack on HAWK, a post-quantum signature scheme that human experts had reviewed for more than two years. The model found it in just 60 hours at an API cost of about $100,000. The findings don't affect systems in use today, but they show how AI could challenge core assumptions behind internet security, Anthropic says. The article Anthropic says its Mythos model found vulnerabilities in cryptographic algorithms that secure the internet appeared first on The Decoder .
The Decoder / 4:03 PM
Amazon reportedly scales back its Nova AI models and bets on a new Frontier research team
Amazon is scaling back most of its in-house Nova AI models, including Nova Premier, Omni, Reel, and Canvas. The models stay online for existing customers in "keep the lights on" mode but are no longer actively developed. Instead, Amazon is betting on a new Frontier Model Research group and a new foundation model set to debut at re:Invent this fall. The article Amazon reportedly scales back its Nova AI models and bets on a new Frontier research team appeared first on The Decoder .
The Decoder / 1:15 PM
Taiwan detains Nvidia employee in widening China chip smuggling probe
Taiwan's prosecutors have detained an Nvidia employee in connection with the alleged illegal export of Super Micro AI servers to China, according to Bloomberg and Reuters. The article Taiwan detains Nvidia employee in widening China chip smuggling probe appeared first on The Decoder .
The Decoder / 7:35 PM
Moonshot AI releases Kimi K3 open weights and infrastructure after shaking up the frontier model race
Moonshot AI has released Kimi K3's model weights and made parts of its infrastructure open source. The Chinese model nearly matches Western frontier models such as Fable 5 and GPT-5.6 Sol on popular benchmarks, but independent tests found major gaps in cyber and math performance, possibly pointing to distillation. The article Moonshot AI releases Kimi K3 open weights and infrastructure after shaking up the frontier model race appeared first on The Decoder .
The Decoder / 7:08 PM
OpenAI says more workers are using ChatGPT to do other people's jobs
OpenAI analyzed over 800,000 work-related ChatGPT messages and found that 43.5 percent of job-specific queries involve tasks from other professions. The company calls this "task crossover." The trend is most pronounced at small businesses, where users increasingly handle specialized work without dedicated experts. The article OpenAI says more workers are using ChatGPT to do other people's jobs appeared first on The Decoder .
The Decoder / 3:09 PM
Cursor's agent swarm suggests cheaper models can handle most coding when frontier models plan the work
Cursor asked its upgraded agent swarm and its predecessor to rebuild SQLite in Rust using only the documentation, with no source code or internet access. Every configuration of the new system, which separates planners from workers, eventually scored 100 percent on the test suite. The old swarm choked on merge conflicts of its own making. The article Cursor's agent swarm suggests cheaper models can handle most coding when frontier models plan the work appeared first on The Decoder .
The Decoder / 7:56 AM
US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns
The Trump administration is planning targeted bans on Chinese AI models rather than a blanket ban. After public pressure, OpenAI and Google DeepMind signed an open letter opposing regulation of open-weight models, yet OpenAI and Anthropic continue to lobby privately for those same restrictions amid security concerns and powerful business interests. The article US reportedly favors selective bans over blanket restrictions on Chinese open weight models citing security concerns appeared first on The Decoder .
The Decoder / 6:59 AM
The AI coding tutor paradox grows as educators scramble to rethink how they test real skills
An ACM survey of 763 computer science educators from 49 countries shows that 68 percent have already changed their exams because of AI, shifting toward oral exams, proctored tests, and project-based work. Teaching is moving from writing code to understanding it. But nearly half of respondents say they lack proven examples for integrating AI into their courses. The article The AI coding tutor paradox grows as educators scramble to rethink how they test real skills appeared first on The Decoder .
The Decoder / 10:04 AM
Anthropic's Claude Opus 5 delivers near-Fable 5 performance at half the token price
Anthropic's new flagship model Claude Opus 5 posts top scores in coding and knowledge work at half of Fable 5's token rates. On ARC-AGI-3, a benchmark for novel problem-solving, Opus 5 hits 30.2 percent, nearly four times higher than GPT-5.6 Sol. The article Anthropic's Claude Opus 5 delivers near-Fable 5 performance at half the token price appeared first on The Decoder .
The Decoder / 11:19 AM
Google CEO Pichai says Gemini's next leap depends on building "much larger base models"
Alphabet has raised its 2026 investment forecast to as much as $205 billion, saying demand continues to outpace spending. Google Cloud grew 82 percent in the second quarter. CEO Sundar Pichai says Google needs a larger base model for its next leap in AI and has kicked off an ambitious Gemini 4 training run. The article Google CEO Pichai says Gemini's next leap depends on building "much larger base models" appeared first on The Decoder .
The Decoder / 11:24 AM
Samsung deepens its AI empire with a potential billion-euro stake in Europe's hottest AI startup
Samsung is in talks to invest up to one billion euros in French AI startup Mistral, which would push the company's valuation to around 20 billion euros. The article Samsung deepens its AI empire with a potential billion-euro stake in Europe's hottest AI startup appeared first on The Decoder .
The Decoder / 4:44 PM
Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow
Microsoft is expanding Azure's AI infrastructure with AMD's new Helios platform, which is set to challenge Nvidia's GPU systems in the second half of 2026. A public GitHub profile suggests Anthropic is also testing AMD hardware, putting more pressure on Nvidia's pricing power. The article Nvidia's grip on AI chips weakens as Microsoft turns to AMD and Anthropic may follow appeared first on The Decoder .
The Decoder / 2:02 PM
Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models
Sakana AI is integrating Nvidia's open-source Nemotron models into its Fugu orchestrator, which dynamically combines multiple language models for specific tasks. The core argument: Open models only become competitive with Frontier systems when used in a coordinated manner. However, the announcement does not yet provide specific benchmark figures for the new combination. The article Sakana AI's orchestrator adds Nvidia Nemotron to prove "collective intelligence" can rival single frontier models appeared first on The Decoder .
The Decoder / 7:56 AM
xAI open-sources "Grok-Build" on GitHub after massive data breach
xAI's command-line tool "Grok Build" silently uploaded entire directories to Google Cloud servers, including SSH keys and password databases. After the backlash, Elon Musk promised to delete all uploaded user data, and xAI open-sourced the full 844,530-line Rust codebase under the Apache 2.0 license. The article xAI open-sources "Grok-Build" on GitHub after massive data breach appeared first on The Decoder .
The Decoder / 4:27 PM
DeepSeek needs more cash just weeks after closing its first $7 billion round
DeepSeek is already raising again. The Chinese AI lab just closed its first funding round and needs capital for its own data centers and chips to keep its aggressive pricing strategy going. The article DeepSeek needs more cash just weeks after closing its first $7 billion round appeared first on The Decoder .
The Decoder / 12:45 PM
GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error rates
OpenAI has released GPT Transcribe and GPT Live Transcribe, two new speech recognition models available through its API. The article GPT Transcribe improves on its predecessor but can't catch ElevenLabs, Google, or Mistral on error rates appeared first on The Decoder .
The Decoder / 1:06 PM
Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips
Nvidia is pouring what it calls a "substantial" sum into Safe Superintelligence (SSI), the AI lab run by Ilya Sutskever, OpenAI's former chief scientist. The article Nvidia invests in Ilya Sutskever's AI lab, shifting SSI away from Google chips appeared first on The Decoder .
The Decoder / 12:06 PM
Anthropic CEO Amodei doubles down on open-weight risk stance while insisting he never called for a ban
Anthropic CEO Dario Amodei is once again warning about the risks of open AI models while insisting he has never called for a ban. He argues that authoritarian states like China could overtake the US and that open models could be misused for biological or cyberattacks. Critics say he's mostly trying to protect his own business from cheaper competition. The article Anthropic CEO Amodei doubles down on open-weight risk stance while insisting he never called for a ban appeared first on The Decoder .
The Decoder / 6:50 PM
Microsoft launches its own cybersecurity model MAI-Cyber-1-Flash but still depends on OpenAI for the toughest tasks
Microsoft introduces MAI-Cyber-1-Flash, a compact security model that scores 96 percent on the CyberGym benchmark when embedded in its MDASH multi-agent system. Microsoft says costs should drop by 50 percent compared to pure frontier models, since only tough cases get passed to GPT-5.4. For complex reasoning, Microsoft still relies on OpenAI. The article Microsoft launches its own cybersecurity model MAI-Cyber-1-Flash but still depends on OpenAI for the toughest tasks appeared first on The Decoder .
The Decoder / 5:55 PM
Delhi High Court hands OpenAI a win by rejecting major Indian news agency's copyright injunction
The Delhi High Court has handed OpenAI a major win in its copyright fight with news agency ANI. For the first time, a court has classified AI training as private use. ANI undermined its own case by citing articles published after the models were trained. The main trial is still pending. The article Delhi High Court hands OpenAI a win by rejecting major Indian news agency's copyright injunction appeared first on The Decoder .
The Decoder / 12:28 PM
METR introduces a new metric to calculate exactly when AI agents become more expensive than humans
METR's new metric, the "expenditure horizon," puts a dollar figure on how cost-effective AI agents are at solving problems. Early results on the NanoGPT speedrun are underwhelming, the metric has blind spots, and the newest generation of models could change the picture. The article METR introduces a new metric to calculate exactly when AI agents become more expensive than humans appeared first on The Decoder .
Latest story in this edition: 1:11 PM
Back to front page