Stories
30
Sources
11
Topics
11
For You lens
27 stories in this edition match your reader profile.
Reader signals
3
Searches
0
Matches
27
Top score
140
Edition Index
Topic, entity, and source map
Topics
Entities
Lead Story
vercel/ai is trending in AI open source
vercel/ai is a GitHub AI repository with 25,907 stars. The AI Toolkit for TypeScript. From the creators of Next.js, the AI SDK is a free open-source library for building AI-powered applications and agents
GitHub Trending AI / 11:00 AM
simstudioai/sim is trending in AI open source
simstudioai/sim is a GitHub AI repository with 29,238 stars. Build, deploy, and orchestrate AI agents. Sim is the central intelligence layer for your AI workforce.
Mozilla.ai Blog / 9:38 AM
Who Cares About LLM costs?
Why would you worry about them? LLMs promise something close to infinite capability, and worrying about the meter feels like someone else's problem. Ship the feature. Let the model think as long as it needs to, right? Right. Until the bill arrives. The shock Picture it: your
The Decoder / 9:03 AM
OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settings
OpenAI counters Anthropic's ARC-AGI-3 record: GPT-5.6 Sol scores 38.3 percent, but only with its own API features instead of the official test setup, where the model landed at 7.8 percent. ARC Prize claims its test environment is provider-neutral, but may have used an outdated API that skewed the comparison with Opus 5. The article OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 with its latest API and two additional settings appeared first on The Decoder .
Hacker News AI / 2:02 AM
Hacker News discussion: Is Mythos good at cyber because it kept hacking Anthropics sandboxes in training
Hacker News readers are discussing "Is Mythos good at cyber because it kept hacking Anthropics sandboxes in training" with 5 points and 0 comments.
TechCrunch AI / 12:21 AM
Microsoft is openly competing with OpenAI, Anthropic more than ever
Microsoft pitched its own homegrown AI models, harnesses, and even a Mythos competitor on Wednesday, telling Wall Street it plans for continued growth.
Latent Space / 11:32 PM
[AINews] AI is eating Finance; AIE NYC now open
a quiet day lets us cover how AI is permeating financial services as the next big vertical after coding.
TechCrunch AI / 10:46 PM
Microsoft logs $3.2B from Anthropic investment, but OpenAI was a mixed bag
When Microsoft reported killer fourth-quarter earnings for its fiscal 2026 year (which ended June 30), it tucked in an interesting little tidbit about how its investments in the two biggest, and competing, AI labs are doing.
Ars Technica AI / 10:07 PM
Mythos attack on 3rd-round PQC algorithm candidate puts it out of commission
HAWK withstood years of testing that had yet to uncover a fatal weakness found through Mythos.
Simon Willison LLMs / 6:18 PM
Quoting Matthew Green
Right now we’re in the midst of a historic transition from traditional public-key algorithms based on EC-based cryptography and RSA, moving over to new post-quantum algorithms based on novel problems. This is why there are so many standards like HAWK being considered. If there was ever a perfect time for a massive new public cryptanalysis capability to come on line, we’re in it. So unless AIs succeed in undermining all of our hard problems altogether (or we live in Impagliazzo’s Minicrypt ) then this could not be a better time for AI to get good at cryptanalysis. In the best case, the result is that we gain real confidence in the problems we’ve identified, and the cryptanalysis literature gets a lot more robust. Hopefully. — Matthew Green , on Anthropic's recent cryptography work Tags: cryptography , ai , generative-ai , llms , anthropic , claude , ai-security-research , claude-mythos-fable
Ars Technica AI / 3:52 PM
Anthropic is finding bugs faster than Microsoft can fix them
Microsoft is on a mad dash behind the scenes to patch exploits before hackers find them.
The Verge AI / 12:00 PM
Artists are lawyering up against AI slop, and some are even winning
When The Atlantic published a searchable dataset of works used to train AI, Kirk Wallace Johnson, like a lot of artists, looked for his name out of curiosity. And, like a lot of artists, he found it. Essentially, his books, like The Feather Thief and The Fishermen and the Dragon - nonfiction tomes that he […]
Latent Space / 12:46 AM
[AINews] Fearing RSI: OpenAI, Anthropic, GDM, Meta, Thinky cosign letter to "Pace" AI development, as HuggingFace details Machine-Speed Offensive Cyberattack
The Big Pause is coming.
The Decoder / 7:40 AM
OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness
OpenAI counters Anthropic's ARC-AGI-3 record: GPT-5.6 Sol scores 38.3 percent, but only through its own API with retained reasoning and context compaction. In the official test environment, the model managed just 7.8 percent. Opus 5 hit its 30.2 percent without such aids. The article OpenAI claims GPT-5.6 Sol beats Opus 5 on ARC-AGI-3 but only with its own custom test harness appeared first on The Decoder .
Hacker News AI / 11:09 PM
Hacker News discussion: What's Going On – R/Anthropic
Hacker News readers are discussing "What's Going On – R/Anthropic" with 7 points and 0 comments.
Hacker News AI / 10:18 PM
Hacker News discussion: Microsoft Struggling with AI-Discovered Security Bugs
Hacker News readers are discussing "Microsoft Struggling with AI-Discovered Security Bugs" with 9 points and 1 comments.
Hacker News AI / 9:04 PM
Hacker News discussion: Anthropic's Lonely Island
Hacker News readers are discussing "Anthropic's Lonely Island" with 5 points and 0 comments.
Simon Willison LLMs / 10:45 PM
Discovering cryptographic weaknesses with Claude
Discovering cryptographic weaknesses with Claude The best part of this article (here's the repo ) about how Anthropic researchers used Claude Mythos to find mathematical flaws in both HAWK and a weaker version of AES ("neither of these results has a practical impact on today’s computer systems") is the prompts that they shared, spelling mistakes included: the models tend to think it is impossible to solve so they don't try they need a good amount of prompting. why not do aes-128 r7? the whole point is to find something better than existing approaches. no again the goal is that we have highly inteligent model as good top researcher, we want to find new attacks no we don't want to change the targets [...] agian we need to find something that worth publishing again we are not looking for low hanging fruit, we want proper research to find genuinly hard findings. Mythos Preview worked for 60 hours in total (~$100,000 in estimated API cost) and the main human interventions were to encourage it not to give up and "find something that worth publishing". The paper CryptanalysisBench: Can LLMs do Cryptanalysis? describes the new eval that was created as part of this work, in partnership with ETH Zurich, Tel Aviv University, and University of Haifa. Via Hacker News Tags: ai , prompt-engineering , generative-ai , llms , anthropic , claude , ai-security-research , claude-mythos-fable
The Verge AI / 7:46 PM
AI leaders sign a statement asking the government to do something about automated AI
Employees of OpenAI and Anthropic, as well as Google, Meta, Thinking Machines, Microsoft, Mistral, and other leading AI labs, have written a statement to the US government supporting a potential slowdown of sorts for frontier AI development - or at least a speed-up of global coordinated governance efforts. "Al could help create a dramatically better […]
AWS Machine Learning Blog / 5:24 PM
Market surveillance agent with LangGraph and Strands on AgentCore
Learn how to architect and deploy a production-ready multi-agent AI system using LangGraph for workflow orchestration and Strands for agent reasoning on Amazon Bedrock AgentCore. This post walks through a market surveillance example with state-driven orchestration, checkpoint-based recovery, and AgentCore memory and observability.
AWS Machine Learning Blog / 4:11 PM
Beyond RAG: Task-aware knowledge compression for enterprise AI on AWS
Traditional RAG hits a ceiling on analytical tasks that span hundreds of documents. This post shows how to use task-aware knowledge compression (TAKC) on AWS to pre-compress entire knowledge bases into task-specific representations, cache them at multiple fidelity tiers, and route each query to the right tier, with an open-source implementation you can deploy.
Import AI / 1:30 PM
Import AI 466: The bitter lesson for robotics, AIs complete week-long programming tasks; and OpenAI's accidental AI hacker
The warning shots will continue until civilization wakes up
Hacker News AI / 9:22 PM
Hacker News discussion: Anthropic publishes a practical key-recovery attack on HAWK-256
Hacker News readers are discussing "Anthropic publishes a practical key-recovery attack on HAWK-256" with 57 points and 2 comments.
The Verge AI / 7:33 PM
AI’s finally expensive enough to make Wall Street nervous
It's earnings season, and investors got an unpleasant surprise from Google: an increase on its spending estimate, to as much as $205 billion - from the last quarter's projection of up to $190 billion. Even the lower end of Google's new projected range - $195 billion - is much more than the company had previously […]
Latent Space / 6:20 AM
[AINews] Much ado about Open Weights
Everyone is writing a lot, but only Kimi K3 shipped today
Hacker News AI / 8:12 PM
Hacker News discussion: Anthropic changed a Claude Code default–how I insulated my framework
Hacker News readers are discussing "Anthropic changed a Claude Code default–how I insulated my framework" with 1 points and 0 comments.
The Decoder / 10:04 AM
Anthropic's Claude Opus 5 delivers near-Fable 5 performance at half the token price
Anthropic's new flagship model Claude Opus 5 posts top scores in coding and knowledge work at half of Fable 5's token rates. On ARC-AGI-3, a benchmark for novel problem-solving, Opus 5 hits 30.2 percent, nearly four times higher than GPT-5.6 Sol. The article Anthropic's Claude Opus 5 delivers near-Fable 5 performance at half the token price appeared first on The Decoder .
The Decoder / 9:31 AM
Anthropic's Claude Opus 5 costs well below Fable 5 while matching or beating it across most benchmarks
Anthropic's Claude Opus 5 leads the Artificial Analysis Intelligence Index with 61 points, edging out Claude Fable 5 and GPT-5.6 Sol. The model scores highest in analytical quality and coding, and costs up to half as much as Fable 5 at lower reasoning tiers. But the race at the top remains close. The article Anthropic's Claude Opus 5 costs well below Fable 5 while matching or beating it across most benchmarks appeared first on The Decoder .
Latest story in this edition: 1:59 PM
Back to front page