Stories
30
Sources
10
Topics
11
For You lens
24 stories in this edition match your reader profile.
Reader signals
3
Searches
0
Matches
24
Top score
116
Edition Index
Topic, entity, and source map
Topics
Entities
Lead Story
simstudioai/sim is trending in AI open source
simstudioai/sim is a GitHub AI repository with 29,626 stars. Sim is the collaborative workspace to build, deploy, and monitor AI agents and workflows. Used by 100,000+ builders.
Hacker News AI / 8:55 AM
The Newsroom EP01 – OpenAI Ships GPT-5 to Azure – DeepSeek Open-Sources 236B Moe
HN 2 pts · 0 comments
Latent Space / 5:56 AM
[AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale
We agree with Sebastian: this should have been DeepSeek v5
The Decoder / 1:50 PM
How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data
Anthropic's new threat intelligence report documents eight months of Claude abuse. Chinese AI labs like Alibaba's Qwen team, DeepSeek, and Moonshot AI relayed requests en masse or extracted training data, with Qwen alone accounting for more than 151 million exchanges. Actors also used Claude for missile software, autonomous kamikaze drones, and nationwide surveillance systems. The article How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data appeared first on The Decoder .
AWS Machine Learning Blog / 9:37 PM
Reduce inference cold starts on Amazon SageMaker HyperPod with model caching
Amazon SageMaker HyperPod now supports model caching for inference, which pre-loads model weights and container images onto cluster nodes so pods read from local NVMe storage instead of downloading over the network. Learn how model caching cuts cold starts from tens of minutes to seconds, how it works, and how to enable it.
TechCrunch AI / 8:57 PM
Anthropic details distillation campaigns from Alibaba, Moonshot AI, and DeepSeek
A new report released Thursday by Anthropic alleges persistent distillation attacks by China-based AI companies, which have escalated in recent months as competition in the space has intensified.
The Decoder / 12:40 PM
New Deepseek model V4.1-Flash cuts memory needs for AI agents
Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of its predecessor. On the DeepSWE coding benchmark, it narrowly beats Opus 5 and GPT-5.6 Sol, even though only 16 billion parameters are active per token. The model ships under the MIT license and targets much cheaper AI agents. The article New Deepseek model V4.1-Flash cuts memory needs for AI agents appeared first on The Decoder .
Bloomberg AI / 10:43 AM
DeepSeek’s New Low-Cost Model Deals a Fresh Blow to OpenAI, Z.ai
DeepSeek rolled out an AI model that charges as little as a fraction of a cent per million tokens, ramping up the pressure on rivals from Anthropic PBC to Z.AI Co.
Latent Space / 3:33 AM
[AINews] not much happened today
a quiet day
Ars Technica AI / 8:06 PM
Six Chinese AI firms accused of aggressively copying US frontier models
US urges AI firms to ID, then secretly switch, Chinese users to less-capable models.
Bloomberg AI / 1:50 AM
US Says Alibaba, DeepSeek Have ‘Systematically’ Siphoned AI Models
US security agencies accused China’s top AI companies including DeepSeek and Kimi maker Moonshot AI of systematically extracting proprietary knowledge from American firms and warned Silicon Valley developers to protect their work.
Simon Willison LLMs / 10:12 PM
Just a rumour of a bug is enough to find a security exploit these days
Just a rumour of a bug is enough to find a security exploit these days Anil Madhavapeddy is a professor of computer science at Cambridge and a core maintainer of the OCaml compiler. In this somewhat alarming post he reports that security issues in OCaml projects are seeing evidence of attempted exploits within minutes of patches being shared for discussion: This normally takes a few days and a release within a week or two is reasonable. Within about ten minutes (!) this website was fielding probes for percent-encoded traversal sequences, indicating that automated watchers are keeping an eye on public repositories. Modern coding agents have become so effective at finding flaws that the slightest hint at a new bug can be enough information for them to find it, something Anil has been able to demonstrate using his own agents, switching to DeepSeek V4 Pro when Claude Fable refused the task. Anil points out that this rate of discovery appears incompatible with existing open source embargo practices for new issues. If an issue can become an exploit this fast, we need to figure out new processes for keeping our communities safe. rclone maintainer Nick Craig-Wood confirms in the Hacker News comments that his project is seeing this problem: In the first 10 years of the rclone project we received about 20 security disclosures through GitHub. We had to deal with over 40 in the last month! That has taken a huge amount of my time, even using AI tools to triage and come up with fixes for review. The hit rate for those security disclosures is pretty good - about 75% of them have a nugget of something which needs looking at. [...] GitHub assigns CVEs for the advisories. Before the AI apocalypse they took 2-3 days for an assignment but now it they are running at 3-4 weeks so I have to send the point releases out with CVE-PENDING in the changelog which isn't ideal. Via Hacker News Tags: open-source , security , ai , generative-ai , llms , coding-agents , ocaml , ai-security-research
AWS Machine Learning Blog / 4:24 PM
Preparing data for supervised fine-tuning Part 2: Advanced data strategies
The advanced side of supervised fine-tuning data prep. This second post in a two-part series covers evaluating data readiness with learning curves, selecting high-value data subsets, augmenting data with synthetic and distilled examples, and mixing data sources to prevent catastrophic forgetting.
Simon Willison LLMs / 11:58 PM
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index
Qwen 3.8 27B scores 52 on the Artificial Analysis Intelligence Index That's the same score as GPT-5.6 Luna (max), and just one point behind GLM-5.2 (max) and DeepSeek V4 Pro 0813 (max) - that GLM is 753B and that DeepSeek is 1.7T parameters , and Luna is size unknown but presumably a whole lot bigger than 27B. Qwen 3.8 27B is a truly astonishing model . Via Hacker News Tags: ai , generative-ai , llms , qwen , ai-in-china , artificial-analysis
Product Hunt AI / 1:08 PM
DeepSeek Harness
Composable agent harness where everything is a plugin Discussion | Link
Hacker News AI / 6:48 AM
DeepSeek Flash v4.1
HN 2 pts · 0 comments
Hacker News AI / 6:41 AM
DeepSeek-v4.1-Flash: Pushing the Limits of KV Cache Compression [pdf]
HN 2 pts · 0 comments
Hacker News AI / 6:13 AM
(Tech Report) DeepSeek-v4.1-Flash: Pushing Limits of KV Cache Compression [pdf]
HN 5 pts · 0 comments
Hacker News AI / 2:12 AM
Busabase for DeepSeek Harness: An Agent database that runs apps and skills
HN 4 pts · 0 comments
Latent Space / 7:46 AM
[AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokens
Queue the usual rush of model launches...
Hacker News AI / 11:02 AM
Show HN: DeepSeekGUI – A Windows desktop client for DeepSeek's coding agent
HN 2 pts · 0 comments
Latent Space / 1:50 AM
[AINews] NVIDIA buys HuggingFace for $13B, as OpenAI publishes their HF incident retro
Open Source wins!
The Decoder / 2:40 PM
Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency"
Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activates just 6 out of 125 billion parameters per token. At one-ninth the training cost, it beats much larger competitors like DeepSeek-V4-Flash and Claude Opus 4.6 on coding and office benchmarks, adding more pricing pressure on OpenAI and Anthropic. The article Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency" appeared first on The Decoder .
The Decoder / 8:40 AM
Taiwanese cybersecurity firm warns that AI tools have more than doubled Chinese state-backed cyberattacks
Chinese state-backed hacking groups have more than doubled their attacks since they started using AI models like DeepSeek to write exploit code and scan networks, according to Taiwanese security firm TeamT5. Hackers also used ChatGPT and Anthropic's Claude Code. A UK study shows the cyber capabilities of open models are catching up fast. The article Taiwanese cybersecurity firm warns that AI tools have more than doubled Chinese state-backed cyberattacks appeared first on The Decoder .
Hacker News AI / 2:55 PM
Gemini 3.7 Flash, Grok 4.6, GLM-5.3 and DeepSeek V4 Pro joined the frontier
HN 1 pts · 0 comments
Latent Space / 7:36 AM
[AINews] 10% worse, 100x cheaper, 10000x faster: Why Simulation is taking over
Did you think RSI stopped at model training?
Latent Space / 5:17 AM
[AINews] Death of Params: Z.ai CEO Jie Tang on GLM 5.3 and the new Post-training Scaling Law
Every lab CEO is on X now
AWS Machine Learning Blog / 1:42 PM
Tiered KV cache for large LLMs on Amazon SageMaker HyperPod with Curvine
Running large language model inference at scale forces a KV cache trade-off: oversized GPU instances or slow time-to-first-token. This post builds a tiered KV cache on Amazon SageMaker HyperPod that extends the cache into a shared, distributed NVMe pool with Curvine, so replicas reuse cache at near-local-disk speeds on cost-efficient instances.
Latest story in this edition: 5:59 AM
Back to front page