Stories
30
Sources
8
Topics
11
For You lens
27 stories in this edition match your reader profile.
Reader signals
3
Searches
0
Matches
27
Top score
120
Edition Index
Topic, entity, and source map
Topics
Entities
Lead Story
Nvidia’s Growing Dependence On a Few Big Customers
Nvidia CEO Jensen Huang has very good reason to try to develop new customers by investing in a variety of neocloud and AI firms. For all its success, Nvidia’s sales have become increasingly dependent on a handful of customers—some of which may cut back their purchases over time. That becomes clear when you look at its own disclosures going back several years, which specify how many customers account for more than 10% of sales. In the first half of this fiscal year, ending in July, three such customers were responsible for 44% of total sales. Last fiscal year, two customers accounted for 36% of total sales. Going back to fiscal 2023, Nvidia had no customers accounting for 10% or more of sales (see the chart below).
The Information AI / 1:04 AM
Nvidia May Invest Up to $10 Billion In Anthropic’s IPO
Nvidia has discussed investing in Anthropic’s upcoming initial public offering that would raise as much as $100 billion at a valuation of around $2 trillion, Reuters reported Friday. Nvidia could invest as much as $10 billion in Anthropic at the IPO price, according to the report. Anthropic last ...
Latent Space / 5:56 AM
[AINews] DeepSeek v4.1-Flash: 763B-P8B-D16B novel causal Encoder–Decoder architecture with vision marks the Return of the Whale
We agree with Sebastian: this should have been DeepSeek v5
AWS Machine Learning Blog / 9:37 PM
Reduce inference cold starts on Amazon SageMaker HyperPod with model caching
Amazon SageMaker HyperPod now supports model caching for inference, which pre-loads model weights and container images onto cluster nodes so pods read from local NVMe storage instead of downloading over the network. Learn how model caching cuts cold starts from tens of minutes to seconds, how it works, and how to enable it.
Latent Space / 3:33 AM
[AINews] not much happened today
a quiet day
AWS Machine Learning Blog / 10:26 PM
Deploying Qwen3.8-2.4T-A95B on Amazon SageMaker HyperPod with vLLM
Learn how to deploy Qwen3.8-2.4T-A95B, a 2.4-trillion-parameter open-weight model, on Amazon SageMaker HyperPod with vLLM. This walkthrough covers cluster provisioning, NVFP4 quantization, and an OpenAI-compatible endpoint with built-in reasoning, tool calling, and native MTP speculative decoding.
The Verge AI / 4:00 PM
Nvidia launches free tool that links idle computers into a personal AI data center
Nvidia is announcing its new Personal AI Router (PAIR), a free tool that syncs up your home computers for tackling local AI inference tasks with tools like Ollama and LM Studio. Let's get the obvious thing out of the way, despite what its name might imply: PAIR is not a hardware router. It's open-source software […]
The Information AI / 4:10 PM
Why AI Companies Are Building Out Wall Street-Style Finance Teams
The financing boom for the AI build-out is getting bigger and more complicated by the day—and AI companies have been staffing up for the challenge. AI labs including OpenAI and Anthropic, as well as neoclouds such as Nscale, are among a growing number of AI companies building out their capital markets teams and hiring specialists in areas like structured finance. That in part reflects the sheer volume of deals these companies are doing, many of which don’t fit neatly into standard corporate debt. This in-house staff can help when it comes to negotiating with lenders and drilling down into construction, power and other key details. Of course, tech and data center companies have long had in-house teams to handle fundraising, deals and other corporate finance needs. And structured finance is nothing new to the infrastructure world. But the scale of the AI build-out, which bankers peg at around $7.5 trillion in spending over the next five years, has pulled relatively young labs and upstart cloud firms into financing arrangements that are new territory. That means finance professionals, from bankers to investors at private equity, private credit and infrastructure firms, have more options in the form of neoclouds and other AI infrastructure startups, some of which are offering significant pre–initial public offering equity. “It's a new avenue for these people,” said James Howl-Newton, founder of Futura Search Partners, a specialist search firm focused on areas including digital infrastructure finance. As a result, “sponsors are having to deal with additional routes to exits for top performers,” he said. AI companies and infrastructure providers are tapping financing frequently and across different instruments, requiring deeper in-house capabilities and expertise than young tech firms have typically needed. One executive overseeing finance hiring at a neocloud noted that leveraged and structured finance backgrounds bring expertise that can help in areas like working through project diligence and getting banks to sign off on deals. Some AI firms may also want to run their own project finance models so they can move quickly through negotiations and have something to compare to lenders’ models. AI companies aren’t always issuing the debt themselves—that can fall to data center developers or special purpose vehicles, with firms like Blackstone and Apollo providing or arranging chip and other financing. And some of the biggest AI deals are using backstops from investment-grade companies like Nvidia or major cloud providers. Even so, commitments from AI customers often underpin much of the borrowing. And the users of the infrastructure will want to understand what they’re signing up for and their risks if a project runs into trouble. “Hiring of people within that business, responsible for the financing of compute, could prove to be an existential decision,” said Dan McCarthy, founder and CEO of One Search, an executive search firm focused on infrastructure finance whose recent clients include OpenAI. “You want someone who knows where all the pitfalls are, where all the bodies are buried in multibillion-dollar loans.” OpenAI, for its part, in July named Sven Semmelmann as head of compute capital markets. He previously led structured finance at Generate Capital, an investment firm that finances and owns infrastructure projects, and he has also held project finance roles at major banks. OpenAI Chief Financial Officer Sarah Friar, when announcing the hire on LinkedIn, said Semmelmann would oversee financing and partnerships to grow the company’s compute resources. Anthropic, meanwhile, has made several finance hires recently to work on capital markets and compute deals, and also has open positions posted including a capital markets infrastructure financing role. AI infrastructure upstarts are staffing up as well. Nscale, which launched in 2024 and is gearing up for a potential IPO , has been hiring across levels for capital markets and treasury as well as legal roles, calling for experience in areas like structured finance and private credit. SB Energy and Crusoe, which are developing major new data centers for OpenAI and other customers, are hiring across levels for jobs focused on project financings and other structured deals, recent postings show, while AI infrastructure startup Fluidstack is hiring a structured finance lead and a more junior counterpart. The good news for AI companies is that private credit and infrastructure teams, as well as investment banking teams focused on structured or project finance, had been growing even prior to the AI boom, providing a pool of skills that could translate into new twists on structured finance, like big graphics processing unit–backed deals. But that kind of finance talent doesn’t come cheap, especially for more senior people who have a track record of working on large transactions. And the normal tech tactic of dangling stock to lure talent won’t necessarily do the trick in all cases, especially for the most seasoned dealmakers and investors. Financiers would have to weigh a cash-heavy Wall Street pay package, albeit one that can depend heavily on how good bonus season is, against betting a portion of their pay on stock in a private or newly public company. Managing directors in investment banking can make north of $1 million in cash a year, with the biggest rainmakers making considerably more. The part of pay they get in stock at big public banks may vest over a few years but is generally easy to sell after that. For people at big infrastructure or private credit firms, senior employees may also receive carried interest, meaning a share of the profits on the funds or investments they work on, which can become worth millions over time. For instance, an investor at a top infrastructure firm may have several million dollars’ worth of carried interest tied up at their current firm they’d have to leave on the table. An AI company could try to make them whole with stock, which could be tantalizing to some, though others might not want to make a bet on equity in a young company. That might make the most experienced investors—those who’ve seen big infrastructure projects through over many years and know all the tricks of the trade—hard to pry away. New From Our Reporters Exclusive Anthropic’s In-House Payments Tech Push Could Chip Away at Stripe By Stephanie Palazzolo Exclusive China Curbs Humanoid IPOs After Unitree’s Volatile Debut By Jing Yang and Qianer Liu
The Information AI / 1:31 PM
The DOJ Is Investigating Nvidia’s Licensing Deal With Chip Startup Groq
The Department of Justice is looking into whether Nvidia tried to avoid antitrust scrutiny in its $20 billion deal to license the technology of chip startup Groq and hire most of its employees in December, the New York Times reported Wednesday. The two companies described the deal as a “ ...
The Information AI / 1:01 PM
Nvidia Deepens Partnership With AI Chip Startup D-Matrix
Chip startup d-Matrix said Thursday it plans to use Nvidia’s networking hardware to connect its chips for running AI models to each other and to Nvidia’s Vera central processing units, which are used in data centers. The startup will use Nvidia’s NVLink Fusion products, including switches, data ...
Hacker News AI / 7:23 AM
Nvidia Personal-AI-Router
HN 1 pts · 0 comments
AWS Machine Learning Blog / 3:51 PM
Simplify and support your TorchServe workloads using Ray Serve Deep Learning Containers
TorchServe is no longer maintained, leaving teams to own the entire GPU inference stack. The AWS Ray Serve Deep Learning Container is a supported, pre-tested container with the framework, GPU drivers, and serving layer already assembled. This post walks through deploying a vision-language model on Amazon EKS using the Ray Serve DLC on a single GPU node.
AWS Machine Learning Blog / 7:12 PM
Pathway’s brain-inspired architecture development on Amazon SageMaker HyperPod
Pathway's Baby Dragon Hatchling (BDH) is a brain-inspired, post-transformer architecture that reasons in latent space instead of emitting chain-of-thought tokens. See how Pathway develops and scales BDH on Amazon SageMaker HyperPod, and how BDH-CQ set a new cost-efficiency mark on the ARC-AGI-1 benchmark.
AWS Machine Learning Blog / 4:21 PM
Benchmarking small LLM inference on SageMaker AI: G7 vs G5 and G6
Benchmark two 30B Mixture-of-Experts models, Qwen3-Coder-30B and NVIDIA Nemotron-3-Nano-30B, across G5, G6, G6e, and G7 GPU instances on Amazon SageMaker AI. Compare throughput, latency, and cost-per-token, and see how G7's NVIDIA Blackwell GPUs deliver measurable price-performance gains for real-time LLM inference.
Hacker News AI / 12:25 PM
Show HN: CUDA/graphics in QEMU-KVM VMs without passing the Nvidia card to them
HN 4 pts · 2 comments
Hacker News AI / 10:14 PM
Nvidia Lambda Circular Deal [video]
HN 1 pts · 0 comments
Latent Space / 4:32 AM
[AINews] Collusion.wiki: A second undisclosed OpenAI agent swarm incident...
AI News for 9/2/2026-9/3/2026.
TechCrunch AI / 5:18 PM
What will Apple’s John Ternus era look like?
It’s officially the Ternus era at Apple. Tim Cook stepped down as CEO this week, handing the company to former hardware chief John Ternus, whose first memo promised a “huge launch next week” — timing that puts Apple’s next iPhone event on his desk before he’s even settled in. Cook isn’t going far, though: he’s staying on as Executive Chairman, focused on the kind of policy […]
The Decoder / 2:25 PM
Nvidia buys the front door to open AI as closed labs increasingly design their own silicon
Nvidia plans to acquire Hugging Face for about $12.9 billion, securing the central platform for open AI models. More than 18 million developers and 200,000 companies use the hub. CEO Jensen Huang promises to keep the platform open and hardware-neutral, but the deal also hands him a powerful distribution channel for compute. The article Nvidia buys the front door to open AI as closed labs increasingly design their own silicon appeared first on The Decoder .
TechCrunch AI / 12:42 PM
Nvidia confirms it will buy Hugging Face for $12.9 billion
Nvidia said Hugging Face hosts over 3 million models and is used by over 18 million developers.
The Verge AI / 12:12 PM
Nvidia is buying Hugging Face for almost $13 billion
Nvidia has agreed to buy Hugging Face for $12.93 billion, bringing one of the most popular hosting platforms for open-source AI models, datasets, and tools under the ownership of the world's biggest AI chipmaker. Hugging Face is an online platform founded in 2016 that gives AI developers a space to share their projects and data […]
Bloomberg AI / 6:37 PM
Anthropic’s Compute Bet, Musk on AI, Apple’s New CEO | Bloomberg Tech 9/01/2026
Bloomberg’s Ed Ludlow breaks down Anthropic's $35 billion computing deal with Nvidia-backed cloud provider Lambda as the AI startup looks to expand its AI capacity. Plus, Elon Musk predicts AI will increase the global economy by up to 30%. And, John Ternus officially takes over as Apple CEO as the company gets ready for a new product unveil next week. (Source: Bloomberg)
AWS Machine Learning Blog / 4:17 PM
From theory to delivery: How Atos upskilled 400 engineers in agentic AI
When Atos set out to upskill 400 engineers in agentic AI, hands-on learning was the missing ingredient. Over three days, engineers built multi-agent systems on AWS through an AI League event. This post explains why Atos chose the format, what engineers built and learned, and what other enterprises should consider.
Bloomberg AI / 11:19 PM
Nvidia Mulls $10 Billion Anthropic IPO Backing, Reuters Says
Nvidia Corp. is considering investing as much as $10 billion in Anthropic PBC’s initial public offering, which could become the biggest IPO of all time, Reuters reported.
TechCrunch AI / 9:51 PM
Jensen Huang explains why Nvidia will grow an astounding 70% next year
Nvidia has its finger in every pie, and sees another year of plenty in its future, Jensen Huang says. But, he insists, its deals are not circular.
The Verge AI / 3:38 PM
Universal Music is launching an AI music platform with ElevenLabs
Universal Music Group is launching a new AI-powered platform that will allow users to draw from its catalog of licensed music to create song remixes, mashups, and new takes on tracks, according to an announcement on Thursday. The record label is developing the platform through a multiyear licensing agreement with ElevenLabs, a company that specializes […]
The Decoder / 12:23 PM
Nvidia and Palantir team up to run supply chains with AI, starting with Nvidia's own million-part operation
Nvidia and Palantir want to run supply chains with AI. The article Nvidia and Palantir team up to run supply chains with AI, starting with Nvidia's own million-part operation appeared first on The Decoder .
Bloomberg AI / 9:03 AM
Nvidia-Backed MediaTek’s Sales Soar 44% With AI Chip Momentum
MediaTek Inc. reported a 44% surge in monthly sales, boosted by its emerging AI chip business from the likes of Alphabet Inc.’s Google.
Latest story in this edition: 3:00 PM
Back to front page