← All Posts

Qualcomm Claims Single-Core Leadership for Its First Server CPU, the Dragonfly C1000, Delivering 250+ Cores & 5 GHz By 2028

The Hot Take: Interesting...

Qualcomm has introduced its first-ever CPU designed for Data Centers, the Dragonfly C1000, which leverages the Oryon architecture. Qualcomm Enters The Agentic AI CPU Race With Dragonfly C1000 Chip, Oryon-Based With Over 5 GHz Clocks, Over 250 Cores, & Aims To Achieve Single-Core Leadership One of the biggest announcements by Qualcomm today was its first release of a CPU for the data center segment, called the Dragonfly C1000. This is a chip purpose-built for Agentic AI & General-Purpose workloads, delivering best-in-class power…

Read the full article

Tensordyne's 3nm Napier AI Chip Promises 13x Higher Token Throughput Than Blackwell & Blazes Past Rubin With 1000 Tokens/s In Multi-Trillion Parameter Models

The Hot Take: I really hope something comes soon to alleviate all this nonsense AI is causing.

US-based AI company, Tensordyne, has announced the successful tape-out of its Napier chip, which it claims to demolish NVIDIA's Blackwell & Rubin chips with leading token throughput and efficiency. Tensordyne’s new Napier AI Chip arrives with one clear mission: to make NVIDIA’s Blackwell and Rubin chips look considerably less impressive The Napier chip will be the core component of the Tensordyne Napier TDN system, which is designed in collaboration with Broadcom and HPE Juniper Networks. The Napier platform has one goal: to unify AI…

Read the full article

AWS Graviton5 Debuts with 192 Arm Cores and PCIe 6.0

The Hot Take: ARM seems to be breaking out from everywhere. Fujitsu, Nvidia, AWS and ARM. Qualcomm seems to be playing catch up in the server market from the looks of it.

AWS has provided a first look at its next-generation Graviton5 processor, a custom server CPU developed by Annapurna Labs for deployment across the company's cloud computing platform and AI inference infrastructure.

Read the full article

Microsoft is killing the Copilot+ PC advantage, brings Windows 11’s local AI to RTX 30+ PCs with 6GB vRAM

The Hot Take: Now we know why M$ is trying to squeeze out every ounce of performance in Windows 11.....

Microsoft says you’ll be able to run Windows 11’s local Language Model APIs on non-Copilot+ PCs as long as you meet the new hardware requirement: an RTX 30+ GPU with 6GB of VRAM. It’s a major change, as it means Copilot+ PCs’ advantages are getting “thin,” and I wouldn’t be surprised if Microsoft drops the NPU requirement entirely in the future. Copilot+ PCs officially debuted on June 18, 2024, and they’ve been driving sales for PC makers. However, it’s not because of the “Copilot” or “NPU” factor. It’s largely because newer PCs are now sold…

Read the full article

Chinese military has been acquiring Nvidia chips, even post-Washington export controls, research claims — multiple institutions linked to the PLA asked for Nvidia AI chips, according to publicly available documents

The Hot Take: Tell me something I didn't know already. Why else would the GPU market go crazy prices wise?

A business-intelligence researcher said that the Chinese military has been actively acquiring Nvidia AI chips, even after the U.S. put export controls on them. Public documents show that some institutions ask for these chips either through the specifications they demand or by directly asking for Nvidia chips by name.

Read the full article

Intel details long-awaited Crescent Island AI GPU at Computex, boasts up to 480 GB of LPDDR5X to combat memory shortages — company shares more details of its Xe3P inference accelerator at Computex

The Hot Take: Intel moving fast to make up lost ground on this front for sure. From the looks trying to hit the $ sweet spot too.

Intel revealed more details of its next-gen Data Center GPU, code-named Crescent Island, at Computex 2026. This inference-optimized chip will feature up to 480GB of LPDDR5X memory for efficient handling of massive AI contexts.

Read the full article

NVIDIA Loses Ground With AI Engineers as Cooling and Power Costs Push Hyperscalers Toward Custom ASICs, Evercore Warns

The Hot Take: When these start getting traction we'll get GPUs to drop in price.....

While AI GPU giant NVIDIA's chips are widely believed to offer superior total cost of ownership (TCO) compared to custom AI chip alternatives, analysts from Evercore ISI believe that AI engineers are unimpressed by them. NVIDIA CEO Jensen Huang has defended his firm's AI chip price points on multiple occasions by claiming that they offer better performance efficiency compared to peers. However, according to the Evercore report, AI engineers are also focused on other metrics, such as the cost of cooling the chips, when deciding which products…

Read the full article

'Changing of the Guard'? AMD, Intel, and Micron Soar While Nvidia Lags

The Hot Take: AMD seems to be out performing Intel & Nvidia on the market, while Nvidia is still the preferred Ai holy-grail? Just seems odd.

While Nvidia has dominated the "infrastructure boom" since 2022's launch of ChatGPT and "the generative AI craze," CNBC writes that "This week offered the starkest illustration yet of what MIzuho analyst Jordan Klein said could be a 'changing of the guard in AI.'" Chipmakers Advanced Micro Devices and Intel notched gains of about 25%, while memory maker Micron jumped more than 37% and fiber-optic cable maker Corning climbed about 18%. All four of those companies have more than doubled in value this year, with Intel leading the way, up…

Read the full article

Claude hitches ride on SpaceX's datacenter capacity

The Hot Take: Ai usage growing pretty steady and fast it would appear.

Anthropic is partnering with SpaceX to ease capacity constraints that have stranded Claude customers, a gesture that may soothe developer discontent about service availability and cost. Ami Vora, chief product officer at Anthropic, announced the expanded rate limits during Code for Claude, a developer event livestreamed from San Francisco. "As of today, we are increasing rate limits for developers on Claude Code and the Claude Platform," said Vora. "More specifically, we are doubling Claude Code's five-hour rate limits for Pro, Max, Team, and…

Read the full article