/

Chips & compute

194 stories

Run Ray on TPU, Part 1: The foundations

Google Developers BlogLabsRun Ray on TPU, Part 1: The foundations Ray 2.55 introduces official, first-class support for Google Cloud TPUs, enabling developers to run distributed Python workloads on Google's accelerators using the familiar Ray task-and-actor APIs. To handle the strict n

September 6
Qiita (LLM)

Qiita (LLM)JapanVRAMは束ねない。NVIDIA PAIRが家庭内PCを推論クラスタに変える仕組み ローカルLLMを動かすPCを増やしても、1本の推論がその台数分だけ速くなるわけではない。 NVIDIA PAIRが減らすのは、複数の独立したリクエストが1台のGPUの前で並ぶ待ち時間だ。 公式デモでは処理時間が18分から8分48秒へ短縮されたが、その数字を正しく読むには「...

September 6
b10819

llama.cpp releasesToolsb10819 metal : fix memory leak in early return ( #28399 ) Website: https://llama.app Attestations: https://github.com/ggml-org/llama.cpp/attestations/45438612 macOS/iOS: macOS Apple Silicon (arm64) macOS Apple Silicon (arm64, K

September 5
Building a Memory-Driven Agent with NVIDIA NemoClaw

NVIDIA Technical BlogSiliconBuilding a Memory-Driven Agent with NVIDIA NemoClaw Enterprise work spans messages, decisions, projects, and obligations that change over time. An AI agent that starts without this context must reconstruct it...

September 4
PyTorch Blog

PyTorch BlogToolsYour Guide to Hardware Acceleration & Compute Infrastructure at PyTorch Conference North America 2026 TL:DR PyTorch Conference North America 2026 (San Jose, October 20–21) is packed with sessions on getting PyTorch to run fast, portably, and reliably across an increasingly diverse silicon landscape –...

September 4
Hacker News Show

Hacker News ShowGitHubShow HN: Open-Source eInk Bike Computer

September 4
Frontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson

NVIDIA Technical BlogSiliconFrontier Reasoning Reaches the Edge: How to Deploy and Optimize Models on NVIDIA Jetson Running reasoning and agentic AI at the edge has been harder than it needs to be. Until recently, models capable of multi-step reasoning were too large to run...

September 4
Build a Physical AI model factory with NVIDIA Cosmos 3 on SageMaker HyperPod

AWS Machine LearningLabsBuild a Physical AI model factory with NVIDIA Cosmos 3 on SageMaker HyperPod Building a Physical AI system takes a continuous pipeline, not a single training job. This post shows how to run that model factory (synthetic data generation, post-training, and closed-loop evaluation with NVIDIA Cosmos

September 4
Who Cares if AI Is Conscious—It’s Basically Alive

Wired AIValleyWho Cares if AI Is Conscious—It’s Basically Alive While philosophers ponder AI consciousness, the models have ideas of their own.

September 4
Guangdong’s MLCC ‘Three Musketeers’ Ride the AI Compute Wave to Stronger Earnings

PandailyAsiaGuangdong’s MLCC ‘Three Musketeers’ Ride the AI Compute Wave to Stronger Earnings Chaozhou Three-circle (Group) posted surging first-half earnings as AI compute drives demand for MLCCs, alongside peers Fenghua and Viiyong in Guangdong’s “three musketeers” component cluster.

September 4
Sugon Unveils the World’s First 64-Thread Mobile Workstation With an Unnamed Domestic x86 Chip

PandailyAsiaSugon Unveils the World’s First 64-Thread Mobile Workstation With an Unnamed Domestic x86 Chip Sugon has launched the N50 Pro, billed as the world’s first 64-thread mobile workstation, powered by an unnamed domestic x86 chip widely expected to be Hygon.

September 4
Nvidia Expands Its Open Source AI Presence With $12.9 Billion Hugging Face Buy

The Next PlatformSiliconNvidia Expands Its Open Source AI Presence With $12.9 Billion Hugging Face Buy

September 3
Nvidia acquires Hugging Face for $12.93 billion — company gains control of major AI model distribution platform

Tom's Hardware AISiliconNvidia acquires Hugging Face for $12.93 billion — company gains control of major AI model distribution platform Nvidia expands beyond AI hardware with its $12.93 billion acquisition of Hugging Face, gains control of a major open AI model platform, vows to preserve its support for competing models, clouds, and hardware platforms.

September 3
Sparks Fly: NVIDIA Accelerates Local AI at IFA 2026

NVIDIA BlogSiliconSparks Fly: NVIDIA Accelerates Local AI at IFA 2026 Frontier intelligence is going local. At IFA 2026, NVIDIA, Microsoft and its partners are teaming up to provide faster inference and new tools that make agents easier to set up and run locally on NVIDIA hardware. New com

September 3
Nvidia PAIR utility joins every GPU in your home into a cluster for agentic AI tasks — tool uses spare cycles to keep agent swarms from hammering one GPU

Tom's Hardware AISiliconNvidia PAIR utility joins every GPU in your home into a cluster for agentic AI tasks — tool uses spare cycles to keep agent swarms from hammering one GPU Nvidia's Personal AI Router (PAIR) clustering utility lets agentic AI workloads take advantage of every spare GPU cycle on a home network, potentially making for faster execution and more private inference.

September 3
‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW

NVIDIA BlogSilicon‘NBA 2K27’ With NVIDIA DLSS 5 Leads 28 New Games Coming to GeForce NOW September is here with 28 more games streaming on GeForce NOW this month, led by a slam dunk: NBA 2K27 with the NVIDIA DLSS 5 3D-Guided Neural Rendering feature. Through NVIDIA’s close collaboration with Visual Concepts

September 3
NVIDIA to Acquire Hugging Face

NVIDIA BlogSiliconNVIDIA to Acquire Hugging Face I’m excited to announce that NVIDIA has agreed to acquire Hugging Face for $12,930,300,000. Together, we will scale Hugging Face’s platform, strengthen its infrastructure and expand access to AI for developers and instit

September 3
Software Engineering Daily

Software Engineering DailyVoicesMoving Beyond RAG with Precomputed Context Retrieval has become one of the central problems in building useful AI systems. The standard approach to grounding a model in one’s own data has been retrieval augmented generation, or RAG, where an agent searches a vect

September 3
Hugging Face Blog

Hugging Face BlogLabsFine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps

September 3
Optics Still Driving Marvell’s AI Business More Than Custom Chips

The Next PlatformSiliconOptics Still Driving Marvell’s AI Business More Than Custom Chips

September 2
Hot Chips 2026: The CPU’s next chapter is being built on Arm

Arm NewsroomSiliconHot Chips 2026: The CPU’s next chapter is being built on Arm From IBM Z and FUJITSU-MONAKA to NVIDIA Vera and Arm AGI CPU, Hot Chips showed why a common Arm software foundation is becoming a strategic advantage

September 2
Taiwan’s six-year hunt for China’s undercover chip labs

Rest of WorldTaiwan’s six-year hunt for China’s undercover chip labs Previously unpublished government data reveals the scale of Taiwan’s crackdown on Chinese companies accused of hiding their ties while recruiting chip talent and pursuing sensitive technology.

September 2
Cash In on the AI Boom by Renting Out Your Spare Compute

IEEE Spectrum AISiliconCash In on the AI Boom by Renting Out Your Spare Compute If you own an at-home server, a gaming computer, or just a laptop that doesn’t get much love, listen up. You can now put that spare computing power to use and earn some passive income in the process. AI companies are hun

September 1
📈 Data to start your week

Exponential ViewVoices📈 Data to start your week Nvidia doubles again; AI’s cover-up habit; the $1.50 data centre bill++

September 1