Nvidia

22 pieces on Nvidia.

News & Analysis

news
Shan2026-08-31
NvidiaMediaTekAI ChipsData CentersNVLink

Nvidia's $3.5B MediaTek Bet Makes Custom Chips Depend on Nvidia Racks

Nvidia invests $3.5B in MediaTek so hyperscaler custom ASICs plug into NVLink Fusion — escaping the GPU but not the platform.

Read more
news
Shan2026-08-31
NvidiaGPUAI InfrastructureData CentersOpen Weights

Nvidia's Vera Rubin Delivers 3x Storage Gains Beyond the GPU

Nvidia's Vera Rubin stack delivers up to 3x storage operation gains via the Vera CPU — shifting its moat from GPU silicon to data orchestration.

Read more
news
Shan2026-08-29
NVIDIAWeather ForecastingEarth2StudioOpen SourceTutorials

NVIDIA Earth2Studio: Custom Batched Ensemble Forecasting Pipeline

Build a full ensemble weather pipeline in Earth2Studio using low-level iterators, custom diagnostics, Zarr I/O, and fair CRPS verification — all in one Colab session.

Read more
news
Shan2026-08-27
NvidiaAmazon Web ServicesData CenterGPUInfrastructure

Amazon Adds 2M More Nvidia GPUs After 5-Month Demand Surge

AWS triples its Nvidia GPU commitment to 3M+ chips spanning Blackwell Ultra, Rubin, and Rubin Ultra, with delivery across 2027–2028.

Read more
news
Shan2026-08-27
NVIDIAAmazon EC2Inference OptimizationASRCUDA MPS

NVIDIA MPS Cuts ASR GPU Count from 16 to 4 on EC2

NVIDIA CUDA MPS on EC2 g7e.4xlarge delivers 92.1 RPS at 352 ms mean latency, reducing Heidi Health's GPU fleet 75% while holding sub-second SLAs.

Read more
news
Shan2026-08-26
AnthropicNvidiaAI InfrastructureComputeOpen Weights

Anthropic Signs $45B Nscale Deal for Vera Rubin Compute by 2027

Anthropic's $45B, six-year Nscale deal adds Vera Rubin capacity from a West Virginia data center, extending a compute spree that now spans six partners.

Read more
news
Shan2026-08-26
OpenAICustom SiliconAI InferenceNvidiaBenchmarks

OpenAI's Jalapeño ASIC Beats Nvidia GB200/GB300 on Latency and Throughput

OpenAI's Jalapeño chip delivers 1.5–1.9× more AI work per watt and 1.7–3.6× lower latency than Nvidia GB200/GB300 across three models.

Read more
news
Shan2026-08-25
PerplexityNVIDIAAgentic AILocal InferenceOpen Weights

Perplexity Portable Computer: Full Agent Harness on DGX Spark, $0 Local Steps

Perplexity ships a full agentic harness on NVIDIA DGX Spark with zero per-token cost for local steps and OS-enforced sandboxing.

Read more
news
Shan2026-08-24
NvidiaData CentersAI InfrastructureInvestment

Nvidia Takes Minority Stake in Data Center Site Developer Cloverleaf

Nvidia acquires a minority equity stake in Cloverleaf Infrastructure, a 2024-founded power-site middleman, in a deal likely worth several hundred million dollars.

Read more
news
Shan2026-08-22
NvidiaAgentic AIBenchmarksOpen WeightsLLM Infrastructure

Nvidia's AVO Harness Takes Claude Opus 5 from 30% to 100% on ARC-AGI-3

Nvidia's custom AVO harness lifted Claude Opus 5 from 30% to 100% on ARC-AGI-3 — without changing the model at all.

Read more
news
Shan2026-08-19
AI HardwareFundraisingInferenceStartupsNVIDIA

Etched Raises $700M at $21B Valuation After Jane Street Deploys Its Hardware

Etched's valuation doubled from $10.3B to $21B in a month after quant firm Jane Street tested its inference chips and deployed a rack in its own datacenter.

Read more
news
Shan2026-08-18
NvidiaData CentersOpenAISoftBankAI Infrastructure

Nvidia Invests $1.5B in SB Energy, Locks In Sole GPU Supply at Ports-Pike

Nvidia's $1.5B equity stake in SB Energy makes it the sole compute supplier at the OpenAI-linked Ports-Pike data center, backed by up to $105B in credit.

Read more
news
Shan2026-08-18
NVIDIATensorRTInferenceOpen SourceEdge AI

NVIDIA TRTMC: Hugging Face to C++ TensorRT in Two Commands, No ONNX

NVIDIA's TensorRT Model Connect converts supported checkpoints to native C++ inference in two CLI commands, no ONNX export, across 76 model families.

Read more
news
Shan2026-08-17
GroqNeocloudNvidiaInferenceFundraisingData Centers

Groq Raises $350M at $3.5B Valuation in Neocloud Pivot

Groq closes $350M at $3.5B — down from $6.9B — as it abandons custom LPU chips and operates Nvidia GPUs across 13 data centers.

Read more
NVIDIA Nemotron 3.5 Lightning: 30B Parameters, 3B Active
news
Shan2026-08-16
NVIDIAAI AgentsOpen WeightsLarge Language ModelsInference

NVIDIA Nemotron 3.5 Lightning: 30B Parameters, 3B Active

NVIDIA's Nemotron 3.5 Lightning activates only 3B of 30B parameters per token, targeting the execution layer of multi-model agent stacks.

Read more
NVIDIA Nemotron 3.5 Lightning: 30B MoE with 3B Active Parameters
news
Shan2026-08-12
NVIDIAOpen WeightsMixture of ExpertsAI AgentsModel Routing

NVIDIA Nemotron 3.5 Lightning: 30B MoE with 3B Active Parameters

NVIDIA ships Nemotron 3.5 Lightning, a 30B MoE with 3B active parameters and 1M-token context, plus NeMo Switchyard for per-step agent routing.

Read more
NVIDIA VoiceChat 11B: Open Full-Duplex Speech Model with 448 ms Latency
news
Shan2026-08-10
NVIDIASpeech ModelsOpen WeightsTool CallingFull-Duplex

NVIDIA VoiceChat 11B: Open Full-Duplex Speech Model with 448 ms Latency

NVIDIA's NemotronLabs VoiceChat 11B unifies ASR, LLM, and TTS into one 11B model with 448 ms turn-taking and live tool calling.

Read more
news
Shan2026-08-09
AI AgentsNVIDIAPythonOpen SourceBenchmarksLLM Infrastructure

NOOA: NVIDIA's Object-Oriented Agent Framework Explained

NVIDIA open-sources NOOA, a Python framework that collapses prompt templates, tool schemas, and workflow graphs into one class—with 82.2% on SWE-bench Verified.

Read more
news
Shan2026-05-22
NvidiaHardwareChipsChinaHuawei

Nvidia 'Largely Concedes' China AI Chip Market to Huawei Following Export Restrictions

CEO Jensen Huang admitted Nvidia has effectively exited the Chinese market due to tightening US export controls, allowing Huawei to solidify its dominance.

Read more
news
Shan2026-05-22
NvidiaHardwareEarningsFinanceSemiconductors

Nvidia Posts Record $58.3 Billion Profit, Rides Data Center Boom Driven by Agentic AI

Driven by the surge in agentic AI capabilities, Nvidia announced staggering quarterly growth, hauling in $81.6 billion in revenue and executing a massive stock buyback.

Read more
articles
Shan2026-05-12
NvidiaHardwareChipsAI InfrastructureSupply Chain

Nvidia Tightens Grip on AI Market with $95 Billion Supply Chain Bet

Nvidia is aggressively locking down the AI chip supply chain, committing nearly $100 billion to vendors to ensure hardware dominance and meet explosive demand.

Read more
articles
Shan2026-04-13
HardwareNvidiaCompute

Nvidia's Blackwell Architecture Finally Hits Global Data Centers

The much anticipated Blackwell chips are spinning up at massive scale, promising a dramatic reduction in AI inference costs.

Read more