❌

Normal view

There are new articles available, click to refresh the page.
Before yesterdayMain stream

NVIDIA’s First Made-In-America GB300 GPUs Roll Off TSMC’s Arizona-Based Fab 21, But Still Fly To Taiwan for Packaging

27 July 2026 at 16:07

NVIDIA and TSMC have just achieved a major milestone that significantly bolsters the prospects of on-shore chip fabrication, insulating the US against geopolitical shocks. Even so, there is still quite a lot left to be done, including on the advanced packaging front. TSMC has just fabricated some of the first NVIDIA GB300 GPUs in its Arizona facility, taking a significant step towards fulfilling the demands of the US to onshore chip manufacturing According to the Commercial Times, TSMC has just fabricated some of the first NVIDIA GB300 AI GPUs on its 4nm node in Arizona's Fab 21. Even so, the […]

Read full article at https://wccftech.com/nvidias-first-made-in-america-gb300-gpus-roll-off-tsmcs-arizona-based-fab-21-but-still-fly-to-taiwan-for-packaging/

DeepSeek CEO Believes NVIDIA Is Now β€œDigging Its Own Grave” Even As 1 NVIDIA GB300 GPU Equals 4 Huawei Acend 950 GPUs

23 July 2026 at 15:06

A computer case and motherboard are displayed against a swirling golden background, with visible ports and slots on the case and a complex circuit design on the motherboard.

DeepSeek's CEO, Liang Wenfeng, has disclosed a number of valuable nuggets in an investor-focused conference call, spilling the beans on China's compute deficit, the superiority of NVIDIA's GB300 GPUs, Huawei's meticulous catch-up, and DeepSeek's ongoing work on data annotation. DeepSeek's Wenfeng Says Huawei's Atlas 950 SuperPoD can fully substitute NVIDIA's GB300 NVL rack, but for every GB300 GPU DeepSeek needs 4x Ascend GPUs DeepSeek's Wenfeng disclosed a slew of valuable information in a recent investor conference call, noting that the AI lab's available compute only allows for models that activate around 10 billion parameters at any given time vs. the […]

Read full article at https://wccftech.com/deepseek-ceo-believes-nvidia-is-now-digging-its-own-grave-even-as-1-nvidia-gb300-gpu-equals-4-huawei-acend-950-gpus/

NVIDIA Blackwell GB300 Continues To Set World Records for MoE Pre-Training While GB200 Sees A 4x Boost In Perf/W Through Continuous AI Software Stack Optimizations

22 July 2026 at 02:30

NVIDIA Blackwell GB300 Continues To Set World Records for MoE Pre-Training While GB200 Sees A 4x Boost In Perf/W Through Continuous AI Software Stack Optimizations

NVIDIA's Blackwell GB300 and GB200 GPUs continue to showcase leading capabilities across AI workloads through continuous optimizations. NVIDIA Isn't Done With Blackwell Yet As GB300 & GB200 Deployments Receive Performance & Efficiency Boosting Upgrades Across The Globe One may think that NVIDIA will move over to its Rubin GPUs now that its next-gen AI platform is being deployed around the world, but no, NVIDIA Blackwell is already deployed in countless data centers, and just like Hopper before it, the Blackwell generation continues to see optimizations, so much so that GB300 is still posting record performance in AI workloads. According to […]

Read full article at https://wccftech.com/nvidia-blackwell-gb300-continues-to-set-world-records-gb200-gets-4x-ai-boost-optimizations/

NVIDIA Brings Local AI Agents To Its Most Powerful Workstation PC, The DGX Station, With The NVIDIA Agent Toolkit & Omniverse

20 July 2026 at 15:00

A cartoon lobster holds a green shield with a lock icon next to a desk setup featuring an unbranded monitor displaying a 3D warehouse simulation with colorful overlays and a humanoid figure.

NVIDIA has enabled DGX Station users to run personal AI agents locally through its NVIDIA Agent Toolkit, which can be set up in just three steps. DGX Station With GB300 Is A Super Workstation PC, & You Can Now Easily Run Super AI Agents On It With NVIDIA's Agent Toolkit Agentic AI has brought new kinds of use cases for PCs, especially powerful ones such as NVIDIA's DGX Station, which packs the GB300 "Blackwell Ultra" GPU. To make full use of its power, NVIDIA is offering its Agent Toolkit software, which unleashes these capabilities in just 30 minutes (the time […]

Read full article at https://wccftech.com/nvidia-brings-local-ai-agents-to-dgx-station-pc-with-agent-toolkit/

SpaceX Awards Foxconn A Part In A Huge $52 Billion Order For 13,000 Racks Of NVIDIA GB300 AI Servers, Where Each Rack Costs $4 Million And The Total Order Spans Nearly 1 Million GPUs

19 July 2026 at 23:58

SpaceX has just given Foxconn a major order for NVIDIA GB300 AI server racks, significantly upping its training- and inference-related compute footprint. Foxconn has now added SpaceX to its growing roster of hyperscaler deals According to a report by a Taiwan-based publication, SpaceX has just awarded Foxconn a part in a new $52 billion order for around 13,000 racks of NVIDIA GB300 AI servers, with shipments expected to begin in Q4 2026 and continuing until Q1 2027, with each rack costing around $4 million. Also, the order spans 936,000 NVIDIA GPUs. Do note that it is not yet clear if […]

Read full article at https://wccftech.com/spacex-awards-foxconn-a-huge-52-billion-order-for-13000-racks-of-nvidia-gb300-ai-servers-where-each-rack-costs-4-million/

NVIDIA Slashes DeepSeek v4 Token Costs By Up To 5x Just One Month After Launch, Through Pure Blackwell Software Tuning

30 June 2026 at 19:30

NVIDIA Slashes DeepSeek v4 Token Costs By Up To 5x Just One Month After Launch, Through Pure Blackwell Software Tuning

NVIDIA Blackwell GPUs continue to see massive optimizations, leading to a 5x drop in token cost in DeepSeek v4 AI models. NVIDIA Cost Per Token Narrative Sees Massive Gain In DeepSeek V4 As AI Model Sees 5x Boost On Blackwell GPUs With Continued Optimizations "Cost Per Token" is the fundamental metric for AI TCO, as NVIDIA highlighted this a few months back, and now, the company is delivering the lowest-ever token cost in DeepSeek v4. Today, NVIDIA announced that its full-stack inference software has brought further optimizations to its hardware stack, such as Blackwell GB200 & GB300, improving their performance […]

Read full article at https://wccftech.com/nvidia-slashes-deepseek-v4-token-costs-by-up-to-5x-one-month-after-launch/

NVIDIA’s Blackwell Ultra GB300 Now Powers Anthropic’s Claude Models on Microsoft Azure, Targeting Autonomous Enterprise Agents

29 June 2026 at 19:05

Anthropic, Microsoft, and NVIDIA logos are displayed on a black background.

Anthropic has announced the general availability of its Claude AI models on Microsoft Azure, powered by NVIDIA's Blackwell Ultra GPUs. Anthropic & NVIDIA Bring Claude AI Models to Microsoft Azure - Running on NVIDIA's Blackwell Ultra GB300 Platform As Agentic workloads continue to drive the AI segment, NVIDIA is working with leading firms to deliver faster and expanded capabilities to end-users. Today, NVIDIA partnered with Anthropic to deliver Claude models on Microsoft Azure. With the general availability of Anthropic's Claude models, Microsoft Azure provides billing, authentication, and governance controls. The two models that are available today are Claude Opus 4.8 […]

Read full article at https://wccftech.com/nvidia-blackwell-ultra-powers-anthropic-claude-models-in-microsoft-azure/

Jim Keller Says Cerebras IPO Was Helpful As Tenstorrent Set To β€œBeat Them on Everything”, Confirms Meeting With Intel & Qualcomm CEOs β€œHoping To Get A Big Deal”

27 June 2026 at 21:45

Samsung To Build Next-Gen Tenstorrent AI Chiplet Leveraging RISC-V Architecture 1

Jim Keller isn't bothered by Cerebras's recent IPO and says that he welcomes it, but Tenstorrent will still beat them on everything. Tenstorrent CEO, Jim Keller, Signals Deal With Intel or Qualcomm While Promising To Beat Cerebras "on everything" Tenstorrent recently introduced its latest BlackHole Galaxy server, a system with which it can disrupt the entire AI segment, with performance levels that crush the competition. We covered the announcement last month when the company demoed its Blackhole server undercutting a NVIDIA GB300 with up to five times better TCO. Keller Accepts The Challenge To Beat NVIDIA, Cerebras & Others At […]

Read full article at https://wccftech.com/jim-keller-cerebras-ipo-was-helpful-tenstorrent-to-beat-them-on-everything/

Taking an Up-Close Look at the Supermicro GB300 Super AI Station

27 June 2026 at 15:00

With NVIDIA DGX Station systems now shipping, Supermicro had their Super AI Station on display a Computex, showing off the GB300 system in all of its 208 billion transistor glory

The post Taking an Up-Close Look at the Supermicro GB300 Super AI Station appeared first on ServeTheHome.

NVIDIA GB300 Dominates Agentic AI Workloads With 20x Performance Leap Over Hopper As Rubin Nears Launch

14 June 2026 at 12:25

A close-up view of an NVIDIA circuit board featuring multiple processing units, mounted on a dark background.

NVIDIA's Blackwell GB300 has posted record performance in AA-AgentPerf, a new benchmark that measures Agentic AI workflows. NVIDIA Blackwell Ultra GB300 is 20 Times Faster Than Hopper In Agentic AI, Records Highest Performance In Latest Benchmarks Artificial Analysis has a new benchmark out called AA-AgentPerf, which measures how many active agents an inference deployment can support under realistic workloads, which include: The AA-AgentPerf benchmark is used to measure three key metrics, which form the basis of modern-day AI deployments, such as: NVIDIA is now publishing its first benchmarks in AgentPerf measures using DeepSeek V4 Pro on its GB300 NVL72 platform. […]

Read full article at https://wccftech.com/nvidia-gb300-dominates-agentic-ai-workloads-20x-performance-leap-over-hopper/

NVIDIA Confirms Vera Rubin Launch In Q3 With Volume Ramp by Q4, As Blackwell Continues To See Massive Demand

21 May 2026 at 03:00

NVIDIA Confirms Vera Rubin Launch In Q3 With Volume Ramp by Q4, As Blackwell Continues To See Massive Demand

NVIDIA confirms its next-gen Vera Rubin AI platform timeline, but existing GPUs continue to see major traction from AI firms, leading to price hikes. NVIDIA Vera Rubin Is All Prepped To Power Agentic AI Factories In Q3, Volume Ramp In Q4 2026 & Bigger Numbers In Early 2027 Today, NVIDIA announced its Q1 FY2027 earnings with a record revenue of $81.6 billion, up 85% from the last year. The revenue was driven primarily by the Data Center segment, which held the mammoth share of $75.24 billion, while Edge Computing amounted to $6.3 billion in revenue. Revenue by Market Platform (in […]

Read full article at https://wccftech.com/nvidia-confirms-vera-rubin-launch-in-q3-volume-ramp-q4-blackwell-continues-to-see-massive-demand/

NVIDIA Feynman GPUs Push Power Semi Content To $191,000, 17 Times Increase Over Blackwell As Industry Embraces 800V DC Architectures

4 May 2026 at 16:00

NVIDIA Feynman GPUs Push Power Semi Content To $191,000, 17 Times Increase Over Blackwell As Industry Pushes 800V DC Architectures

As compute requirements grow in AI datacenters, so do the power requirements, which are estimated to reach 17x higher with NVIDIA's Feynman. NVIDIA Feynman Racks Estimated To Feature 17x Higher Power Semi Costs Per Rack Versus Blackwell NVIDIA Feynman GPUs feature several groundbreaking features and will launch in 2028, after Rubin. The company has been working hard to deliver more efficient AI solutions, but as requirements grow, power requirements have increased tremendously. Morgan Stanley Research has published a chart that visualizes the total power semi content of three AI rack solutions from NVIDIA. Starting with the baseline Blackwell or B200, […]

Read full article at https://wccftech.com/nvidia-feynman-gpus-push-power-semi-content-17-times-higher-vs-blackwell/

Agentic AI Pushes CPUs to Pack 400 GB of Memory, 4x More Than Today, as DRAM Shortage Spirals Toward 2027

2 May 2026 at 10:10

CPUs or GPUs, but require lots of memory for running Agentic AI, and this demand is spiraling to unseen levels as DRAM constraints persist. CPUs Running Agentic AI Will Be Equipped With Up to 400 GB of Memory, Further Crushing The DRAM Supply Chain Memory makers are earning big profits but are also unable to meet the demand. We have seen reports on how major manufacturers are rapidly expanding their production facilities, but these are yet to become operational, and Samsung itself has stated that 2027 will be worse for the DRAM industry than 2026, so it's looking like a […]

Read full article at https://wccftech.com/agentic-ai-pushes-cpus-to-pack-400-gb-of-memory-4x-more-than-today/

NVIDIA Beats Everyone To DeepSeek V4 With Day-0 Blackwell Support, Pushing 3,500 Tokens Per Second On 1.6T Models

26 April 2026 at 09:10

A person stands next to a large NVIDIA data center server rack with multiple GPUs and visible branding.

DeepSeek V4 is out, bringing major optimizations, including up to 1.6T model sizes, and NVIDIA is ready with Day-0 support on Blackwell GPUs using NVFP4. NVIDIA Blackwell NVFP4 Architecture Delivers Major Speed-Ups In DeepSeek v4 With More Optimizations On The Way With the launch of DeepSeek V4, we saw some major optimizations in compute & memory requirements. The updated AI modelΒ uses just 27% of single-token inference FLOPs & 10% of the KV cache when running a one-million-token context window. Two new models were also introduced, one being a Pro model with a parameter size of 1.6T, and a Flash version […]

Read full article at https://wccftech.com/nvidia-beats-everyone-to-deepseek-v4-day-0-blackwell-support-pushing-3500-tokens-on-1-6t-models/

MSI XpertStation WS300 NVIDIA GB300 Station and More at the MSI NVIDIA GTC 2026 Booth

6 April 2026 at 17:00

At NVIDIA GTCX 2026, we saw MSI servers ranging from the desktop EdgeXpert and XpertStation WS300 to air and liquid-cooled multi-GPU servers

The post MSI XpertStation WS300 NVIDIA GB300 Station and More at the MSI NVIDIA GTC 2026 Booth appeared first on ServeTheHome.

NVIDIA Is Among the First to Submit MLPerf Inference v6.0 Benchmarks With Blackwell Ultra, and It’s Total Domination Over Competitors

1 April 2026 at 18:46

Man showcasing NVIDIA GPU on stage with server racks in the background.

NVIDIA has become one of the first to submit the 'extensive' MLPerf Inference v6.0 benchmarks, delivering the highest performance relative to "all competitors" combined. NVIDIA's Blackwell Ultra, Combined With Extreme Co-Design Laws, Manages to Dominate With MLPerf v6.0 Benchmarks When it comes to benchmark submissions and showcasing the 'prowess' of its computing platforms, NVIDIA has been at the forefront, particularly with MLPerf, where the firm is one of the few entities to complete a rigorous round of benchmarks. This time, according to the company's latest blog post, NVIDIA has discussed its latest submission to MLPerf v6.0, noting that, with Blackwell […]

Read full article at https://wccftech.com/nvidia-is-among-the-first-to-submit-mlperf-inference-v6-0-benchmarks/

Gigabyte NVIDIA Vera Rubin and More at NVIDIA GTC 2026

31 March 2026 at 04:34

We take you around the Gigabyte booth at NVIDIA GTC 2026 and see NVIDIA Vera Rubin platforms, and tons of new systems and components

The post Gigabyte NVIDIA Vera Rubin and More at NVIDIA GTC 2026 appeared first on ServeTheHome.

❌
❌