Normal view

There are new articles available, click to refresh the page.
Before yesterdayMain stream

AMD Unleashes Instinct MI455X GPU, A 320 Billion Transistor Behemoth That Is Designed to Tackle NVIDIA’s Rubin With 50% More HBM4 Memory & Up to 40 PFLOPs of AI Compute

23 July 2026 at 18:25

AMD Instinct chip installed on a server motherboard.

AMD's Instinct MI455X GPUs extend the company's AI roadmap, bringing leading HBM4 capacities and over 40 PFLOPs of compute for Agentic AI, rivaling NVIDIA's Rubin chip. AMD Has An Answer To NVIDIA's Rubin, It's Called Instinct MI455X & It's An Engineering Marvel For Agentic AI With More HBM4 Than Any Other AI Chip On The Planet The AMD Instinct MI455X is the GPU that will power the Helios AI rack. This GPU is based on the latest CDNA 5 architecture and is made using TSMC's 2nm and 3nm process technologies. The chip packs 320 billion transistors, just 16 billion transistors […]

Read full article at https://wccftech.com/amd-instinct-mi455x-gpu-320b-behemoth-tackle-nvidia-rubin-with-432gb-hbm4-40-pflops-ai/

NVIDIA Vera CPU Is Architected For The Agentic AI Era, as It Delivers Max Single-Core & Single-Thread Performance Versus x86; Full Architectural Breakdown Shows

21 July 2026 at 15:00

NVIDIA Pivots to CPUs to Salvage China Revenue After GPU Restrictions, As Several Customers Show Interest In Vera CPUs

The NVIDIA Vera CPU is designed for Reinforcement Learning and Agentic AI systems, offering 88 Olympus cores for max single-core performance. NVIDIA's x86-Killer, Vera CPU, Is Now Available, Built Purely With Agentic AI In Mind Surge in agentic AI workloads and reinforcement learning emerging as a key scaling mechanism for improving model capability have made CPUs a vital component for AI systems. Without a fast CPU (execution, evaluation, orchestration), a GPU won't be able to be supplied with context, and that leads to major drawbacks and inefficiencies. This is why NVIDIA designed Rubin, a CPU that is purpose-built for agents, […]

Read full article at https://wccftech.com/nvidia-vera-cpu-architecture/

Tensordyne’s 3nm Napier AI Chip Promises 13x Higher Token Throughput Than Blackwell & Blazes Past Rubin With 1000 Tokens/s In Multi-Trillion Parameter Models

15 June 2026 at 17:40

A person wearing gloves holds a TensorDyne TDN AIP chip, with visible circuitry and labeled sections.

US-based AI company, Tensordyne, has announced the successful tape-out of its Napier chip, which it claims to demolish NVIDIA's Blackwell & Rubin chips with leading token throughput and efficiency. Tensordyne’s new Napier AI Chip arrives with one clear mission: to make NVIDIA’s Blackwell and Rubin chips look considerably less impressive The Napier chip will be the core component of the Tensordyne Napier TDN system, which is designed in collaboration with Broadcom and HPE Juniper Networks. The Napier platform has one goal: to unify AI through novel logarithmic AI math, a tightly integrated memory architecture, and a high-performance scale-up interconnect that […]

Read full article at https://wccftech.com/tensordyne-3nm-napier-ai-chip-13x-higher-token-throughput-blackwell-blazes-past-rubin/

NVIDIA’s Rubin AI Platform Alone Will Devour More LPDDR Memory in 2027 Than Apple and Samsung Combined, Starving Smartphone Supply

17 May 2026 at 14:15

LPDDR Memory Demand For NVIDIA Rubin In AI To Exceed The Combined Demand of Samsung & Apple In 2027

LPDDR memory demand in the AI & server segment is going to outstrip smartphone demand massively in 2027, with next-gen platforms such as NVIDIA Rubin and AMD MI400 requiring way more DRAM to meet the needs of AI firms. AI LPDDR Demand Is Going To Push Back The Smartphone Market Massively, NVIDIA Rubin's Demand Alone To Exceed Both Samsung & Apple Incoming memory demand from Agentic AI use cases is going to be insane. We know that DRAM is going to retain its crucial status in powering upcoming AI servers, and future platforms are going to house more capacity as […]

Read full article at https://wccftech.com/nvidia-rubin-ai-platform-will-devour-more-lpddr-memory-in-2027-than-apple-samsung-combined/

AMD MI430X Is The “Highest Performance” FP64 GPU Ever Built, Surpassing NVIDIA’s Rubin By 6x In Classic-HPC Workloads

7 May 2026 at 12:30

AMD MI430X Will Be The "Highest Performance" FP64 GPU Ever Built, Surpassing NVIDIA's Rubin By 6x For Classic HPC-Focused Data Centers

AMD's Instinct MI430X GPU will be the fastest FP64 chip, delivering up to 6x the performance of Rubin in classic HPC workloads. AMD Delivers The Biggest Leap In FP64 Compute With MI430X, 6x Rubin Performance & Deployment In The Discovery System at ORNL By 2028 The AI segment continues to push up compute FLOPS in the exascale range through low-precision formats such as FP4, FP6, and FP8. These formats lead the AI spectrum, and while they are important for neural networks, higher-precision formats such as FP64 still hold a lot of value in High-Performance Computing workloads. AMD has always been […]

Read full article at https://wccftech.com/amd-mi430x-highest-performance-fp64-gpu-ever-built-surpassing-nvidia-rubin-by-6x/

130,000 Rubin GPUs Are Being Deployed at Nscale For Microsoft, Further Showing Massive Interest In NVIDIA’s Next-Gen AI Chips

21 April 2026 at 13:45

130,000 NVIDIA Rubin GPUs Are Being Deployed at Nscale For Microsoft, Further Showing Massive Interest In Next-Gen AI Chips 1

Nscale will be adding 30,000 NVIDIA Rubin GPUs in addition to the 100,000 chips that are already being deployed at its facilities in 2027. Nscale Scales Up Its AI Datacenters With 30,000 Additional NVIDIA Rubin GPUs, Bringing The Total Count To 130,000 NVIDIA's Vera Rubin AI chips are hot in demand, and the reason behind this is simple: they are aiming to be the fastest solution for AI inferencing, fueling the race towards Agentic AI. Josh Payne, Founder and CEO at Nscale said: “Customer demand for advanced AI infrastructure continues to accelerate across markets, and our focus is on bringing the […]

Read full article at https://wccftech.com/130000-nvidia-rubin-gpus-deployed-at-nscale-for-microsoft-showing-massive-interest-in-next-gen-ai-chips/

NVIDIA’s Rubin Ultra Reportedly Scaled Back to Dual-Die Design, Instead of the Ambitious Four-Die One, Amid Supply Chain Concerns

1 April 2026 at 13:56

A circuit board features two NVIDIA chips at its center with four large yellow components in a row above them.

NVIDIA's Rubin Ultra GPU has undergone a significant design revision, according to reports from Taiwanese media, which claim that Team Green wants to ensure supply chain complexities are minimal with the new generation. NVIDIA's Rubin Ultra Will Not Feature Four Dies On One Package, But Rather On a Single Board NVIDIA operates with a highly aggressive product cadence, and while the firm's official disclosure is to be at an annual cycle, the supply chain partners need to act a lot more quickly, which is why the timeline is shorter for them, likely at eight to ten months. With that, we […]

Read full article at https://wccftech.com/nvidias-rubin-ultra-reportedly-scaled-back-to-dual-die-design-instead-of-the-ambitious-four-die-one-amid-supply-chain-concerns/

NVIDIA Unveils Vera Rubin With Groq’s LPX to Break Into Inference, a Market Where It Has Never Been First

16 March 2026 at 19:48

A presenter on stage with three open computer servers, showcasing internal components against a black background.

NVIDIA's Groq partnership is now formalizing, as Jensen unveils a hybrid compute tray featuring Groq's third-generation LPU units in a Rubin rack. NVIDIA's Idea With Groq Is to Target 'High-Speed' Workloads, Hoping to Crack the Inference Competition The debate over what NVIDIA would do with Groq has been ongoing for quite some time, and we have maintained a key lead on developments. At GTC 2026, NVIDIA unveiled a new Vera Rubin hybrid compute tray, the Groq 3 LPX, which features eight of the 'unannounced' Groq3 units, which we'll discuss ahead. According to NVIDIA, LPX and Rubin together deliver unprecedented inference […]

Read full article at https://wccftech.com/nvidia-unveils-vera-rubin-with-groq-lpx-to-break-into-inference/

NVIDIA Sees Compute Revenue Exploding to $1 Trillion in Just Two Years, as AI Hits an ‘Inflection Point’ With Inference

16 March 2026 at 19:22

A man wearing dollar sign glasses in front of a semiconductor wafer background.

NVIDIA's CEO, Jensen Huang, has talked about what we anticipate in terms of compute demand growing in the coming years, and the figure he projects is beyond shocking. The AI World Has Brought In 'Wild' Compute Demand With the Inference Crazy, And NVIDIA's Here to Capitalize On It The AI industry is at a defining point, as we are seeing a massive shift from training to inference, driving increased compute demand and aggressive revenue growth. Back at GTC 2025, NVIDIA talked about achieving $500 billion in revenue in three quarters, credited to the company's performance with Blackwell and Vera Rubin, […]

Read full article at https://wccftech.com/nvidia-anticipates-compute-revenue-to-skyrocket-to-1-trillion-in-just-two-years/

Intel Lands Inside NVIDIA’s DGX Rubin NVL8 Systems, With Xeon 6 Becoming the Mission-Critical Host CPU

16 March 2026 at 18:00

Intel's partnership with NVIDIA has finally begun to formalize with the recent announcement, as the company's Xeon 6 CPU series is being integrated into Rubin systems. Intel's Xeon 6 Server CPU Gets Integrated Into NVIDIA's Baseline Rubin Offering; Larger Rack Adoption Remains Uncertain For Now With today's agentic workloads, CPUs have become the next area of focus for hyperscalers and manufacturers like NVIDIA, as functions such as "orchestration, memory access, and model security" have become increasingly dominant. We did talk about Intel attending NVIDIA's GTC this year as a key partner, and with that, the company's recent announcement focuses on […]

Read full article at https://wccftech.com/intel-finally-finds-a-spot-in-nvidia-rubin-systems/

NVIDIA May Finally Abandon Its “One GPU Does Everything” Mantra at GTC 2026, and Here’s What to Expect

15 March 2026 at 20:52

A person is standing on stage showcasing various open server units with visible cooling systems and hardware components.

We are heading towards GTC 2026, one of the most important events within the AI world, and this year, we are expecting a massive shift in how computing is perceived. The race for AI infrastructure has evolved signifcantly over the past few years, as evolving compute requirements have forced companies like NVIDIA and AMD to innovate in what they offer. Since 2022, we have seen training workloads gain massive popularity, which Hopper and Blackwell capitalized on. Now, moving into 2026, agentic workloads are the next area to focus on for compute providers, which is why the upcoming GTC announcements from […]

Read full article at https://wccftech.com/nvidia-may-finally-abandon-its-one-gpu-does-everything-mantra-at-gtc-2026/

NVIDIA’s CEO Says OpenClaw Did in 3 Weeks What Linux Took 30 Years to Achieve; Proof of How Big Agentic AI Really Is

5 March 2026 at 15:23

A man in a black leather jacket with lobster claws for hands holds a device labeled 'CUE.'

NVIDIA's CEO has talked about the 'agentic AI' inflection point at the Morgan Stanley conference, and he has called out OpenClaw as the "most important" software release of our times. NVIDIA's CEO Says that Agentic AI Has Brought Uses 1,000x Higher Tokens, Bringing In Immense Compute Demand Jensen has talked about AI being a "5-layer cake", and one of the more interesting layers that yields the most returns to hyperscalers and frontier labs is the applications layer. OpenClaw and AI agents are examples of how AI, when placed in a hyper-personalized environment, yields results that replicate human workloads. NVIDIA's CEO […]

Read full article at https://wccftech.com/nvidia-ceo-says-openclaw-did-in-3-weeks-what-linux-took-30-years/

❌
❌