The Latest News in AI

We publish news articles on Forbes, which are copied here for your convenience.  

AI Cambrian Explosion: 2021 Predictions

For the last two years, in January, I have published my predictions about the new AI silicon I expect in the coming year. Like the weatherman, I’ve gotten some details wrong but I’ve been reasonably accurate, at least directionally. So, here we go again with a look at...

Neutral-Atom Quantum: What Is It, And Why Infleqtion Stands Out

Late last year, I wrote a Forbes piece about IBM, the leader in quantum computing using superconducting circuits, and promised an update on other modalities. This is the first follow-on piece and explores neutral-atom (NA) quantum, which operates at room temperature,...

Enhanced Memory Grace Hopper Superchip Could Shift Demand To NVIDIA CPU And Away From X86

The company’s new high bandwidth memory version is only available with the CPU-GPU Superchip. In addition, a new dual Grace-Hopper MGX Board offers 282GB of fast memory for large model inferencing. The AI landscape continues to change rapidly, and fast memory (HBM)...

The Graphcore Data Center Architecture

The Graphcore disaggregated accelerator could be a game changer. I have recently finished a research paper looking into the data center architecture for deploying the Graphcore IPU-Machine, which is a network-attached accelerator for highly-parallel workloads. Lets...

The Raging Debate: When Will Quantum Arrive?

There has been considerable discussion and stock price volatility of late surrounding the expected timing of useful applications and hardware for quantum computing. One month after Google created excitement around its Willow quantum chip, Nvidia CEO Jensen Huang and...

d-Matrix Emerges From Stealth With Strong AI Performance And Efficiency

Startup launches “Corsair” AI platform with Digital In-Memory Computing, using on-chip SRAM memory that can produce 30,000 tokens/second at 2 ms/token latency for Llama3 70B in a single rack. Using Generative AI, called inference processing, is a memory-intensive...

NVIDIA Adds New Software That Can Double H100 Inference Performance

TensorRT-LLM adds a slew of new performance-enhancing features to all NVIDIA GPUs. Just ahead of the next round of MLPerf benchmarks, NVIDIA has announced a new TensorRT software for Large Language Models (LLMs) that can dramatically improve performance and efficiency...

GrAI Matter Labs: Brain-Inspired AI For The Edge

This article was written by Alberto Romero, Cambrian-AI Analyst, and Karl Freund, Cambrian-AI Founder.We recently tweeted about the startup GrAI Matter Labs (GML) and received a lot of questions about the company’s products and strategy. As one of the first startups...

HotChips Preview: Nvidia Scales AI Beyond The Data Center

The annual HotChips conference starts this Sunday, Aug. 24, in San Francisco. Nvidia is scheduled to present six sessions covering topics of interest to AI data center users and operators and will make several key announcements I’ll cover in this article. (Like most...

Intel’s New Chips Focus On AI

Today, Intel launched and disclosed new technologies across its portfolio of processors, with special emphasis on enhanced AI capabilities. Unlike its many competitors, who either produce a CPU, a GPU, an FPGA or an AI-specific accelerator, Intel’s strategy is “all of...