Infrastructure

InfrastructureAMD Handed Half of Inference to a Rival Chip

AMD Handed Half of Inference to a Rival Chip

On July 23 AMD and Cerebras built one inference pipeline from two rival chips: AMD Helios handles the prompt, the Cerebras wafer writes the answer. The 5x efficiency figure is a model, not a benchmark. Here is what disaggregated inference changes for your compute budget.

3 min read
Infrastructure256 Cores on One Chip Changes Your Rack Math

256 Cores on One Chip Changes Your Rack Math

At its Advancing AI event on July 22 AMD launched EPYC Venice, a 256-core server chip on TSMC 2nm claiming a 70 percent efficiency gain. Here is what that does to your rack count, your power bill and your negotiating position.

3 min read

Page 4 / 8