FuriosaAI Ditches GPU Playbook For 2nm Broadcom-Built Inference Chip, Claims HBM4/E Bandwidth Beats Even The Most Efficient GPUs
… The company claims that its focus on bandwidth rather than thread management required by GPUs will help it deliver higher efficiency and higher token throughput than modern GPU designs. …
