GenStorAIGE AI90 Reduces First Token Latency by Up to 50x
…Combined with intelligent peer-to-peer GPU communication, the company says the platform can accelerate inference by up to 5.8× on systems equipped with eight NVIDIA GeForce RTX 5090 graphics cards…
