Gimlet Labs and Cerebras Team Up for AI at 3,000 Tokens Per Second
Gimlet Labs and Cerebras Systems announced a collaboration to deliver ultrafast AI inference at up to 3,000 tokens per second through the Gimlet Cloud, combining Cerebras’ Wafer Scale Engine with GPUs via disaggregated inference. The first Cerebras-powered Gimlet Cloud datacenter is expected online later this year. Sources: GlobeNewswire via The Manila Times and AiThority, September 28, 2026.