“A hybrid system combining GPUs and Groq's LPUs can achieve the optimal cost-per-token and capacity across any desired inference speed.”