Code · Groq
Custom-chip inference, 500+ tokens/sec
GroqCloud serves open-weight models on custom LPU chips, delivering some of the fastest inference speeds in the industry at low per-token cost.
[ Pros ]
[ Cons ]
[ Tags ]