Groq® - Hyperfast LLM running on custom built GPUs

An LPU Inference Engine, with LPU standing for Language Processing Unit™, is a new type of end-to-end processing unit system that provides the fastest inference at ~500 tokens/second.

Add a comment

Replies

Best
This seems extremely interesting- I’m curious what you’ve seen to be the biggest use case for this LLM?
It is fast, that is for sure. Where can I get more information about the chips and hardware? Is there a GPU cloud service?
oh, thank you. Will dig in.
Congratulations! speed/accuracy is incredible, no wonder NVDA took a dip 😯
Groq is a promising product, and I believe your detailed insights could attract even more supporters, helping people better understand its value.
Man, that IS fast...Already loving it : )
this will be incredible for the future of LLMs and all the products benefiting from them. super excited with all the new things that will come
Congrat on the launch? Do you have any plan when to support custom training?
Good luck! I am really excited about this hardware stuff for LLMs!
Going to give this a try, team Groq®. Looks interesting.
Don't know why its ranked so low as of now, the speed is awesome. It does what it says.
12
Next