Groq® LPU™ Inference Engine Leads in First Independent LLM Benchmark
ArtificialAnalysis.ai Adjusts Chart Axes to Accommodate Groq Performance Levels
"ArtificialAnalysis.ai has independently benchmarked Groq and its Llama 2 Chat (70B) API as achieving throughput of 241 tokens per second, more than double the speed of other hosting providers," said ArtificialAnalysis.ai Co-creator
Groq has run several internal benchmarks, reaching 300 tokens per second consistently, setting a new speed standard for AI solutions that has yet to be achieved by legacy solutions and incumbent providers. ArtificialAnalysis.ai benchmarks confirm Groq superiority over other providers, especially regarding throughput at 241 tokens per second and total time to receive 100 output tokens at 0.8 seconds according to the benchmark techniques of input prompt size and output prompt size. For more benchmark details please visit https://groq.link/aabenchmark.
"Groq exists to eliminate the 'haves and have-nots' and to help everyone in the AI community thrive," said Groq CEO and founder
ArtificialAnalysis.ai benchmarks are conducted independently and are 'live' in that they are updated every three hours (eight times per day). Prompts are unique, around 100 tokens in length, and generate ~200 output tokens. This is designed to reflect real-world usage and measures changes to throughput (tokens per second) and latency (time to first token) over time. Benchmarks are also present on ArtificialAnalyis.ai with longer prompts to reflect retrieval augmented generation (RAG) use cases.
The LPU Inference Engine is available through the Groq API. For access, please complete the request form at https://groq.link/contact.
About Groq
Groq® is a generative AI solutions company and the creator of the LPU™ Inference Engine, the fastest language processing accelerator on the market. It is architected from the ground up to achieve low latency, energy-efficient, and repeatable inference performance at scale. Customers rely on the LPU Inference Engine as an end-to-end solution for running Large Language Models (LLMs) and other generative AI applications at 10x the speed. The LPU Inference Engine is available via the GroqCloud, an API that enables customers to purchase Tokens-as-a-Service for experimentation and production-ready applications.
Media Contact for Groq
[email protected]
View original content to download multimedia:https://www.prnewswire.com/news-releases/groq-lpu-inference-engine-leads-in-first-independent-llm-benchmark-302060263.html
SOURCE Groq
Serious News for Serious Traders! Try StreetInsider.com Premium Free!
You May Also Be Interested In
- Kelley Blue Book Brand Watch: Hybrid and SUV Interest Reaches New Highs as Audi, Kia and Subaru Gain Shopper Consideration in First Half of 2026
- Southern Power Foundation distributes grant to Haskell County, Texas, fire department
- VizSense Launches "Go Out. Do Good." -- the Inaugural Campaign of Its Influencing for Good Series
Create E-mail Alert Related Categories
PRNewswire, Press ReleasesSign up for StreetInsider Free!
Receive full access to all new and archived articles, unlimited portfolio tracking, e-mail alerts, custom newswires and RSS feeds - and more!



Tweet
Share