GROQ SUPERCHARGES FAST AI INFERENCE FOR META LLAMA 3.1
New Largest and Most Capable Openly Available Foundation Model 405B Running on GroqCloud™
"I'm really excited to see Groq's ultra-low-latency inference for cloud deployments of the Llama 3.1 models. This is an awesome example of how our commitment to open source is driving innovation and progress in AI. By making our models and tools available to the community, companies like Groq can build on our work and help push the whole ecosystem forward."
"Meta is creating the equivalent of Linux, an open operating system, for AI – not only for the Groq LPU which provides fast AI inference, but for the entire ecosystem. In technology, open always wins, and with this release of Llama 3.1, Meta has caught up to the best proprietary models. At this rate, it's only a matter of time before they'll pull ahead of the closed models," said
The Llama 3.1 models are a significant step forward in terms of capabilities and functionality. As the largest and most capable openly available Large Language Model to date, Llama 3.1 405B rivals industry-leading closed-source models. For the first time, enterprises, startups, researchers, and developers can access a model of this scale and capability without proprietary restrictions, enabling unprecedented collaboration and innovation. With Groq, AI innovators can now tap into the immense potential of Llama 3.1 405B running at unprecedented speeds on GroqCloud to build more sophisticated and powerful applications.
With Llama 3.1, including 405B, 70B, and 8B Instruct models, the AI community gains access to increased context length up to 128K and support across eight languages. Llama 3.1 405B is in a class of its own, with unmatched flexibility, control, and state-of-the-art capabilities in general knowledge, steerability, math, tool use, and multilingual translation. Llama 3.1 405B will unlock new capabilities, such as synthetic data generation and model distillation, and deliver new security and safety tools to further the shared Meta and Groq commitment to build an open and responsible AI ecosystem.
With unprecedented inference speeds for large openly available models like Llama 3.1 405B, developers are able to unlock new use cases that rely on agentic workflows to provide a seamless, yet personalized, human-like response for use cases such as: patient coordination and care; dynamic pricing by analyzing market demand and adjusting prices in real-time; predictive maintenance using real-time sensor data; and customer service by responding to customer inquiries and resolving issues in seconds.
GroqCloud has grown to over 300,000 developers in five months, underscoring the importance of speed when it comes to building the next generation of AI-powered applications at a fraction of the cost of GPUs.
To experience Llama 3.1 models running at Groq speed, visit groq.com, and learn more about this launch from Groq and Meta.
About Groq
Groq builds fast AI inference technology. Groq® LPU™ AI inference technology is a hardware and software platform that delivers exceptional AI compute speed, quality, and energy efficiency. Groq, headquartered in
Media Contact
[email protected]
View original content to download multimedia:https://www.prnewswire.com/news-releases/groq-supercharges-fast-ai-inference-for-meta-llama-3-1--302204185.html
SOURCE Groq
Serious News for Serious Traders! Try StreetInsider.com Premium Free!
You May Also Be Interested In
- Best Graphic Design Software for Print and Production Work (2026): CorelDRAW Recognized for Print-Friendly, Production-Ready Workflows by Better Business Advice
- Life Peptides Launches Company-Wide Transparency Initiative to Advance Research Peptide Quality and Buyer Education
- Lennar Debuts Sutton and Hollis, Two New Home Collections at Valencia in the Santa Clarita Valley
Create E-mail Alert Related Categories
PRNewswire, Press ReleasesRelated Entities
Mark ZuckerbergSign up for StreetInsider Free!
Receive full access to all new and archived articles, unlimited portfolio tracking, e-mail alerts, custom newswires and RSS feeds - and more!



Tweet
Share