Build Faster with Groq + Hugging Face
Simplicity of Hugging Face + Efficiency of GroqExciting news for developers and AI enthusiasts! Hugging Face is making it easier than ever to access Groq’s lightning-fast and efficient inference with...
View ArticleGroqCloud™ Now Supports Qwen3 32B
Delivering Fast Inference with the Full 131k Context WindowGroqCloud now supports Qwen3 32B, a cutting-edge, dense 32.8 billion parameter causal language model from Alibaba’s Qwen3 series. This...
View ArticleIntroducing GroqCloud™ LoRA Fine-Tune Support: Unlock Efficient Model...
GroqCloud now supports Low-Rank Adaptation (LoRA) fine-tunes, exclusively by request, for our Enterprise tier customers. LoRA enables businesses to deploy adaptations of base models customized to...
View ArticleFrom Speed to Scale: How Groq Is Optimized for MoE & Other Large Models
You know Groq runs small models. But did you know we run large models including MoE uniquely well? Here’s why. The Evolution of Advanced Openly-Available LLMs There’s no argument that Artificial...
View ArticleHow to Build Your Own AI Research Agent with One Groq API Call
Simplifying the Complexity of AI Agents with Server-Side Tool Use Large Language Models (LLMs) are powerful but constrained by static training data, lacking the ability to access real-time information...
View ArticleThe Official Llama API, Accelerated by Groq
The official Llama API is now accelerated by Groq. Served on the world’s most efficient inference chip, it’s the fastest way to run the world’s most trusted openly available models with no tradeoffs....
View ArticleNow in Preview: Groq’s First Compound AI System
Build with access to the internet and the ability to run code with a one line change to your model string. Compound Beta is Groq’s first compound AI system, released under preview on GroqCloud. It...
View ArticleLlama 4 Live Today on Groq — Build Fast at the Lowest Cost, Without Compromise
Meta’s Llama 4 Scout and Maverick models are live today on GroqCloud, giving developers and enterprises day-zero access to the most advanced open-source AI models available. Today, Meta released the...
View ArticleBuild Fast with Text-to-Speech
Groq & PlayAI partner to bring Dialog, a leading TTS model, to GroqCloud for real-time voice applications One of the most popular emerging applications for applied AI has been generative voice...
View ArticleGroq & Vercel Partner To Make Building Fast and Simple
Connect Your Vercel Projects Directly to GroqCloud, With Billing All in One Place What do you get when you combine the power of speed with simplicity? The newly announced Groq and Vercel Marketplace...
View ArticleBatch Processing with GroqCloud™ for AI Inference Workloads
GroqCloud provides fast inference for complex AI solutions that require instant responsiveness. But what happens when your use cases expand and require features beyond speed? That’s where Batch...
View ArticleBuild Fast with Word-Level Timestamping
Imagine if you could search a group of audio recordings and jump directly to the exact word you were looking for. Or, if you were captioning a video and you wanted the words to appear precisely as...
View ArticleA Guide to Reasoning with Qwen QwQ 32B
Author: Hatice Ozen Scaling Reinforcement Learning is All You Need for the Rise of Smaller, Smarter Models On March 5th, Alibaba Cloud’s Qwen team broke the internet with the release of QwQ-32B less...
View ArticleWhat is a Language Processing Unit?
OverviewGroq LPU AI Inference Technology Groq builds fast AI inference. Groq® LPU AI inference technology delivers exceptional AI compute speed, quality, and affordability at scale.Groq AI inference...
View ArticleQwen QwQ 32B Running Same Day As Release
With a community of over one million developers who build FAST, Groq can’t help but want to keep up. That’s why we ship fast, like today’s launch of Qwen-qwq-32b on GroqCloud. Performance The 32B...
View ArticleHow to Win Hackathons with Groq
Winning a hackathon can be a life-changing experience with a platform to showcase your skills, network with industry leaders, and gain recognition for your innovative solutions, the possibilities are...
View ArticleMistral Saba Added to GroqCloud™ Model Suite
GroqCloud has added another openly-available model to our suite – Mistral Saba. Mistral Saba is Mistral AI’s first specialized regional language model, custom-trained to serve specific geographies,...
View ArticleEnabling LLMOps with Fast AI Inference
Customer Use Case: Orq About Orq.ai: Orq.ai is the end-to-end platform for serious software teams to control GenAI at scale. Build, ship, and scale LLM applications – all in one place. The Challenge...
View ArticleGroq Customer Use Case: Fintool
About Fintool: Fintool is an AI equity research copilot for institutional investors. Fintool uses Large Language Models (LLMs) to discover financial insights beyond the reach of timely human analysis....
View ArticleGroq Customer Use Case: Data Leaders
About Data Leaders: Your custom AI sales associate answers questions, books meetings, gathers lead info, and more to close more deals. The Challenge We all know the drill. You call a business to learn...
View Article