Quantcast
Channel: Groq
Browsing index pages (102 articles)

Build Faster with Groq + Hugging Face

Simplicity of Hugging Face + Efficiency of GroqExciting news for developers and AI enthusiasts! Hugging Face is making it easier than ever to access Groq’s lightning-fast and efficient inference with...

View Article


Image may be NSFW.
Clik here to view.

GroqCloud™ Now Supports Qwen3 32B

Delivering Fast Inference with the Full 131k Context WindowGroqCloud now supports Qwen3 32B, a cutting-edge, dense 32.8 billion parameter causal language model from Alibaba’s Qwen3 series. This...

View Article


Introducing GroqCloud™ LoRA Fine-Tune Support: Unlock Efficient Model...

GroqCloud now supports Low-Rank Adaptation (LoRA) fine-tunes, exclusively by request, for our Enterprise tier customers. LoRA enables businesses to deploy adaptations of base models customized to...

View Article

Image may be NSFW.
Clik here to view.

From Speed to Scale: How Groq Is Optimized for MoE & Other Large Models

You know Groq runs small models. But did you know we run large models including MoE uniquely well? Here’s why. The Evolution of Advanced Openly-Available LLMs There’s no argument that Artificial...

View Article

Image may be NSFW.
Clik here to view.

How to Build Your Own AI Research Agent with One Groq API Call

Simplifying the Complexity of AI Agents with Server-Side Tool Use Large Language Models (LLMs) are powerful but constrained by static training data, lacking the ability to access real-time information...

View Article


The Official Llama API, Accelerated by Groq

The official Llama API is now accelerated by Groq. Served on the world’s most efficient inference chip, it’s the fastest way to run the world’s most trusted openly available models with no tradeoffs....

View Article

Image may be NSFW.
Clik here to view.

Now in Preview: Groq’s First Compound AI System

Build with access to the internet and the ability to run code with a one line change to your model string. Compound Beta is Groq’s first compound AI system, released under preview on GroqCloud. It...

View Article

Image may be NSFW.
Clik here to view.

Llama 4 Live Today on Groq — Build Fast at the Lowest Cost, Without Compromise

Meta’s Llama 4 Scout and Maverick models are live today on GroqCloud, giving developers and enterprises day-zero access to the most advanced open-source AI models available. Today, Meta released the...

View Article


Image may be NSFW.
Clik here to view.

Build Fast with Text-to-Speech

Groq & PlayAI partner to bring Dialog, a leading TTS model, to GroqCloud for real-time voice applications One of the most popular emerging applications for applied AI has been generative voice...

View Article


Image may be NSFW.
Clik here to view.

Groq & Vercel Partner To Make Building Fast and Simple

Connect Your Vercel Projects Directly to GroqCloud, With Billing All in One Place What do you get when you combine the power of speed with simplicity? The newly announced Groq and Vercel Marketplace...

View Article

Image may be NSFW.
Clik here to view.

Batch Processing with GroqCloud™ for AI Inference Workloads

GroqCloud provides fast inference for complex AI solutions that require instant responsiveness. But what happens when your use cases expand and require features beyond speed? That’s where Batch...

View Article

Image may be NSFW.
Clik here to view.

Build Fast with Word-Level Timestamping

Imagine if you could search a group of audio recordings and jump directly to the exact word you were looking for. Or, if you were captioning a video and you wanted the words to appear precisely as...

View Article

Image may be NSFW.
Clik here to view.

A Guide to Reasoning with Qwen QwQ 32B

Author: Hatice Ozen Scaling Reinforcement Learning is All You Need for the Rise of Smaller, Smarter Models On March 5th, Alibaba Cloud’s Qwen team broke the internet with the release of QwQ-32B less...

View Article


Image may be NSFW.
Clik here to view.

What is a Language Processing Unit?

OverviewGroq LPU AI Inference Technology  Groq builds fast AI inference. Groq® LPU AI inference technology delivers exceptional AI compute speed, quality, and affordability at scale.Groq AI inference...

View Article

Image may be NSFW.
Clik here to view.

Qwen QwQ 32B Running Same Day As Release 

With a community of over one million developers who build FAST, Groq can’t help but want to keep up. That’s why we ship fast, like today’s launch of Qwen-qwq-32b on GroqCloud.  Performance The 32B...

View Article


Image may be NSFW.
Clik here to view.

How to Win Hackathons with Groq

Winning a hackathon can be a life-changing experience with a platform to showcase your skills, network with industry leaders, and gain recognition for your innovative solutions, the possibilities are...

View Article

Image may be NSFW.
Clik here to view.

Mistral Saba Added to GroqCloud™ Model Suite

GroqCloud has added another openly-available model to our suite – Mistral Saba. Mistral Saba is Mistral AI’s first specialized regional language model, custom-trained to serve specific geographies,...

View Article


Image may be NSFW.
Clik here to view.

Enabling LLMOps with Fast AI Inference

Customer Use Case: Orq About Orq.ai: Orq.ai is the end-to-end platform for serious software teams to control GenAI at scale. Build, ship, and scale LLM applications – all in one place. The Challenge...

View Article

Image may be NSFW.
Clik here to view.

Groq Customer Use Case: Fintool

About Fintool: Fintool is an AI equity research copilot for institutional investors. Fintool uses Large Language Models (LLMs) to discover financial insights beyond the reach of timely human analysis....

View Article

Image may be NSFW.
Clik here to view.

Groq Customer Use Case: Data Leaders

About Data Leaders: Your custom AI sales associate answers questions, books meetings, gathers lead info, and more to close more deals. The Challenge We all know the drill. You call a business to learn...

View Article
Browsing index pages (102 articles)


Latest Images