Forgeworker guide

GroqCloud

Low-latency inference for supported language models. A hosted inference API using Groq hardware for supported text, speech, and multimodal models.

About GroqCloud

A hosted inference API using Groq hardware for supported text, speech, and multimodal models.

How you access it

Use the provider service or API. Hosted API, Low latency

Capabilities

  • Text
  • Reasoning
  • Speech

Published pricing

Per-model usage pricing

Review note

Directory reviewed on 2026-07-28. Review the provider privacy policy before sending sensitive information.

Runware

Fast image and media inference through one API.

OpenRouter

One API for a wide catalog of language models.

This server-rendered summary is available to search engines and no-JavaScript visitors. JavaScript loads the full interactive experience.