AI in the product

LLM API

Call large language models from your product.

Ask your AI about this, with this page as the source:ChatGPT ↗Claude ↗Perplexity ↗

The real choice

Which provider, and whether to go through a gateway. Going direct to a provider gets its newest models and features first and the lowest latency; a gateway gives you one key and one bill for hundreds of models, with fallbacks when a provider is down, for a small markup or a proxy you run. The code that calls the model — an AI SDK or agent framework — is its own decision.

Pick by situation

Tap the ones that are you — the tools that fit light up below.
  • You want the strongest general models and the widest ecosystem of examples and integrationsOpenAI APIClaudeGemini API
  • Token cost dominates your budget, for example high-volume batch or agent workloadsDeepSeekGroqMistral AI
  • You want to try or mix many models with one key and one bill, and fall back when a provider is downOpenRouter
  • You want that routing and spend control but self-hosted, with keys and logs on your serversLiteLLM

The contenders

Grouped by the side of the choice they answer, not ranked. Open a row for when to use it, the trade-off and what makers say.

ToolBest for
Model provider5
OpenAI APIapiThe broadest model lineup — text, images, speech, embeddings — and the largest ecosystem.734+481 open source$10GPT-6 Luna34.1M/wk+61% vs npm

Use it whenYou want one account that covers text, images, speech and embeddings, with the most examples to copy from.

Trade-offFrequent model and API changes to keep up with, and no option to run its hosted models elsewhere.

llms.txtIn Lovable, Replit

Used by 2dto3D, 8base, Agentplace.io, AgentX, Agnost AI and 729 more · in 481 open-source projects.

We use GPT-5.6 to power the Basedash AI data analyst. It's incredibly intelligent and, with the right harness, allows us to rank #1 on BI Bench for solving real-world BI scenarios.
Basedash, the makerSep 2026 ↗
ClaudeapiCoding, long documents and agent-style tool use.251+285 open source$10Claude Haiku 5.531.6M/wk4.1× vs npm

Use it whenYour feature is an agent that calls tools, writes code or works through long documents.

Trade-offNo image generation or embedding models of its own, so those need a second provider.

llms.txtIn Lovable, Replit

Used by 8base, Accordio AI, Agentcard, Agentplace.io, AgentX and 246 more · in 285 open-source projects.

Taskade Genesis works beautifully with Claude. It brings strong reasoning and long-range thinking to any Genesis app, helping systems plan, adapt, and make sense of complex inputs.
Taskade, the makerSep 2026 ↗
Gemini APIapi · free tierVery long context and multimodal input, with a free tier to start.99+227 open source$28Gemini 3.1 Flash-Lite19.2M/wk6.5× vs npm

Use it whenYou feed in whole documents, video or audio, or want to start without paying.

Trade-offFree-tier prompts may be used to improve Google's products, so paid tier is the one for user data.

In Lovable, Replit

Used by 1752vc Pitch Deck Analyzer, 8base, Accorata, Agentplace.io, Appear on AI and 94 more · in 227 open-source projects.

Powers all agent conversations on Konfide. Fast, cost-effective, handles unlimited concurrent chats. Every user message goes through Gemini. Chose it for speed and quality at scale.
Konfide, the makerSep 2026 ↗
Mistral AIapi · free tierA European provider whose API covers chat, coding, OCR and speech models.15+69 open source$6.00Ministral 3 (3B)7.1M/wk6× vs npm

Use it whenYou want an EU-based provider, or document OCR alongside text models.

Trade-offA smaller ecosystem of third-party integrations than the largest providers.

llms.txtIn Replit

Used by Internet.io, Kuration AI, MeetMinutes, Officely AI, Omnifact and 10 more · in 69 open-source projects.

Talespinner uses Mistral Large as a writing model. It is great for writing novels with more mature topics, since it doesn't censor your writing that much.
Talespinner, the makerSep 2026 ↗
DeepSeekapiLow per-token prices on strong reasoning and coding models, with OpenAI- and Anthropic-format endpoints.11+24 open source$27deepseek-flash1.9M/wk4.9× vs npm

Use it whenToken cost dominates your budget, for example high-volume batch or agent workloads.

Trade-offServed from China, which matters for some customers' data requirements; peak-hour prices are double off-peak.

Used by Genstore.ai, IFTTT, MGX (Now Atoms), Read & Give by Bono, Shipable AI by CNTXT AI and 6 more · in 24 open-source projects.

DeepSeek delivers serious reasoning at a fraction of the cost of comparable models. We chose it because powerful AI shouldn't have a premium price tag attached to every use case.
IFTTT, the makerSep 2026 ↗
Open models, hosted1
Groqapi · free tierVery low-latency inference on open models, for real-time features.39+62 open source$6.75GPT OSS 20B1.8M/wk3.9× vs npm

Use it whenResponse speed matters more than having the most capable model, as in voice or autocomplete.

Trade-offOnly the open models it chooses to host, and no frontier closed models.

Used by Antispace, Automaticall, Bebop.ai, ClarityUX, Extra Thursday and 34 more · in 62 open-source projects.

Shoutout to Groq — the lightning fast inference engine running Voicr's AI under the hood. Without Groq's speed, that under 3 seconds promise wouldn't be possible.
Voicr for Mac, the makerSep 2026 ↗
Gateway (many providers, one API)2
OpenRouterapi · free tierTrying many models from many providers with one API key and one bill, without opening an account at each.61+31 open source—2.1M/wk4.8× vs npm

Use it whenYou want to compare models quickly or offer users a model picker.

Trade-offA fee on top of provider prices, and your traffic and data pass through a third party.

llms.txtIn Replit

Used by Actx0, AI Emaily, Breadcrumb, ChikitAI, Clado and 56 more · in 31 open-source projects.

Talespinner uses OpenRouter to allow for switching between different language models. They make it super easy to do this, because they offer one API format for all language models.
Talespinner, the makerSep 2026 ↗
LiteLLMlibrary · open source · free tierRunning your own OpenAI-compatible proxy in front of your own provider keys, with retries, routing and per-key budgets.21+97 open source—148.1M/wk

Use it whenYou want keys and request data to stay on your infrastructure.

Trade-offYou deploy, monitor and upgrade it yourself; SSO and audit logs are in the paid enterprise tier.

llms.txtMCP

Used by CamelAI, Crossnode, Fume, PandaProbe, PandaProbe Cloud and 16 more · in 97 open-source projects.

We use LiteLLM mainly so we’re not tied to a single model provider. Makes it super easy to switch models or add fallbacks when running AI services for clients.
Crossnode, the makerSep 2026 ↗

Cost: the cheapest plan that fits input tokens 50 million tokens, output tokens 10 million tokens, from list prices. Try your own numbers →

Cost as you grow

Each contender's cheapest usable plan as usage rises.

$0$100$500$1,000$2,0001510501005001k5k
DeepSeekMistral AIGroqGemini APIClaudeOpenAI APIx: input tokens per month (million tokens), other usage scaled with it · cheapest usable plan at each point, list prices · try your own numbers

Who switches to what

Public pull requests on GitHub since Oct 2024 whose title says "X to Y" — real code changes, by developers in general rather than makers only. Pick a flow to see its pull requests.

OpenAI API → Gemini API: 179+ pull requestsGemini API → Groq: 152 pull requestsOpenAI API → Claude: 127 pull requestsGemini API → Claude: 108+ pull requestsClaude → Gemini API: 105+ pull requestsGemini API → OpenAI API: 90+ pull requestsClaude → OpenAI API: 89 pull requestsGroq → Gemini API: 64 pull requestsGemini API 350OpenAI API 306Claude 194Groq 64Gemini API 348Claude 235OpenAI API 179Groq 152Moving fromMoving to

Before you choose

How to approach it

Put model calls behind one function in your code so switching providers is a one-line change. Log prompts, responses and cost from day one. Set spending limits on every provider account.

Common mistakes
  • Shipping with no spending limit or per-user rate limit, so one abusive user or a retry loop runs up the bill overnight.
  • Hard-coding one model name across the codebase, so moving to a newer or cheaper model means a search-and-replace and a round of surprises.
  • Putting the provider API key in client-side code, where anyone can copy it from the browser.

Other options

Real choices most makers here won't need to weigh.

Azure OpenAIThrough your cloudOpenAI models inside an Azure account, with regional and EU/US data-zone deployments.23 makers' productsTogether AIOpen models, hostedPer-token access to a wide catalog of open-weight models, with fine-tuning and dedicated GPUs from the same vendor.12 makers' productsGrok APIModel providerGrok models with built-in web and X search, plus image, video and voice APIs.6 makers' productsMiniMaxModel providerLow per-token prices on open-weight coding and agent models, with speech, image, video and music models in the same account.6 makers' productsFireworks AIOpen models, hostedServerless and dedicated inference on popular open-weight models, with fine-tuning and batch jobs at half the serverless price.5 makers' productsVertex AIThrough your cloudGemini and partner models such as Claude on Google Cloud, with its regions, IAM and billing.3 makers' productsPortkeyGateway (many providers, one API)A gateway that also handles guardrails, prompt management and request logs in one dashboard.2 makers' productsEden AIGateway (many providers, one API)One API key and bill for LLMs and non-chat AI — OCR, speech, translation, vision — across many providers, with an EU endpoint.2 makers' productsDeepInfraOpen models, hostedLow per-token prices on a large catalog of open-weight models, plus embeddings, speech and image models, behind an OpenAI-compatible API.2 makers' productsAmazon BedrockThrough your cloudClaude, Llama, Mistral, Amazon Nova and other models inside an AWS account, billed with the rest of your AWS usage.1 makers' productReplicateModel providerOpen-source image, video and audio models without running GPUs yourself.Vercel AI GatewayGateway (many providers, one API)One key and one bill for hundreds of models at provider list prices, with fallbacks and budgets, built into the AI SDK.BifrostGateway (many providers, one API)A self-hosted, open-source gateway written in Go that routes one OpenAI-compatible API to many providers with failover, load balancing and caching.Qwen APIModel providerQwen models plus DeepSeek, Kimi and GLM under one Alibaba Cloud account, with OpenAI- and Anthropic-format endpoints.Kimi APIModel providerKimi K3, an open-weight model with 1M-token context, and coding models behind OpenAI- and Anthropic-format endpoints.GLM APIModel providerGLM coding and agent models through Z.ai, with free Flash models and Anthropic-format access for Claude Code.Doubao APIModel providerByteDance's Seed models alongside its Seedream image and Seedance video models, with an OpenAI-compatible API.ERNIE APIModel providerBaidu's ERNIE models through Qianfan, with an OpenAI-compatible endpoint.Hunyuan APIModel providerTencent's HY models plus DeepSeek, GLM, Kimi and MiniMax on one TokenHub key, in OpenAI and Anthropic formats.Cerebras Inference APIOpen models, hostedVery fast output — thousands of tokens per second — on a small set of open-weight models.Meta Model APIModel providerMeta's Muse Spark models for agents and coding, with a 1M-token context, multimodal input and OpenAI- and Anthropic-format endpoints.Xiaomi MiMo APIModel providerVery low per-token prices on open-weight models with a 1M-token context and image, audio and video input, through OpenAI- and Anthropic-format endpoints.

Decided alongside

What the 919 makers' products here chose for their other decisions.