Skip to main content

Supported Providers

BoostGPT supports 14 AI providers with 105 models spanning from ultra-fast nano models to advanced reasoning models. Thirteen are hosted for you; Ollama runs on your own machine.

OpenAI

20 models — GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna and more

Google

9 models — Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.5 Flash Lite and more

Anthropic

9 models — Claude Fable 5, Claude Opus 4.8, Claude Sonnet 5 and more

xAI

5 models — Grok 4.5, Grok 4.3, Grok 4.20 Reasoning and more

Zhipu AI

5 models — GLM-4.6, GLM-4.7, GLM-4.7 Flash and more

MiniMax

3 models — MiniMax M3, MiniMax M2.7, MiniMax M2.5

DeepSeek

2 models — DeepSeek V4 Flash, DeepSeek V4 Pro

Moonshot AI

5 models — Kimi K3, Kimi K2.7 Code, Kimi K2.7 Code Highspeed and more

Mistral AI

3 models — Mistral Small, Mistral Medium, Mistral Large

Cohere

5 models — Command A, Command R+, Command R and more

Groq

4 models — DeepSeek Llama (70B), GPT-OSS (20B), GPT-OSS (120B) and more

Fireworks AI

12 models — DeepSeek V3, DeepSeek R1, Llama V3 (8B) and more

Together AI

2 models — Qwen 2.5 Coder (32B), Qwen 2.5 Turbo (72B)

Ollama

4 models — Llama 3.3 (70B) Local, Llama 3.1 (8B) Local, Llama 3.1 (70B) Local and more

Key Features

Use your own API keys for OpenAI, Google, Anthropic, xAI, DeepSeek, Mistral, Cohere, Groq, or Ollama. You control costs and rate limits.
Don’t have API keys? Use our pooled infrastructure on paid plans. We handle provisioning, rate limits, and scaling.
Access advanced reasoning models like O1, O3 Mini, DeepSeek R1, Gemini Flash Thinking, and Claude Extended Thinking for complex problem-solving.
Choose models based on speed, cost, reasoning capability, and context window. Use fast models for simple queries, reasoning models for complex tasks.
Run models locally using Ollama. Keep data private, eliminate API costs, and use custom fine-tuned models.

Model Categories

Speed-Optimized (Nano/Mini)

Ultra-fast models for simple tasks, high-volume applications, and real-time responses. Best for: FAQs, basic chat, high-traffic bots
Examples: GPT-4.1 Nano, Llama 3.1 Instant (8B), Gemini 2.5 Flash Lite

Balanced (Standard)

Great all-around models balancing speed, cost, and capability. Best for: Most production use cases, customer support, content generation
Examples: GPT-4o Mini, Claude Sonnet 4.5, Gemini 2.5 Flash

Advanced (Pro/Large)

Powerful models for complex tasks requiring deep understanding. Best for: Complex queries, creative writing, technical analysis
Examples: GPT-5, Claude Opus 4.1, Gemini 2.5 Pro

Reasoning Models

Specialized models with extended thinking for complex problem-solving. Best for: Math, coding, logic puzzles, multi-step reasoning
Examples: O1, DeepSeek R1, Gemini Flash Thinking, Claude Extended Thinking
Learn more about reasoning models →

Model Comparison

*Local hosting costs (compute) not included View full comparison →

Using Your Own Models (Ollama)

The only way to use custom models in BoostGPT is through Ollama.
1

Install Ollama

Download and install Ollama on your server or local machine.
2

Pull Your Model

3

Configure BoostGPT

Set your Ollama host URL in the dashboard or via SDK:
4

Deploy

Your agent will now use your self-hosted Ollama model. No API costs, full control.
Learn more about Ollama →

Credits System

Each model has a credit cost per message. Credits are consumed when your agent generates responses.
  • Nano/Mini models: 0.5-1 credits (cheapest)
  • Standard models: 2-3 credits
  • Pro/Large models: 4-5 credits
  • Reasoning models: 3-6 credits (most expensive, but highest quality)
Credits are only for agent responses. Incoming messages, training data, and API calls don’t consume credits.

Choosing the Right Model

For Creators (No-Code Dashboard)

  1. Go to your bot settings
  2. Click “Model Selection”
  3. Choose based on your needs:
    • Fast responses → GPT-4o Mini, Gemini Flash Lite
    • Best quality → GPT-5, Claude Opus, Gemini Pro
    • Complex reasoning → O1, DeepSeek R1
    • Low cost → Use nano/mini variants

For Developers (SDK/API)

Provider-Specific Setup

Bring Your Own Keys

Use your existing API keys from any provider

Use Our Infrastructure

Let us handle provisioning (paid plans only)

Next Steps

OpenAI

23 models — GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna and more

Google

10 models — Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash and more

Anthropic

10 models — Claude Fable 5, Claude Opus 5, Claude Opus 4.8 and more

xAI

6 models — Grok 4.6, Grok 4.5, Grok 4.3 and more

Zhipu AI

8 models — GLM-5.3, GLM-5.3 Flash, GLM-5.2 and more

MiniMax

5 models — MiniMax M2.7 Highspeed, MiniMax M2.5 Highspeed, MiniMax M3 and more

DeepSeek

3 models — DeepSeek V4 Flash, DeepSeek V4 Flash Vision, DeepSeek V4 Pro

Moonshot AI

5 models — Kimi K3, Kimi K2.7 Code, Kimi K2.7 Code Highspeed and more

Mistral AI

3 models — Mistral Small, Mistral Medium, Mistral Large

Cohere

5 models — Command A, Command R+, Command R and more

Groq

4 models — GPT-OSS (20B), Qwen 3.8 (27B), Groq Compound and more

Fireworks AI

12 models — DeepSeek V3, DeepSeek R1, Llama V3 (8B) and more

Together AI

7 models — GLM-5.3, GLM-5.2, Kimi K3 and more

Ollama

4 models — Llama 3.3 (70B) Local, Llama 3.1 (8B) Local, Llama 3.1 (70B) Local and more