Supported Providers
BoostGPT supports 14 AI providers with 105 models spanning from ultra-fast nano models to advanced reasoning models. Thirteen are hosted for you; Ollama runs on your own machine.OpenAI
20 models — GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna and more
9 models — Gemini 3.6 Flash, Gemini 3.5 Flash, Gemini 3.5 Flash Lite and more
Anthropic
9 models — Claude Fable 5, Claude Opus 4.8, Claude Sonnet 5 and more
xAI
5 models — Grok 4.5, Grok 4.3, Grok 4.20 Reasoning and more
Zhipu AI
5 models — GLM-4.6, GLM-4.7, GLM-4.7 Flash and more
MiniMax
3 models — MiniMax M3, MiniMax M2.7, MiniMax M2.5
DeepSeek
2 models — DeepSeek V4 Flash, DeepSeek V4 Pro
Moonshot AI
5 models — Kimi K3, Kimi K2.7 Code, Kimi K2.7 Code Highspeed and more
Mistral AI
3 models — Mistral Small, Mistral Medium, Mistral Large
Cohere
5 models — Command A, Command R+, Command R and more
Groq
4 models — DeepSeek Llama (70B), GPT-OSS (20B), GPT-OSS (120B) and more
Fireworks AI
12 models — DeepSeek V3, DeepSeek R1, Llama V3 (8B) and more
Together AI
2 models — Qwen 2.5 Coder (32B), Qwen 2.5 Turbo (72B)
Ollama
4 models — Llama 3.3 (70B) Local, Llama 3.1 (8B) Local, Llama 3.1 (70B) Local and more
Key Features
Bring Your Own API Keys
Bring Your Own API Keys
Use your own API keys for OpenAI, Google, Anthropic, xAI, DeepSeek, Mistral, Cohere, Groq, or Ollama. You control costs and rate limits.
Use Our Infrastructure (Optional)
Use Our Infrastructure (Optional)
Don’t have API keys? Use our pooled infrastructure on paid plans. We handle provisioning, rate limits, and scaling.
Reasoning Models
Reasoning Models
Access advanced reasoning models like O1, O3 Mini, DeepSeek R1, Gemini Flash Thinking, and Claude Extended Thinking for complex problem-solving.
Smart Model Selection
Smart Model Selection
Choose models based on speed, cost, reasoning capability, and context window. Use fast models for simple queries, reasoning models for complex tasks.
Self-Hosted with Ollama
Self-Hosted with Ollama
Run models locally using Ollama. Keep data private, eliminate API costs, and use custom fine-tuned models.
Model Categories
Speed-Optimized (Nano/Mini)
Ultra-fast models for simple tasks, high-volume applications, and real-time responses. Best for: FAQs, basic chat, high-traffic botsExamples: GPT-4.1 Nano, Llama 3.1 Instant (8B), Gemini 2.5 Flash Lite
Balanced (Standard)
Great all-around models balancing speed, cost, and capability. Best for: Most production use cases, customer support, content generationExamples: GPT-4o Mini, Claude Sonnet 4.5, Gemini 2.5 Flash
Advanced (Pro/Large)
Powerful models for complex tasks requiring deep understanding. Best for: Complex queries, creative writing, technical analysisExamples: GPT-5, Claude Opus 4.1, Gemini 2.5 Pro
Reasoning Models
Specialized models with extended thinking for complex problem-solving. Best for: Math, coding, logic puzzles, multi-step reasoningExamples: O1, DeepSeek R1, Gemini Flash Thinking, Claude Extended Thinking Learn more about reasoning models →
Model Comparison
*Local hosting costs (compute) not included
View full comparison →
Using Your Own Models (Ollama)
The only way to use custom models in BoostGPT is through Ollama.1
Install Ollama
Download and install Ollama on your server or local machine.
2
Pull Your Model
3
Configure BoostGPT
Set your Ollama host URL in the dashboard or via SDK:
4
Deploy
Your agent will now use your self-hosted Ollama model. No API costs, full control.
Credits System
Each model has a credit cost per message. Credits are consumed when your agent generates responses.- Nano/Mini models: 0.5-1 credits (cheapest)
- Standard models: 2-3 credits
- Pro/Large models: 4-5 credits
- Reasoning models: 3-6 credits (most expensive, but highest quality)
Credits are only for agent responses. Incoming messages, training data, and API calls don’t consume credits.
Choosing the Right Model
For Creators (No-Code Dashboard)
- Go to your bot settings
- Click “Model Selection”
- Choose based on your needs:
- Fast responses → GPT-4o Mini, Gemini Flash Lite
- Best quality → GPT-5, Claude Opus, Gemini Pro
- Complex reasoning → O1, DeepSeek R1
- Low cost → Use nano/mini variants
For Developers (SDK/API)
Provider-Specific Setup
Bring Your Own Keys
Use your existing API keys from any provider
Use Our Infrastructure
Let us handle provisioning (paid plans only)
Next Steps
OpenAI
23 models — GPT-5.6 Sol, GPT-5.6 Terra, GPT-5.6 Luna and more
10 models — Gemini 3.7 Flash, Gemini 3.6 Flash, Gemini 3.5 Flash and more
Anthropic
10 models — Claude Fable 5, Claude Opus 5, Claude Opus 4.8 and more
xAI
6 models — Grok 4.6, Grok 4.5, Grok 4.3 and more
Zhipu AI
8 models — GLM-5.3, GLM-5.3 Flash, GLM-5.2 and more
MiniMax
5 models — MiniMax M2.7 Highspeed, MiniMax M2.5 Highspeed, MiniMax M3 and more
DeepSeek
3 models — DeepSeek V4 Flash, DeepSeek V4 Flash Vision, DeepSeek V4 Pro
Moonshot AI
5 models — Kimi K3, Kimi K2.7 Code, Kimi K2.7 Code Highspeed and more
Mistral AI
3 models — Mistral Small, Mistral Medium, Mistral Large
Cohere
5 models — Command A, Command R+, Command R and more
Groq
4 models — GPT-OSS (20B), Qwen 3.8 (27B), Groq Compound and more
Fireworks AI
12 models — DeepSeek V3, DeepSeek R1, Llama V3 (8B) and more
Together AI
7 models — GLM-5.3, GLM-5.2, Kimi K3 and more
Ollama
4 models — Llama 3.3 (70B) Local, Llama 3.1 (8B) Local, Llama 3.1 (70B) Local and more