Available Models & Pricing
Complete list of AI models available through Sub2API with transparent pricing.
INFO
Prices shown are in USD per 1 million tokens. Actual pricing may vary. Check the public API endpoint for current rates.
Model Categories
Models available through Sub2API span multiple providers and use cases.
Chat Completion Models
Conversational AI for text generation:
- OpenAI: GPT-4 Turbo, GPT-4, GPT-3.5 Turbo
- Anthropic: Claude 3.5 Sonnet, Claude 3 Opus, Claude 3 Sonnet, Claude 3 Haiku
- Google: Gemini 1.5 Pro, Gemini 1.5 Flash, Gemini 1.0 Pro
Pricing Structure
Sub2API uses transparent per-token pricing:
- Input tokens: Your prompts and conversation context
- Output tokens: Generated responses
- No hidden fees: What you see is what you pay
- Real-time tracking: Monitor usage in the console
Example Pricing
Typical models pricing (reference only, check API for current rates):
GPT-4 Turbo:
- Input: ~$10 per 1M tokens
- Output: ~$30 per 1M tokens
GPT-3.5 Turbo:
- Input: ~$0.50 per 1M tokens
- Output: ~$1.50 per 1M tokens
Claude 3.5 Sonnet:
- Input: ~$3 per 1M tokens
- Output: ~$15 per 1M tokens
Claude 3 Haiku:
- Input: ~$0.25 per 1M tokens
- Output: ~$1.25 per 1M tokens
Gemini 1.5 Pro:
- Input: ~$1.25 per 1M tokens
- Output: ~$5 per 1M tokens
Gemini 1.5 Flash:
- Input: ~$0.075 per 1M tokens
- Output: ~$0.30 per 1M tokens
WARNING
Prices shown are approximate and for reference only. Use the public models API endpoint to fetch current pricing programmatically, or check the console for the latest rates.
Model Selection Guide
By Use Case
Creative Writing: GPT-4, Claude 3 Opus
- High-quality, nuanced content
- Character consistency
- Complex narratives
Code Generation: GPT-4 Turbo, Claude 3.5 Sonnet
- Accurate syntax
- Good at debugging
- Multi-language support
Data Analysis: GPT-4 Turbo, Claude 3.5 Sonnet
- Complex reasoning
- Pattern recognition
- Detailed explanations
Customer Support: GPT-3.5 Turbo, Claude 3 Haiku
- Fast responses
- Cost-effective at scale
- Handles common queries
Document Processing: Gemini 1.5 Pro, Claude 3 models
- Long context windows
- Comprehensive analysis
- Multi-format support
Rapid Prototyping: GPT-3.5 Turbo, Gemini 1.5 Flash
- Quick iterations
- Low cost per request
- Good for testing
By Context Length
Huge (200K+ tokens):
- Claude 3 models (200K)
- Gemini 1.5 models (1M)
Large (32K-128K tokens):
- GPT-4 Turbo (128K)
- GPT-4-32K (32K)
Standard (8K-16K tokens):
- GPT-4 (8K)
- GPT-3.5 Turbo (16K)
Getting Current Pricing
Via Public API
Fetch live pricing without authentication:
curl https://api.tokenlio.ai/api/v1/public/modelsReturns JSON with current pricing, savings percentages, and model capabilities.
Via Console
Log in to app.tokenlio.ai to view:
- Current pricing for all models
- Your workspace's available models
- Usage costs and projections
Model Identifiers
Use these identifiers in API requests:
OpenAI
gpt-4-turbogpt-4gpt-3.5-turbo
Anthropic
claude-3-5-sonnetclaude-3-opusclaude-3-sonnetclaude-3-haiku
Google
gemini-1.5-progemini-1.5-flashgemini-1.0-pro
Usage Examples
Choose by Task Complexity
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.tokenlio.ai/v1"
)
# Simple task - use cheaper model
response = client.chat.completions.create(
model="gpt-3.5-turbo",
messages=[{"role": "user", "content": "Summarize in one sentence"}]
)
# Complex task - use powerful model
response = client.chat.completions.create(
model="gpt-4-turbo",
messages=[{"role": "user", "content": "Analyze this legal document..."}]
)Dynamic Model Selection
def select_model(task_complexity: str) -> str:
if task_complexity == "simple":
return "gpt-3.5-turbo"
elif task_complexity == "moderate":
return "claude-3-haiku"
elif task_complexity == "complex":
return "gpt-4-turbo"
else:
return "claude-3-5-sonnet"FAQ
Q: Can I request access to a specific model?
A: Yes, email support@tokenlio.ai with your use case.
Q: Are prices guaranteed?
A: Prices may change with notice. Current usage is always charged at current rates.
Q: Do you offer volume discounts?
A: Contact sales@tokenlio.ai for enterprise pricing.
Q: How do I know which model to use?
A: Start with GPT-3.5 Turbo for testing, then upgrade to GPT-4 or Claude for better quality if needed.