Skip to content

Models

Fastlane routes your requests to the best model for each task. Below are all models available across BYOK providers, OpenRouter, and Marketplace.

Model Context Price (Input) Best For Modes
GPT-4o 128k $2.50/1M Complex reasoning, coding Accurate, Balanced
GPT-4o Mini 128k $0.15/1M General tasks, quick answers Cheap, Balanced, Eco
Model Context Price (Input) Best For Modes
Claude Sonnet 4 200k $3.00/1M Analysis, writing, coding Accurate, Balanced
Claude 3.5 Haiku 200k $1.00/1M Fast responses, summarization Cheap, Balanced
Model Context Price (Input) Best For Modes
Llama 3.3 70B 128k $0.59/1M Open-source quality, fast Balanced, Accurate
Llama 3.1 8B 128k $0.20/1M Quick tasks, high volume Eco, Cheap
Mixtral 8x7B 32k $0.27/1M Balanced open-source Cheap, Balanced
Gemma 2 9B 8k $0.20/1M Ultra-fast, simple tasks Eco, Cheap

When you send a request with fastlane-auto, Fastlane:

  1. Analyzes your prompt
  2. Scores every available model using the ONNX ranking model
  3. Picks the best model based on your mode (Eco, Cheap, Balanced, Accurate)
  4. Routes the request and returns the response

You can also send requests directly to any model by name.

Feature BYOK Marketplace
Requires API key Yes No
Pricing Provider rates 3% markup
Model selection Your keys only All models
Affiliation N/A 1% commission

All prices shown are per 1 million input tokens. Output tokens are typically 2-4x more expensive.

Prices include the 3% Marketplace markup where applicable. BYOK prices are direct from the provider.