🚀 One API for every frontier model on FastInfra

Browse models

Transparent per-token pricing → compare providers on FastInfra

View pricing

OpenAI-compatible chat completions → start in minutes

Read the docs

⚡ Server Auction → Enterprise bare metal from $78.15/mo (173 in stock)

Browse deals

Gpt 4o Mini 2024 07 18 vs L3 8B Lunaris v1 Turbo

Live API pricing and availability, side by side. Both models run on FastInfra's OpenAI-compatible endpoint — switching between them is a one-line change.

Pricing

Price per 1M tokens

Prices are live and sync automatically from wholesale providers.

Gpt 4o Mini 2024 07 18 L3 8B Lunaris v1 Turbo
Input $0.16 $0.04
Output $0.63 $0.05
Default route route-07 route-21
Routes available 1 1

On input tokens, L3 8B Lunaris v1 Turbo is currently 3.8× cheaper than Gpt 4o Mini 2024 07 18 on FastInfra. Output-token pricing and quality trade-offs differ per workload — test both with the free tier.

Quickstart

Try both in 30 seconds

Same endpoint, same SDK — only the model string changes.

from openai import OpenAI

client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")

for model in ["openai/gpt-4o-mini-2024-07-18", "Sao10K/L3-8B-Lunaris-v1-Turbo"]:
    response = client.chat.completions.create(
        model=model,
        messages=[{"role": "user", "content": "Hello!"}]
    )
    print(model, "->", response.choices[0].message.content)