Meta Llama 3.1 8B Instruct AWQ INT4 API
Run togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 through FastInfra's API.
Pay per token, no subscription, routed to the cheapest available provider.
Meta Llama 3.1 8B Instruct AWQ INT4 pricing
Billed per token used. Prices sync automatically from wholesale providers.
| Direction | Price per 1M tokens |
|---|---|
| Input | — |
| Output | — |
1 route available
Requests route to route-03 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.
| Route label | Model ID |
|---|---|
route-03 |
togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 |
Call Meta Llama 3.1 8B Instruct AWQ INT4 in 30 seconds
Works with any OpenAI SDK — change the base URL and API key only.
Python
from openai import OpenAI
client = OpenAI(api_key="YOUR_API_KEY", base_url="https://api.fastinfra.ai/v1")
response = client.chat.completions.create(
model="togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4",
messages=[{"role": "user", "content": "Hello!"}]
)
print(response.choices[0].message.content)
curl
curl https://api.fastinfra.ai/v1/chat/completions \
-H "Authorization: Bearer YOUR_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4",
"messages": [{"role": "user", "content": "Hello!"}]
}'
Meta Llama 3.1 8B Instruct AWQ INT4 — common questions
How much does the togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 API cost?
On FastInfra, togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 costs Pay-per-token pricing via automatic routing. Billing is per token used, with no subscription.
Is togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 compatible with the OpenAI SDK?
Yes. FastInfra exposes togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 through an OpenAI-compatible endpoint at https://api.fastinfra.ai/v1 — point any OpenAI SDK at that base URL with a FastInfra API key and keep your existing code.
How does routing work for togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4?
togethercomputer/meta-llama-3.1-8B-Instruct-AWQ-INT4 is available on 1 route(s): route-03. Requests use route-03 by default (lowest cost). FastInfra fails over automatically if a route is unavailable.
Related models
Compare side by side: Meta Llama 3.1 8B Instruct AWQ INT4 vs Gemma2:2b · Meta Llama 3.1 8B Instruct AWQ INT4 vs Llama3.1:8b · Meta Llama 3.1 8B Instruct AWQ INT4 vs Llama3.2:1b