Llama 3.1 8B Instruct on Novita AI

Llama 3.1 · AI at Meta

ServerlessOpen Weights

Last refreshed 2026-07-09. Next refresh: weekly.

Why use Llama 3.1 8B Instruct on Novita AI?

Novita AI offers Llama 3.1 8B Instruct with pay-as-you-go pricing at $0.02/1M input tokens. Novita AI offers a GPU-based inference API for image, video, and language model generation with a broad catalog of open-source models.

Compare Llama 3.1 8B Instruct across 17 providers to find the best fit for your use case
Input / 1M
$0.020
Output / 1M
$0.050
Cache
Not sourced
Batch
Not sourced

Setup recipe

Docs fallback
Install
Use the provider REST API or SDK
Auth
Create a provider API key
Call
model: llama-3.1-8b-instruct
Model ID
llama-3.1-8b-instruct

Request example

Curated snippets for this provider have not been sourced yet.

Gotchas

  • Use provider model ID "llama-3.1-8b-instruct", not the LLMReference slug "llama3.1-8b-instruct".

Compare Llama 3.1 8B Instruct Across Providers

ProviderInput (per 1M)Output (per 1M)
Cloudflare Workers AI——
OctoAI API (Deprecated)——
Together AI$0.18$0.18
Fireworks AI$0.20$0.20
NVIDIA NIM——
View all 17 providers →

Pricing

TypePrice (per 1M)
Input tokens$0.02
Output tokens$0.05

Capabilities

Structured Outputs

About Llama 3.1 8B Instruct

The Llama 3.1 8B Instruct model, released on July 23, 2024, is a multilingual large language model with 8 billion parameters, optimized for instruction-following tasks. It features an enhanced transformer architecture, supporting languages like English, German, French, and others. The model excels in dialogue applications, having been fine-tuned using supervised fine-tuning and reinforcement learning with human feedback. Trained on approximately 15 trillion tokens with a December 2023 data cutoff, it outperforms many existing open-source and closed chat models in various benchmarks.

Get Started