LLM ReferenceLLM Reference
Fireworks AI

CodeLlama 13B on Fireworks AI

Code Llama · AI at Meta

ProvisionedOpen Source

Last refreshed 2026-04-19. Next refresh: weekly.

Why use CodeLlama 13B on Fireworks AI?

Fireworks AI offers CodeLlama 13B with pay-as-you-go pricing at $0.20/1M input tokens. Fireworks AI offers a generative AI platform as a service, focusing on rapid product iteration and cost-efficient AI deployment.

Compare CodeLlama 13B across 4 providers to find the best fit for your use case
Input / 1M
$0.20
Output / 1M
$0.20
Cache
Not sourced
Batch
Not sourced

Setup recipe

Python + curl
Install
pip install openai
Auth
export FIREWORKS_API_KEY=...
Call
import os
from openai import OpenAI
client = OpenAI(
    api_key=os.environ["FIREWORKS_API_KEY"],
Model ID
codellama-13b

Request example

import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["FIREWORKS_API_KEY"],
    base_url="https://api.fireworks.ai/inference/v1"
)
response = client.chat.completions.create(
    model="codellama-13b",
    messages=[{"role": "user", "content": "Hello"}]
)
print(response.choices[0].message.content)

Gotchas

  • Fireworks model IDs use "accounts/fireworks/models/{model-name}" format, e.g. "accounts/fireworks/models/llama4-scout-instruct-basic" or "accounts/fireworks/models/deepseek-r1".
  • The examples expect FIREWORKS_API_KEY; rename it only if your application config maps the new variable.

Compare CodeLlama 13B Across Providers

ProviderInput (per 1M)Output (per 1M)
Together AI$0.30$0.30
Fireworks AI$0.20$0.20
Microsoft Foundry$0.81$0.94
Replicate API$0.10$0.50

Pricing

TypePrice (per 1M)
Input tokens$0.20
Output tokens$0.20

Capabilities

Structured Outputs

About CodeLlama 13B

CodeLlama 13B is a state-of-the-art generative text model developed by Meta, specifically designed for code synthesis and understanding tasks. Released on August 24, 2023, this 13-billion-parameter model excels in general code generation and comprehension, making it suitable for a wide range of programming tasks, including code completion, infilling, and instruction following. It utilizes an optimized transformer architecture and has been trained on a diverse dataset similar to Llama 2, ensuring robust understanding of programming languages and coding practices. AI engineers can integrate CodeLlama 13B into various coding environments and tools for both commercial and research applications, leveraging its powerful capabilities to enhance productivity and streamline the coding process.

FAQ

What does CodeLlama 13B cost on Fireworks AI?

On Fireworks AI, CodeLlama 13B costs $0.20 per 1M input tokens and $0.20 per 1M output tokens.

What is the context window for CodeLlama 13B on Fireworks AI?

CodeLlama 13B supports a 100,000 token context window on Fireworks AI.

How does Fireworks AI compare to other CodeLlama 13B providers?

CodeLlama 13B is available from 4 providers. The cheapest input pricing is $0.10/1M tokens from Replicate API.

Who created CodeLlama 13B?

CodeLlama 13B was created by AI at Meta as part of the Code Llama model family.

Is CodeLlama 13B open source?

CodeLlama 13B is open source according to the seed data.

Get Started