Using OctoML CodeLlama-70b-Instruct on OctoML (Deprecated)

Implementation guide · Code Llama · AI at Meta

ServerlessOpen Weights

OctoML (Deprecated) exposes OctoML CodeLlama-70b-Instruct through model ID octoml-codellama-70b-instruct. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-07-09. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at OctoML (Deprecated) and generate an API key.
  2. 2
    Use the OctoML (Deprecated) SDK or REST API to call octoml-codellama-70b-instruct — see the documentation for request format.
  3. 3
    You'll be billed $0.40/1M input, $0.60/1M output tokens. See full pricing.

Code Examples

See OctoML (Deprecated) documentation for integration details.

Pricing on OctoML (Deprecated)

TypePrice (per 1M)
Input tokens$0.40
Output tokens$0.60

Capabilities

No model capability flags are currently sourced.

About OctoML CodeLlama-70b-Instruct

OctoML CodeLlama-70b-Instruct is Meta's Code Llama model. It offers a 100K-token context window with weights openly available for self-hosting.

Model Specs

Released2023-07-18
Parameters70B
Context100k
ArchitectureDecoder Only
Knowledge cutoff2022-09

Provider

OctoML (Deprecated)

OctoML

Seattle, Washington, United States