Using Nous Hermes 13B on Lepton AI API

Implementation guide · Hermes · Nous Research

ServerlessOpen Source

Lepton AI API exposes Nous Hermes 13B through model ID nous-hermes-13b. Use the setup steps, sourced pricing, capabilities, and official provider links below to validate this route before deployment.

Last refreshed 2026-10-02. Next refresh: weekly.

Quick Start

  1. 1
    Create an account at Lepton AI API and generate an API key.
  2. 2
    Use the Lepton AI API SDK or REST API to call nous-hermes-13b — see the documentation for request format.
  3. 3
    You'll be billed $0.13/1M input, $0.13/1M output tokens. See full pricing.

Code Examples

See Lepton AI API documentation for integration details.

Pricing on Lepton AI API

TypePrice (per 1M)
Input tokens$0.13
Output tokens$0.13

Capabilities

No model capability flags are currently sourced.

About Nous Hermes 13B

The Nous Hermes 13B is a cutting-edge large language model fine-tuned on over 300,000 instructions and built upon the Llama 13B architecture. Developed by Nous Research, with significant contributions from Teknium, Karan4D, and Redmond AI, the model excels in generating coherent long-form responses and performs comparably to GPT-3.5-turbo. It avoids OpenAI's content censorship, offering more open dialogues. Additionally, Nous Hermes 13B achieves top rankings in various benchmark tests and is available for download on Hugging Face, with future plans for model quantization and further benchmarking 123.

Model Specs

Released2023-12-15
Parameters13B
ArchitectureDecoder Only

Provider

Lepton AI API

Lepton AI

Sacramento, California, United States