LLM Reference

Phind Models by Phind

PhindProprietary
2 models2024Up to 32k ctx

Last refreshed 2026-05-19. Next refresh: weekly.

Details

ResearcherPhind
LicenseProprietary
Commercial useCommercial use: conditional
Models2
Released2024
Max context32k

About

Phind's family of large language models excel in code generation, with enhancements over the open-source CodeLlama-34B foundation model. Notably, the Phind Model V7 achieves a HumanEval score of 74.7%, outperforming GPT-4 in coding tasks and operating at five times its speed by leveraging NVIDIA H100 GPUs and the TensorRT-LLM library for rapid processing of up to 100 tokens per second 12. With these advancements, Phind models support extensive context lengths of up to 16,000 tokens, prioritizing user input and web results. Furthermore, earlier versions like Phind-CodeLlama-34B-v2 are open-source on Hugging Face, allowing for independent capability assessments 310.

Decision facts

Best fit
coding
Capability starting point
Phind Instant with 32k context
Lowest tracked input
Not tracked
Closest related family
Phind CodeLlama

Current Variants

Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.

2 in view

Use when the workload needs 32k context.

2024-0232k context
Phind 70BCurrent

Use when the workload needs 32k context and 70B parameters.

2024-0232k context70B parameters

Release Timeline

1 release group
2024-02
2 current
Phind 70B
32k context70B parameters
Current
Phind Instant
32k context
Current

Specifications(2 models)

Phind model specifications comparison
ModelReleasedContextParameters
Phind Instant2024-0232k
Phind 70B2024-0232k70B