Phind Models by Phind
Last refreshed 2026-05-19. Next refresh: weekly.
Details
About
Phind's family of large language models excel in code generation, with enhancements over the open-source CodeLlama-34B foundation model. Notably, the Phind Model V7 achieves a HumanEval score of 74.7%, outperforming GPT-4 in coding tasks and operating at five times its speed by leveraging NVIDIA H100 GPUs and the TensorRT-LLM library for rapid processing of up to 100 tokens per second 12. With these advancements, Phind models support extensive context lengths of up to 16,000 tokens, prioritizing user input and web results. Furthermore, earlier versions like Phind-CodeLlama-34B-v2 are open-source on Hugging Face, allowing for independent capability assessments 310.
Decision facts
- Best fit
- coding
- Capability starting point
- Phind Instant with 32k context
- Lowest tracked input
- Not tracked
- Closest related family
- Phind CodeLlama
Current Variants
Use-when guidance is based on each model's tracked capabilities, context window, release date, and replacement status.
Use when the workload needs 32k context and 70B parameters.
| Model | Use when | Released | Signals | Status |
|---|---|---|---|---|
| Phind Instant | Use when the workload needs 32k context. | 2024-02 | 32k context | Current |
| Phind 70B | Use when the workload needs 32k context and 70B parameters. | 2024-02 | 32k context70B parameters | Current |
Release Timeline
1 release groupSpecifications(2 models)
| Model | Released | Context | Parameters |
|---|---|---|---|
| Phind Instant | 2024-02 | 32k | — |
| Phind 70B | 2024-02 | 32k | 70B |

