Snorkel Mistral PairRM
Snorkel Mistral PairRM is worth evaluating for general LLM work when its provider route and context window match the workload.
Use it for
- Teams evaluating general LLM work
- Workloads that can use a 32k context window
- Buyers comparing 2 tracked provider routes
Do not use it for
- Vision or document-understanding workloads
- Strict JSON or tool-calling flows
- Family
- Snorkel
- Released
- 2023-11-15
- Context
- 32k
- Parameters
- 7B
- Architecture
- Decoder Only
- Specialization
- general
- Openness
- Proprietary
- Weights
- Not released
- Code
- Unknown
- Training
- Fine-tuned
Cheapest of 2 routes · Fireworks AI
About
The Snorkel Mistral PairRM-DPO is a chat-optimized large language model, leveraging the Mistral-7B-Instruct-v0.2 architecture. Designed to interpret and respond efficiently to user inputs, it employs Direct Preference Optimization alongside the Pairwise Reward Model (PairRM) to enhance its alignment with human preferences. Exclusively trained on the UltraFeedback dataset without input from other LLMs, it excels in generating text for conversational contexts, ranking third on the AlpacaEval 2.0 leaderboard at 30.22. Post-processing with PairRM-best-of-16 boosts its score to 34.86.
Top use-case fit
No primary decision-task fit is mapped for this model yet.
Provider price ladder
Compare all 2Compare API pricing across 2 providers for input and output tokens, batch, and cached reads when available.
| Provider | Input / 1M | Output / 1M | Route |
|---|---|---|---|
| Fireworks AI | $0.200 | $0.200 | Provisioned |
| Together AI | $0.200 | $0.200 | Serverless |
Available via routers & gateways(1)
Capabilities
No model capability flags are currently sourced.
Benchmark peer barsfor Coding
No task-mapped benchmark peers are available for this model yet.
Migration checks
No linked migration route is available for this model yet.
Cheapest of 2 routes · Fireworks AI