O

LangMart: DeepSeek: R1 Distill Qwen 32B

Openrouter
64K
Context
$0.2400
Input /1M
$0.2400
Output /1M
N/A
Max Output

LangMart: DeepSeek: R1 Distill Qwen 32B

Model Overview

Property Value
Model ID openrouter/deepseek/deepseek-r1-distill-qwen-32b
Name DeepSeek: R1 Distill Qwen 32B
Provider deepseek
Released 2025-01-29

Description

DeepSeek R1 Distill Qwen 32B is a distilled large language model based on Qwen 2.5 32B, using outputs from DeepSeek R1. It outperforms OpenAI's o1-mini across various benchmarks, achieving new state-of-the-art results for dense models.\n\nOther benchmark results include:\n\n- AIME 2024 pass@1: 72.6\n- MATH-500 pass@1: 94.3\n- CodeForces Rating: 1691\n\nThe model leverages fine-tuning from DeepSeek R1's outputs, enabling competitive performance comparable to larger frontier models.

Description

LangMart: DeepSeek: R1 Distill Qwen 32B is a language model provided by deepseek. This model offers advanced capabilities for natural language processing tasks.

Provider

deepseek

Specifications

Spec Value
Context Window 64,000 tokens
Modalities text->text
Input Modalities text
Output Modalities text

Pricing

Type Price
Input $0.24 per 1M tokens
Output $0.24 per 1M tokens

Capabilities

  • Frequency penalty
  • Include reasoning
  • Max tokens
  • Min p
  • Presence penalty
  • Reasoning
  • Repetition penalty
  • Response format
  • Seed
  • Stop
  • Structured outputs
  • Temperature
  • Top k
  • Top p

Detailed Analysis

DeepSeek R1 Distill Qwen 32B - Variant Analysis