GPT-6.1 Sol are now live on CometAPI →
ai-comparisons/CometAPI 리서치

DeepSeek API 요금제: V4 Pro vs. V4.1 Flash

I don’t have real-time access to DeepSeek’s latest pricing or peak/off-peak schedules beyond Oct 2024. “Current” rates can also differ by provider (DeepSeek’s own API vs aggregators like OpenRouter, Together, etc.). If you share the pricing page or specify the provider, I can fill in exact numbers and compute workload costs. In the meantime, here’s a precise checklist and calculator you can plug rates into. What to collect for each model - Provider and endpoint: e.g., DeepSeek official vs OpenRouter (peak/off-peak often differs by provider). - Model IDs: the exact strings you call in the API (e.g., v4-pro vs v4.1-flash; provider-specific IDs may differ). - Context window and output cap: informs max tokens and cache effectiveness. - Prices per 1K tokens: - Off-peak input price (P_in_off) - Off-peak output price (P_out_off) - Peak input price (P_in_peak) or a peak multiplier (M_peak) - Peak output price (P_out_peak) or use same M_peak - Prompt/cache pricing: - Cached-input price per 1K tokens (P_cache_in), if supported - Cache retention and eviction rules (affects reuse rate) - Operational details: - Peak window definition and timezone - Minimum billing increments and rounding - Rate limits, retry policies, and any free tier credits Comparison template (fill with your numbers) - Model IDs: - V4 Pro: <provider’s exact model ID> - V4.1 Flash: <provider’s exact model ID> - Off-peak prices per 1K tokens: - V4 Pro: input P_in_off = <…>, output P_out_off = <…> - V4.1 Flash: input P_in_off = <…>, output P_out_off = <…> - Peak prices per 1K tokens: - Either list P_in_peak/P_out_peak directly, or give M_peak (e.g., 1.25×) - Cache pricing (if available): - P_cache_in (per 1K tokens), note whether output is cache-billed (usually not) - Any first-hit vs subsequent-hit differences - Latency/throughput notes: - V4 Pro: typically higher quality, higher latency/cost - V4.1 Flash: typically lower latency, lower cost, better for high-throughput Real workload cost calculator Define: - N = total requests - f_peak = fraction of requests during peak (0–1) - T_in = average input tokens per request - T_out = average output tokens per request - r_cache = fraction of input tokens served from cache (0–1; subsequent calls that reuse a cached system/prompt segment) - P_in_off, P_out_off = off-peak prices per 1K tokens - P_cache_in = cache price per 1K tokens (if supported) - M_peak = peak multiplier (if peak prices are given directly, use those instead) Per-request off-peak cost: - C_off = (T_in_noncache × P_in_off) + (T_in_cache × P_cache_in) + (T_out × P_out_off) - where T_in_noncache = T_in × (1 − r_cache), T_in_cache = T_in × r_cache Per-request peak cost: - If using multiplier: C_peak = M_peak × C_off - If using separate peak prices: replace P_* with peak equivalents and recompute Total cost: - C_total = N × [ (1 − f_peak) × C_off + f_peak × C_peak ] Operational adjustments to reflect “real” costs - Retries/timeouts: multiply N by (1 + retry_rate) - Token rounding: some providers bill to nearest 1K tokens; apply ceil(T_in/1000) and ceil(T_out/1000) - Streaming partials: still billed as output tokens; ensure T_out includes them - Cache warm-up: first call pays full P_in_off; subsequent calls pay P_cache_in for reused segments - Mixed traffic: compute separate T_in/T_out and r_cache by route (chat vs tool calls), then sum costs How I can finalize this for you - Tell me your provider (e.g., DeepSeek official API vs OpenRouter) and paste the current price table or a link. - Provide your workload profile: N, T_in, T_out, r_cache (if using caching), f_peak, and retry rate. - I’ll return a filled comparison with exact peak/off-peak and cache savings, plus your total monthly cost and per-request cost for V4 Pro vs V4.1 Flash.

CometAPI
Deon GoodwinAI 모델 및 API 리서치 팀
업데이트됨 Oct 2, 2026 7 분 읽기
DeepSeek API 요금제: V4 Pro vs. V4.1 Flash
이 패턴 사용

첫 API 호출하기.

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_COMETAPI_KEY",
    base_url="https://api.cometapi.com/v1",
)

response = client.chat.completions.create(
    model="gpt-5-mini",
    messages=[{"role": "user", "content": "Build this workflow."}],
)

print(response.choices[0].message.content)

요약

DeepSeek V4 Pro와 V4.1 Flash는 서로 다른 가격의 별도 라이브 API 제공 모델입니다. 공식 모델 및 가격 표는 deepseek-v4-pro를 DeepSeek-V4-Pro-0813로, deepseek-flash를 DeepSeek-V4.1-Flash로 명시합니다. DeepSeek의 9월 10일 업데이트는 9월 14일 이후에도 V4 Pro API 서비스를 계속 제공하며 과금 방식은 변하지 않는다고 밝혔습니다. V4 Pro 요청이 Flash로 라우팅되어 Flash 요금으로 청구된다는 이전 주장은 현재 공식 문서로는 더 이상 뒷받침되지 않습니다.

핵심 요약

  • V4.1 Flash는 오프피크 기준 캐시 히트 입력 $0.003/M, 비캐시 입력 $0.15/M, 출력 $0.60/M입니다.
  • V4 Pro는 오프피크 기준 캐시 히트 입력 $0.022/M, 비캐시 입력 $0.66/M, 출력 $1.98/M입니다.
  • 두 모델 모두 피크 요금은 오프피크의 2배입니다. 공식 일정에 따르면 중국 공휴일은 종일 오프피크입니다.
  • V4.1 Flash에는 deepseek-flash, V4 Pro에는 deepseek-v4-pro를 사용하세요. 퇴역한 V4 Flash 모델명만 V4.1 Flash로의 호환 라우팅에 포함되어 있습니다.

현재 DeepSeek API 가격은?

DeepSeek의 최신 가격 표는 각 모델과 시간대별로 캐시 히트 입력, 캐시 미스 입력, 출력 가격을 별도로 나열합니다. 아래 표는 직접 제공자 요금을 그대로 재현한 것이며, 제3자 가격을 설명하지 않습니다.

ModelTime bandCache-hit inputCache-miss inputOutput
V4.1 FlashOff-peak$0.003/M$0.15/M$0.60/M
V4.1 FlashPeak$0.006/M$0.30/M$1.20/M
V4 ProOff-peak$0.022/M$0.66/M$1.98/M
V4 ProPeak$0.044/M$1.32/M$3.96/M

피크 시간은 월요일부터 금요일까지 UTC 01:00–04:00 및 06:00–10:00이며, 중국 공휴일은 제외됩니다. 베이징 시간으로는 평일 09:00–12:00 및 14:00–18:00입니다. 주말과 중국 공휴일은 하루 종일 오프피크입니다. 예산 책정 전 공식 표를 확인하세요. 가격은 변경될 수 있습니다.

DeepSeek API 요금제: V4 Pro vs. V4.1 Flash

공식 DeepSeek V4.1 Flash API 가격. 출처: [DeepSeek release announcement].

현재 DeepSeek V4 Pro 비용은?

DeepSeek의 릴리스 업데이트는 2026년 9월 14일 이후에도 V4 Pro API 서비스를 계속 제공하며 과금 방식은 변하지 않는다고 명시합니다. 최신 모델 표에는 deepseek-flash가 DeepSeek-V4.1-Flash로, deepseek-v4-pro가 DeepSeek-V4-Pro-0813로 나란히 기재되어 있습니다. 따라서 두 이름을 동일 모델의 두 가지 가격 라벨로 간주해서는 안 됩니다.

API model nameModel listed by DeepSeekBilling
deepseek-flashDeepSeek-V4.1-FlashV4.1 Flash rates
deepseek-v4-proDeepSeek-V4-Pro-0813V4 Pro rates
deepseek-v4-flashV4.1 Flash via compatibility routingV4.1 Flash rates

호환성 안내는 사용 중단된 deepseek-v4-flash 및 deepseek-v4-flash-vision-exp에 적용됩니다. 공식 공지에는 deepseek-v4-pro가 해당 라우팅 그룹에 포함되어 있지 않습니다.

DeepSeek V4 Pro vs. V4.1 Flash 가격

Price stateCached inputUncached inputOutput
Former V4 Pro off-peak$0.022/M$0.66/M$1.98/M
V4.1 Flash off-peak$0.003/M$0.15/M$0.60/M
Current V4 Pro route off-peak$0.003/M$0.15/M$0.60/M

비캐시 입력의 경우, V4 Pro의 현재 오프피크 $0.66/M은 Flash의 $0.15/M 대비 4.4배입니다. 출력의 경우 $1.98/M은 $0.60/M 대비 3.3배입니다. 이 직접 제공자 비교에서 Flash는 비캐시 입력이 약 77.3% 저렴하고 출력이 약 69.7% 저렴합니다. 이 퍼센트는 동시에 표에 등재된 두 모델을 비교한 것이며, V4 Pro 라우트가 Flash로 전환되어 절감된 비용을 의미하지 않습니다.

V4 Pro의 $0.022/M 캐시 히트 입력 요금도 현재 요금입니다. Flash의 $0.003/M 캐시 히트 요금은 약 86.4% 낮습니다. 공식 표에는 피크 요금도 모델별로 별도 표시됩니다. V4 Pro는 캐시 미스 입력 $1.32/M, 출력 $3.96/M이며, Flash는 각각 $0.30/M, $1.20/M입니다.

실제 워크로드에서 가격은 무엇을 의미하나?

10,000,000개의 비캐시 입력 토큰과 2,000,000개의 출력 토큰을 사용하는 워크로드를 가정합니다. 현재 직접 제공자 요금 기준:

ModelOff-peak costPeak cost
deepseek-flash$2.70$5.40
deepseek-v4-pro$10.56$21.12

오프피크 계산은 Flash가 10 × $0.15 + 2 × $0.60 = $2.70, V4 Pro가 10 × $0.66 + 2 × $1.98 = $10.56입니다. 현재 시간대 요금 구조에서는 피크 합계가 두 배가 됩니다.

캐싱은 현재 청구액을 더 줄일 수 있습니다. 10,000,000개 입력 토큰 중 80%가 캐시에 적중하면, 오프피크 계산은 캐시 히트 입력 $0.024, 비캐시 입력 $0.30, 출력 $1.20로 합계 $1.524가 됩니다(V4.1 Flash). 동일한 토큰 구성에서 V4 Pro는 $0.176 + $1.32 + $3.96 = $5.456입니다.

DeepSeek API 가용성은 어떻게 변했나?

V4.1 Flash는 공식 DeepSeek API에서 deepseek-flash라는 모델명으로 제공됩니다. 기존 V4 Flash와 V4 Flash Vision Exp 모델은 사용 중단되었지만, 레거시 ID는 일시적으로 수용되어 V4.1 Flash로 라우팅됩니다.

V4 Pro는 deepseek-v4-pro로 계속 제공됩니다. DeepSeek는 과금 방식이 변하지 않았으며 변경 시 별도 공지를 제공하겠다고 밝혔습니다. 오래된 V4 Flash 명칭의 사용 중단을 이유로 V4 Pro의 사용 중단이나 Flash로의 마이그레이션을 추정하지 마세요.

CometAPI 변경 로그에는 DeepSeek Flash와 V4.1 Flash가 기재되어 있습니다. V4.1 Flash 모델 페이지에는 현재 입력 토큰 기준 시작가 $0.12/M가 표시되어 있으며, 별도의 요청 시 조정이 존재합니다. 이는 CometAPI 가격이며 DeepSeek의 직접 API 요금이 아닙니다. 특정 워크로드는 실제 결제 조건을 확인하세요.

DeepSeek V4.1 Pro의 비용은 얼마일 수 있나?

DeepSeek는 향후 V4.1 Pro 출시를 확인했지만, 공식 가격은 발표되지 않았습니다. 오늘 구체적인 숫자를 제시하는 것은 추측에 불과합니다. 더 안전한 계획 방법은 라이브 V4.1 Flash 가격에 대해 여러 프리미엄 시나리오를 모델링하는 것입니다.

Planning scenarioOff-peak uncached inputOff-peak output10M input + 2M output
2× Flash$0.30/M$1.20/M$5.40
3× Flash$0.45/M$1.80/M$8.10
4× Flash$0.60/M$2.40/M$10.80

이는 예산 책정 시나리오이며, 유출 또는 공식 V4.1 Pro 가격이 아닙니다. 현재의 피크/오프피크 구조가 유지되고 Flash 대비 2×–4× 프리미엄이 적용된다고 가정합니다. 회사는 다른 비율, 캐싱 요금, 론치 프로모션, 과금 설계를 선택할 수 있습니다.

과거 V4 Pro 가격은 유용한 상한 기준을 제공합니다. 비캐시 입력 $0.66/M와 출력 $1.98/M는 새로운 V4.1 Flash 오프피크 요금 대비 각각 약 4.4배와 3.3배였습니다. 따라서 2×–4× 시나리오 밴드는 내부 예산에 합리적이지만, DeepSeek의 최종 결정을 예측하지는 않습니다.

팀은 DeepSeek API 비용을 어떻게 통제해야 하나?

  • 의도한 모델을 선택하세요. V4.1 Flash에는 deepseek-flash, 별도 모델인 V4 Pro에는 deepseek-v4-pro를 사용하세요.
  • 유연한 작업은 오프피크로 스케줄링하세요. 나열된 피크 토큰 요금은 오프피크의 두 배이며, 중국 공휴일 예외를 고려하세요.
  • 재사용 가능한 프롬프트 접두부를 유지하세요. 현재 Flash 요금에서 캐시 히트 입력은 비캐시 입력보다 50배 저렴합니다.
  • 출력 길이를 관리하세요. 많은 에이전트 워크로드에서 출력이 여전히 가장 큰 비용 요소입니다.
  • 실시간 가격 표를 확인하세요. 프로덕션 예산 확정 전에 모델별 표와 제3자 약관을 재검증하세요.

FAQ

DeepSeek V4.1 Flash는 얼마인가요?

공식 오프피크 가격은 캐시 히트 입력 $0.003/M, 비캐시 입력 $0.15/M, 출력 $0.60/M입니다. 피크 요금은 각각 $0.006/M, $0.30/M, $1.20/M입니다.

DeepSeek V4 Pro는 여전히 자체 가격으로 이용할 수 있나요?

예. DeepSeek는 V4 Pro API 서비스를 계속 제공하며 과금 방식은 변하지 않았다고 밝혔습니다. 가격 표에는 오프피크 기준 캐시 히트 입력 $0.022/M, 캐시 미스 입력 $0.66/M, 출력 $1.98/M가 나열되어 있으며, 피크 요금은 각각 $0.044/M, $1.32/M, $3.96/M입니다.

V4 Pro와 V4.1 Flash가 현재 동일한 가격인가요?

아니요. 공식 표는 이들을 별도 모델로, 별도 직접 API 요금으로 나열합니다. 오래된 Flash 호환 명칭은 V4.1 Flash로 라우팅되지만, V4 Pro는 그 공지에 포함되어 있지 않습니다.

DeepSeek가 V4.1 Pro 가격을 발표했나요?

현재 공식 모델 및 가격 표에 V4.1 Pro 가격은 없습니다. 목록에 없는 미래 모델에 V4 Pro 가격을 대입하지 마세요.

CometAPI를 통해 DeepSeek V4.1 Flash를 사용할 수 있나요?

예. CometAPI는 Chat API 형식으로 deepseek-flash 및 deepseek-v4.1-flash를 지원합니다. 표시된 가격과 시간대 조정은 DeepSeek의 직접 API 청구와 별개입니다.

결론

핵심 비교 대상은 현재 라이브 DeepSeek API의 두 모델입니다. V4.1 Flash는 오프피크 기준 비캐시 입력 $0.15/M, 출력 $0.60/M에서 시작합니다. V4 Pro는 각각 $0.66/M와 $1.98/M로 계속 나열되어 있습니다. 두 라우트는 공식적으로 동일한 가격을 공유하지 않습니다. 모델의 기능에 맞춰 선택한 다음, 해당 모델의 캐시 및 시간대 요금을 적용해 비용을 계산하세요.

게시 또는 예산 책정 전에 라이브 DeepSeek 가격 표와 릴리스 업데이트를 확인하세요. 제3자 제공사의 오퍼는 해당 제공사의 가격 페이지에서 확인해야 합니다.

SEO 메타데이터

메타 제목: DeepSeek API Pricing: V4 Pro vs. V4.1 Flash

메타 설명: 현재 DeepSeek V4 Pro와 V4.1 Flash API 요율, 모델 ID, 피크/오프피크 가격, 캐시 절감 효과, 실제 워크로드 비용을 비교합니다.

키워드: DeepSeek API price, DeepSeek V4 Pro price, DeepSeek V4.1 Flash price, DeepSeek API pricing, V4 Pro vs V4.1 Flash, DeepSeek V4.1 Pro price

URL 슬러그: deepseek-api-price-v4-pro-vs-v4-1-flash

학습 계속하기

이 글을 다음 결정과 연결하세요.

모든 주제 보기
게시일 Oct 2, 2026
최종 업데이트 Oct 2, 2026
0 회 조회
명확성, 출처 표기 및 최신 API 용어에 대해 검토되었습니다.

더 보기