Observed Signal · May 25, 2026 · Pricing Update · Source: t3n · Impact: 3/5 · Sentiment: Positive

Deepseek Makes V4‑Pro 75% Price Cut Permanent

Executive Signal Summary

Chinese AI developer Deepseek announced that the temporary 75% discount on its new flagship model, Deepseek V4‑Pro, is now permanent. The company said API input-token prices for V4‑Pro are reduced to between $0.003625 and $0.435 per million (previously $0.0145–$1.74), and output-token costs to $0.87 per million (previously $3.48). The move places Deepseek substantially below Western rivals — OpenAI’s GPT‑5 and Anthropic’s Claude Opus 4.7 charge several dollars per million tokens — and aims to win enterprise customers that need very large context windows. Observers note potential cost savings for large users (e.g., Salesforce) but warn of geopolitical and technical risks for US and European companies sending sensitive data to a Chinese provider.

Polaris7 AgentPolaris7 Strategic Assessment
High Confidence

A permanent large price cut for a foundation-model provider materially changes cost dynamics for enterprise AI usage and could influence vendor selection and TCO for AI-enabled MarTech applications, though geopolitical adoption limits blunt its immediate industry-wide impact.

SIGNAL RADAR

Track DeepSeek Signals & Market Shifts in Real-Time

Polaris7 autonomous intelligence agents track regulatory filings, primary sources, executive changes, and deal flow 24/7. Create your free Explorer workspace to monitor these entities.

Start Free in Explorer
Free Explorer tierNo credit card requiredInstant watchlist setup

Key Takeaways & Evidence Grounding

  • Deepseek announced the V4 model family (V4‑Pro and V4‑Flash) on 2026-04-24.
  • Deepseek made an initial 75% temporary discount for V4‑Pro and has now made that 75% price reduction permanent.
  • Deepseek V4‑Pro API input-token prices per 1M tokens now range $0.003625–$0.435 (previously $0.0145–$1.74); output-token price is $0.87 per 1M (previously $3.48).
  • OpenAI’s GPT‑5 charges ~$2.50 per 1M input and $10 per 1M output; Anthropic’s Claude Opus 4.7 charges ~$5 per 1M input and $25 per 1M output, per The Next Web.
  • Deepseek is positioning V4‑Pro at enterprise customers that require very large context windows and long-context processing.
Primary Source Grounding & Direct Attribution
Direct Origin Attribution
Primary Reporting: t3n•Published: May 25, 2026
Original Coverage Title: “Deepseek: Aggressiver Preisnachlass setzt Anthropic und OpenAI unter Druck”

Related Market Signals & Shifts

Recent verified developments and strategic activity across this market segment.

Large Language Models (LLM) & AIJun 4, 2026

DeepSeek Enters U.S. Corporate AI Spending

Chinese AI model provider DeepSeek has begun appearing in U.S. corporate vendor payments, according to Ramp’s June 2026 trending vendors list. The move follows DeepSeek’s May decision to make a 75% price cut on its V4‑Pro model permanent, setting cached-input pricing at RMB 0.025 (about $0.0035) per million tokens — roughly 1% of comparable cached-input costs cited for Anthropic. Ramp’s trending data contrasts with earlier developer-side signals from OpenRouter, which in May showed Chinese-built models dominating developer routing and weekly token volumes rising to ~25 trillion. Ramp cautioned that a monthly trending list is not conclusive market-share proof, but the appearance of other inference platforms (Fireworks AI, fal AI, DeepInfra) and GPU provider Vast.ai on the list suggests a growing cost-optimization infrastructure around alternatives to premium U.S. providers. Publication date: 2026-06-04.

Read assessment
Model ReleaseDec 5, 2025

DeepSeek V3.2 Matches Gemini-3, Cuts Costs

DeepSeek released V3.2 and V3.2-Speciale, claiming frontier-level reasoning that rivals Gemini-3.0-Pro and GPT-class “High” models while dramatically lowering inference costs. The team says a new attention mechanism reduces token-processing cost by about 70% (pricing example: processing 128,000 tokens now ~ $0.70 per million tokens vs $2.40 previously). V3.2 preserves reasoning across multiple external tool calls, improving agent/workflow reliability, and the Speciale variant achieved top scores on several competitive reasoning contests. Benchmarks reported include AIME 2025 and Terminal Bench 2.0 where DeepSeek variants compare favorably to GPT-5-High on math and coding agent tasks. DeepSeek also made models freely available under an MIT license, though it acknowledges token efficiency and world-knowledge remain behind some proprietary frontiers. The newsletter also summarizes practitioner advice on AI pricing, emphasizing retention over pure price points.

Read assessment
Large Language Models (LLM) & AIApr 24, 2026

DeepSeek previews V4 open-source LLM

Deepseek on April 24, 2026 published its long‑anticipated Deepseek V4 (variants Pro and Flash), an open‑source large language model built on a new architecture with 1.6 trillion parameters. The company highlights significant gains in reasoning and autonomous code generation, claims benchmark-leading performance in mathematics, STEM and programming among open models, and says V4 supports context windows up to one million tokens while reducing compute and memory costs. Deepseek positions V4 Pro as materially cheaper on coding tasks versus OpenAI’s GPT‑5.5. The rollout also involves a partnership with Huawei, which supplies "Supernode" clusters of Ascend‑950 chips; Deepseek and analysts note a strategic focus on Huawei and Cambricon domestic chips to relieve reliance on Nvidia/AMD. Market reaction is expected to be more muted than Deepseek’s earlier 2025 breakthrough R1 shock.

Read assessment

Track Real-Time Market Signals & Shifts

Set up custom watchlists to receive automated, evidence-grounded executive digests whenever material signals or shifts occur across your tracked landscape.