{"rewrite":{"id":"r_708f236342ac84cd2b691904","clusterId":"c_c86660f18e201c1c1e72b509","slug":"deepseek-v4-flash-official-release-makes-weights-free-for-commercial-use","model":"deepseek-v4-flash:free","headline":"DeepSeek-V4-Flash Official Release Makes Weights Free for Commercial Use","summary":"DeepSeek has released the official version of its AI model DeepSeek-V4-Flash-0731, replacing the earlier preview version. The weights are available under the MIT license, allowing free commercial use. The model is a Mixture-of-Experts type with 284 billion total parameters, activating only 13 billion during inference, and supports a context length of 1 million tokens. According to benchmark scores, it surpasses the preview version of the higher-tier DeepSeek-V4-Pro, which is more than five times larger, across all nine agent-based benchmark items. Among open models, it also outperformed GLM-5.2 from China in all eight items where scores were published, though it does not reach Anthropic's Claude Opus 4.8 in any item. Community quantized versions are available, including GGUF-format versions from Unsloth in 13 steps ranging from 1-bit at 82.5GB to 8-bit at 162GB. For paid cloud usage, DeepSeek's documentation lists the price at $0.14 per million tokens for input, $0.0028 per million tokens for cached input, and $0.28 per million tokens for output, with prices doubling during peak times.","whyItMatters":"The release came hours after OpenAI cut GPT-5.6 Luna prices by 80 percent, and DeepSeek has now notified users of a significant price increase for the model, suggesting the low-cost strategy drew demand that strained its computing resources.","webCardHtml":"\u003cp\u003e\u003c/p\u003e\u003cul\u003e\n  \u003cli\u003eDeepSeek notified users of the price increase on August 6, 2026, without specifying the new amount, and asked them to plan usage accordingly.\u003c/li\u003e\n  \u003cli\u003eThe price hike follows reports of the model\u0026#39;s inference speed occasionally slowing to extreme levels as demand surged, with DeepSeek\u0026#39;s API status log showing frequent access difficulties around the release.\u003c/li\u003e\n  \u003cli\u003eCEO Liang Wenfeng said in July 2026 that DeepSeek\u0026#39;s computing resources equal about 20,000 NVIDIA H100 GPUs, which some reports say the demand spike has strained.\u003c/li\u003e\n  \u003cli\u003eBloomberg reports DeepSeek is in the middle of a large fundraising round, targeting about $8 billion, with a post-round valuation near $74 billion, and a possible IPO within 2026.\u003c/li\u003e\n  \u003cli\u003eThe fundraising is the second round, resumed on August 6 after trading was temporarily halted in late July following the leak of the CEO\u0026#39;s investor remarks on US-China AI competition.\u003c/li\u003e\n  \u003cli\u003eOpenAI cut GPT-5.6 Luna prices by 80 percent on July 31, bringing input to $0.20 per million tokens and output to $1.20 per million tokens, hours before DeepSeek\u0026#39;s release.\u003c/li\u003e\n  \u003cli\u003eOpenAI said the cut came from extracting further architectural efficiency, though some observers called it a pricing strategy aimed at Chinese AI labs.\u003c/li\u003e\n  \u003cli\u003eThe price-cut GPT-5.6 Luna became more cost-efficient than closed models like Claude Sonnet 5 and Gemini 3.6 Flash, as well as Chinese open models including DeepSeek V4 Pro and GLM-5.2.\u003c/li\u003e\n\u003c/ul\u003e\u003cp\u003e\u003c/p\u003e","blueskyPost":"DeepSeek-V4-Flash-0731 beats the 5x-larger V4-Pro preview on all nine agent benchmarks, but the price hike notice came a week after launch.","twitterPost":"OpenAI cut GPT-5.6 Luna prices 80%. Hours later DeepSeek shipped a cheaper model. Now DeepSeek says prices are going up. The AI price war has a new front.","threadsPost":null,"newsletterBlurb":"DeepSeek released the official DeepSeek-V4-Flash-0731 with MIT-licensed weights, beating larger models on benchmarks. Days later, the company announced a significant API price increase amid reported demand surges.","attributionJson":"[{\"source\":\"GameBusiness.jp\",\"url\":\"https://www.gamebusiness.jp/article/2026/08/06/27624.html\",\"title\":\"Official Release of \\\"DeepSeek-V4-Flash\\\": Weights Free for Commercial Use, Quantized Versions from 82.5GB (Generative AI Close-Up)\"},{\"source\":\"GIGAZINE\",\"url\":\"https://gigazine.net/news/20260807-deepseek-raise-prices/\",\"title\":\"DeepSeekがAPI料金の大幅値上げを予告、低価格戦略による需要急増が背景か\"}]","lintFlagsJson":null,"lintHits":0,"costUsd":0,"inputTokens":13378,"outputTokens":1414,"status":"published","repairAttempts":0,"nextRepairAt":null,"factsAttemptedAt":1786085661,"createdAt":"2026-08-07T06:40:24.000Z","publishedAt":"2026-08-07T06:51:44.000Z","updatedAt":"2026-08-07T06:40:24.000Z"},"cluster":{"id":"c_c86660f18e201c1c1e72b509","canonicalTitle":"「DeepSeek-V4-Flash」の正式版、重みが商用利用可能で無料公開。量子化版は82.5GBから（生成AIクローズアップ）","representativeArticleId":"a_e42b3665eb314215831be379","sourceCount":2,"writtenSourceCount":2,"writeAttempts":0,"isSolo":false,"entitiesJson":"{\"anime_titles\":[],\"manga_titles\":[],\"work_titles\":[\"DeepSeek-V4-Flash\"],\"studios\":[],\"people\":[],\"type\":\"news\",\"domain\":\"other\",\"is_roundup\":false}","contentType":"news","status":"published","firstSeenAt":"2026-08-05T21:30:04.000Z","lastSeenAt":"2026-08-07T06:25:00.000Z","updatedAt":"2026-08-07T06:51:45.000Z"},"attribution":[{"source":"GIGAZINE","url":"https://gigazine.net/news/20260807-deepseek-raise-prices/","title":"DeepSeekがAPI料金の大幅値上げを予告、低価格戦略による需要急増が背景か"},{"source":"GameBusiness.jp","url":"https://www.gamebusiness.jp/article/2026/08/06/27624.html","title":"「DeepSeek-V4-Flash」の正式版、重みが商用利用可能で無料公開。量子化版は82.5GBから（生成AIクローズアップ）"}],"entities":{"anime_titles":[],"manga_titles":[],"work_titles":["DeepSeek-V4-Flash"],"studios":[],"people":[],"type":"news","domain":"other","is_roundup":false},"keyFacts":["DeepSeek released DeepSeek-V4-Flash-0731 as an official open model under a commercially usable license on July 31, 2026.","The model has 284 billion total parameters, activates 13 billion during inference, and supports a 1 million token context length.","Unsloth released GGUF-format quantized versions of DeepSeek-V4-Flash in 13 steps from 1-bit at 82.5GB to 8-bit at 162GB.","DeepSeek's API pricing for deepseek-v4-flash is $0.14 per million input tokens, $0.0028 per million cached input tokens, and $0.28 per million output tokens, doubling during peak times.","DeepSeek notified users on August 6, 2026 of a significant price increase for DeepSeek-V4-Flash-0731 without specifying new amounts."]}
