जो संख्या मायने रखती है वह है $0.14: DeepSeek के V4-Flash-0731 की प्रति million input tokens कीमत, 31 जुलाई को Hugging Face पर ungated MIT license के साथ जारी, output $0.28 प्रति million, V4-Pro tier की कीमत का लगभग एक-तिहाई। Base है 284 billion parameters का mixture-of-experts, हर token पर 13 billion activated, और 1 million token की context window; model card का 304 billion आँकड़ा attached DSpark speculative-decoding draft module को शामिल करता है। Official line यह है कि preview से architecture और size अपरिवर्तित हैं और gains सिर्फ़ re-post-training से आए हैं।
स्वतंत्र माप, और यह जानबूझकर छोटा है: Simon Willison ने OpenRouter के ज़रिए अपना standard pelican test चलाया और default reasoning effort पर एक निराशाजनक pelican पाया और, उनके शब्दों में, reasoning high रखने पर कुछ काफ़ी बेहतर। Artificial Analysis इस model को MiniMax के M3, एक 428 billion model, से आगे rank करती है और इसे संभवतः उपलब्ध सबसे अच्छा value-per-intelligence बताती है। DeepSeek की अपनी benchmark table बड़ी है और marketing जैसी पढ़ी जाती है; साथ लगी उपयोगी चेतावनी MarkTechPost की है कि agent scores harness-sensitive हैं, इसलिए independent runs अलग हो सकते हैं। इस newsroom की shorthand में: बहुत कुछ claim किया गया है, थोड़ा मापा गया है, और जो मापा गया वह अनुकूल है। इन कीमतों पर खुद चलाना सस्ता रहता है।
उसी दिन launch हुआ MiniMax का H3 सप्ताह का दूसरा भाग है। यह एक omni-modal video model है जो text, images, video और audio को एक unified context की तरह पढ़ता है, और इसका असली differentiator native stereo sound है: audio first-class input है, तो आप इसे एक reference clip देकर किसी image के character से उसे गवा सकते हैं, और यह अलग dubbing stage पर निर्भर रहने के बजाय video के साथ stereo audio emit करता है। Output 2K है, 4 से 15 seconds, सिर्फ़ integer durations। हुड के नीचे: एक captioning scheme जो source material के करीब 100K tokens को 4K में distill करता है, एक नया VAE जिसे वह effective sequence length में 4x gain का श्रेय देता है, और bolt-on super-resolution module के बजाय in-context regeneration step।
Caveats ही उपयोगी हिस्सा हैं। H3 आज केवल MiniMax के API और Hailuo app से उपलब्ध है; open weights 'आने वाले दिनों में' वादा किए गए हैं और प्रकाशन तक मौजूद नहीं हैं, ठीक वैसा वादा जिसे यह newsroom ship होने तक शून्य पर price करती है। Price claims MiniMax के अपने हैं (2K पर mainstream models के एक-तिहाई से कम); third-party trackers pay-as-you-go को करीब $0.13 प्रति second बताते हैं, confirmed नहीं reported। और South China Morning Post, Artificial Analysis का हवाला देते हुए, H3 को video editing में आगे बताती है जबकि text-to-video में Google's Gemini Omni Flash से पीछे और image-to-video में Seedance 2.0 से पीछे रखती है। दो drops, एक pattern: इस हफ़्ते models में दिलचस्प action किनारों पर है, open weights और video, बंद frontier पर नहीं।
