frontpage.
newsnewestaskshowjobs

Open Source @Github

fp.

Open in hackernews

DeepSeek launches v4.1 flash, postpones v4 Pro discontinuation

2•Audiolite•8h ago
DeepSeek has officially released the V4.1 Flash model on September 10, 2026 (Beijing Time). In the meanwhile, we plan to postpone the discontinuation of the V4 Pro service to 12:00 Beijing Time on September 14, 2026. At that time, all requests to the Pro model will be routed to V4.1 Flash and billed at Flash's price. After extensive internal and external testing, V4.1 Flash has comprehensively surpassed V4 Pro across all key metrics, including performance, cost, speed, and task completion time. If you encounter any issues during your comparative testing between V4 Pro and V4.1 Flash, please do not hesitate to reach out to us with your feedback.

The price of V4 Pro remains unchanged during the service period. The price of V4.1 Flash will take effect at 4:00 UTC, September 10, 2026.

Peak hours (in UTC): 1:00–4:00 AM and 6:00–10:00 AM, Monday to Friday. (All other hours are off-peak) If you continue to use our services after the billing adjustment, you will be deemed to have accepted the adjusted billing terms. If you do not agree, you may choose to cancel your service and apply for a refund. Should you have any questions or require further information, please do not hesitate to contact us.

Comments

AlexWApp•3h ago
I m not sure I would describe this as blanket win over GPT-5.6 Sol. In DeepSeek’s own table, V4.1 Flash is ahead on Terminal-Bench 2.1, DeepSWE, NL2Repo, and AutomationBench, but it is behind on GPQA Diamond, Terminal-Bench 3.0 and 4.0, and SEC-Bench Pro.

The architecture is probably part of the explanation for the lower cost and faster inference. DeepSeek says V4.1 Flash uses a new Causal Encoder–Decoder design, with 8B active parameters for input processing and 16B for decoding, along with much smaller KV caches.

But I hope it is just not benchmaxxed and genuinely good model

benchmarks: https://media2url.com/m/52a77a33347c48

Google to discontinue allowing send as an external account in Jan. 2027

4•AuthorizedCust•1h ago•0 comments

DeepSeek launches v4.1 flash, postpones v4 Pro discontinuation

2•Audiolite•8h ago•1 comments

Anyone miss a flip phone with buttons? Anyone?

3•wwolfson97•8h ago•9 comments

Ask HN: Fable hacked my piano, can I release the results?

297•jmpman•4d ago•164 comments

Ask HN: How do you manage skills files?

312•imadtaieber•3d ago•291 comments

Ask HN: Why can I only downvote select comments?

6•socalgal2•12h ago•9 comments

Apparently CodePen 2.0 sends data to their servers as you type

113•maxim-fin•3d ago•59 comments

Ask HN: Software Licenses that prevent LLMs from training on open source?

7•mattm•21h ago•5 comments

Ask HN: Would you read a statistics textbook?

107•usernametaken29•4d ago•67 comments

Ask HN: Next day thoughts on the millenium prize threads?

3•seizethecheese•18h ago•0 comments

Reddit saves your keystrokes in text boxes

17•thomasjeff1•22h ago•4 comments

Ask HN: Do you regret using Social Networking apps?

7•julienreszka•20h ago•9 comments

Ask HN: Show your micro-SaaS

25•genekrapivin•3d ago•23 comments

Ask HN: Is there any solution to the AI infestation?

5•yathern•21h ago•6 comments

Ask HN: How to know when Claude has changed model itself during processing?

2•mfaisalghufran•22h ago•1 comments

FinOps for LLM Spend – why your bill has changed

2•leanroute_ai•22h ago•0 comments

Ask HN: What non-IT/Software industries have benefitted the most from AI?

2•simonsarris•1d ago•5 comments

Ask HN: 3.5 inch diskette read errors, would a period correct drive do better?

34•rietta•3d ago•42 comments

Is Google planning for Gemini 4 rather than 3.5 pro?

4•Qhloi•1d ago•1 comments

Ask HN: Are you leveraging Spec-Driven / Spec-Anchored development?

6•locusofself•1d ago•3 comments

Ask HN: Any Software Engineers here who enjoy their AI-native dev workflow?

13•pkos98•3d ago•13 comments

Ask HN: What's your AI coding set up?

6•asxndu•13h ago•3 comments

Ask HN: Connecting Kubernetes dependencies to application telemetry

6•chipfixer•3d ago•3 comments

Ask HN: Is ageism in tech still a problem in 2026?

4•leonagano•1d ago•2 comments

Tell HN: OpenAI brings back 5 hour limit for plus and business standard users

128•spwa4•2d ago•147 comments

Tell HN: I want to see the same moon as you

17•UnderABlueMoon•5d ago•16 comments

Ask HN: Would you hire a new engineer, or get 300k worth of tokens for the team?

5•websap•1d ago•9 comments

Ask HN: UK Rescue Rocket Sheds/Houses Information

24•burnt-resistor•3d ago•6 comments

The universal programming language of LLMs

5•senorqa•1d ago•9 comments

Ask HN: (Why) Was the LLM breakthrough useful for images, audio, etc.?

5•rogerrogerr•3d ago•5 comments