Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

According to their post:

  (Input / Output / Cache Read, [$/M])
  DeepSeek-V4-Flash: 
    Prev: 0.14 / 0.28 / 0.0028 
    Off-Peak: 0.22 (1.6x) / 0.66 (2.4x) / 0.007 (2.5x)
    Peak: 0.44 (3.1x) / 1.32 (4.7x) / 0.014 (5.0x)

  DeepSeek-V4-Pro: 
    Prev: 0.435 / 0.87 / 0.003625
    Off-Peak: 0.66 (1.5x) / 1.98 (2.3x) / 0.022 (6.1x) 
    Peak: 1.32 (3.0x) / 3.96 (4.6x) / 0.044 (12.1x)
gpt-5.6-luna: $0.20 / $1.20 / $0.02 / $0.25 (In / Out / Cache Read / Cache Write)

EDIT: formatting

EDIT2: giving up on the formatting :-/



> EDIT: formatting

Keep at it, I believe in you.


My apologies to any mobile users, but for the desktop folk:

  Provider, Model    Billing       Input           Output          Cache read       Cache write
  DeepSeek                                                                            
    V4-Flash         Old           $0.1400         $0.2800         $0.0028          -      
    V4-Flash         New Off-Peak  $0.2200 (1.6x)  $0.6600 (2.4x)  $0.0070  (2.5x)  -      
    V4-Flash         New Peak      $0.4400 (3.1x)  $1.3200 (4.7x)  $0.0140  (5.0x)  -      
    V4-Pro           Old           $0.4350         $0.8700         $0.0036          -      
    V4-Pro           New Off-Peak  $0.6600 (1.5x)  $1.9800 (2.3x)  $0.0220  (6.1x)  -      
    V4-Pro           New Peak      $1.3200 (3.0x)  $3.9600 (4.6x)  $0.0440 (12.1x)  -      
                                                                                      
  OpenAI                                                                              
    GPT-5.6 Sol                    $5.0000         $30.000         $0.5000          $6.2500
    GPT-5.6 Terra                  $2.0000         $12.000         $0.2000          $2.5000
    GPT-5.6 Luna                   $0.2000         $1.2000         $0.0200          $0.2500
                                                                                      
  Anthropic                                                                           
    Claude Fable 5                 $10.000         $50.000         $1.0000          $12.500
    Claude Opus 5                  $5.0000         $25.000         $0.5000          $6.2500
    Claude Sonnet 5                $2.0000         $10.000         $0.2000          $2.5000
                                                                                      
  Moonshot                                                                            
    Kimi K3                        $3.0000         $15.000         $0.3000          -      
                                                                                      
  Z.AI                                                                                
    GLM 5.2                        $1.4000         $4.4000         $0.2600          -


Seems to still be cheaper at worst case scenario on a pay as you schedule. The price increase isn't ideal, but still seems like a good deal to me.


There are many nuances like amount of cache hit, the time of the day you use it and how other providers or even competitors respond, so it is hard to say concretely, but it might change my monthly usage close to one of the $20 USD plans.

I’m curious about how openrouter and Luna prices will change in response.


Yeah, seems the same for me! Except the subscription plans you can get like Kimi K3 or GLM Coding Plan are still giving you more value if you need that many tokens (and can use those plans).


Forgive my naivety... What are the units? $1.4000 for how many input tokens?


Everything is per 1 million tokens.


TIL: formatting tables on HN is computationally impossible. lol

p.s. Thanks DSv4-Flash, for your hard work of converting a messy table into plain text.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: