GPT-5.6 Sol Pricing Cut by 50%

(openrouter.ai)

100 points | by Topfi 4 hours ago

11 comments

  • CompoundEyes 46 minutes ago
    I used over a billion tokens per day of gpt-5.6 sol xhigh starting last Wednesday through Sunday before reaching my reset limit. The $200 pro plan is still the best deal.
    • ec109685 2 minutes ago
      I spent $800 in a few hours when my sub maxed out because I was trying to get something done and had a long car ride to let it churn.

      Their api pricing is absurdly expensive.

    • infinite_spin 20 minutes ago
      I have mine churning like butter and I'm rarely hitting a billion tokens per day, what's your workflow look like?
      • ramraj07 2 minutes ago
        Some people just do crazy stuff. For example this now ex yc guy who said he has agents constantly scanning Sf govt apis and forming dashboards just because
    • SJMG 11 minutes ago
      A billion a day? How many agents are you running?
    • jm4 31 minutes ago
      I can’t sign up for that. I tried authorizing Codex a couple days ago. For some reason, their system says my phone number has been used for verification 3 times even though it definitely has not. I’ve had this phone number for over 20 years. OpenAI support is useless. They just keep repeating the policy without actually helping me.
  • Fergusonb 1 hour ago
    Luna saw a huge jump after the price cut and is one of the more competitive models at the new price on openrouter.

    Maybe they want to see how much market they can grab with Sol?

    This might help but there are already cheaper models with Sol's intelligence more or less, the most notable being Grok 4.6 at $6/m which makes it a tougher sell

    • xmonkee 1 hour ago
      It's really only between Anthropic and OpenAI for many of my use cases, since I have a Zero Data Retention agreement with both. I'm not trusting random inference providers and especially not Elmo with sensitive data.
    • OutOfHere 1 hour ago
      Since when does Grok 4.6 have Sol 5.6's intelligence? I don't believe it.
      • redox99 12 minutes ago
        It doesn't.
      • dimgl 47 minutes ago
        Why not?
        • chaos_emergent 36 minutes ago
          Because it’s good on benchmarks but not on real usage?
          • dimgl 31 minutes ago
            But OP said they've never used it. How would they know?
      • mohamedkoubaa 1 hour ago
        I wonder if xAI is A/B testing routing some difficult grok 4.6 queries to Sol to seed some true believers.
    • jLaForest 1 hour ago
      I don't care how capable or how cheap Grok is, I refuse to financially support a company owned by a white supremacist that is actively working to disenfranchise me and millions of my fellow citizens.
      • bko 58 minutes ago
        Give it a rest dude.
  • ComputerGuru 15 minutes ago
    Does OpenRouter eat this cost to get their hands on a copy of the conversations people are using with the model?
  • netsec_burn 18 minutes ago
    After using Claude for a long time, I tested Sol 5.6 for the first time today. Love it, its an incredibly capable model and uses far fewer tokens/time thinking. Its what I imagine Fable would be if I haven't been downgraded on every conversation - even after completing the verification program. I think I may cancel my Claude subscription finally.
    • ec109685 1 minute ago
      Fable is still the best there is. Sol close second but I find it gets way to stuck on details.

      Also Opus 5 is fine if your codebase is simple.

    • Aargau 5 minutes ago
      I'm also in Anthropic Cyber Verification Program, but they specifically exclude Fable, just goes up to Opus 5.

      I hear you on the downgrades, I'm 13/13 on downgrades, and last downgraded me to Sonnet for asking for reasoning chain.

    • Razengan 10 minutes ago
      I recently tried Claude again after several months, to see if it was any better at something Codex has been struggling with…

      They STILL don't have an option to "Sign in with Apple" on the website, but they do for Google??!? (and on iPhone of course)

      Screw that asinine UX

      (and no it wasn't better than Codex at this particular task)

  • m4rtink 40 minutes ago
    Price wars did wonders for many businesses, like the bike sharing industry in China.

    Overgrown datacenters or mounds of GPUs dumped into the harbour next ?

    • Moto7451 33 minutes ago
      I would in such a scenario expect the GPUs to be dumped to industrial breakers who would send them to China for refurbishment and repackaging before being sold again on Amazon, AliExpress, and Taobao as last gen gaming cards from weird brands and specs.

      This is what happened after the great crypto GPU dumping.

      • m4rtink 23 minutes ago
        Yeah, I ment it as a joke - I agree with you. Watched the Gamers Nexus GPU investigation recently, where they were shown how a chinese soldering shop can transplant GPU chips to a new board, including memory chip reuse.

        Hopefully we can look forward to all that useless datacenter AI crap gets repurposed in a similar manner into something actually useful for users.

        • kajaktum 16 minutes ago
          I wouldnt be so hopeful because they dont use commodity hardware afaik How are you going to use a h100 at home?
  • josh-wrale 1 hour ago
    Is this motivated by the value of the thinking traces gleaned from the traffic?
    • ec109685 0 minutes ago
      They can’t decrypt the thinking traces.
  • dgunay 40 minutes ago
    I'm loving this race to the bottom.
    • infinite_spin 13 minutes ago
      I'm not having that experience. So far each major model update has been at least slightly better than the last, in ways I've found useful. Can't say it's perfect, or able to do exactly what I want without a decent amount of instruction/implementation/docs, but it's been useful enough to keep paying for it.
      • fn-mote 5 minutes ago
        GP means race to the bottom in price not quality.
  • z_rho_one 36 minutes ago
    If they can cut the price of Sol by 50% and the price of Luna by 80%, then the original price might have carried a massive operating margin. They might still be serving the models at a profit after these price cuts, but we will never know.
    • paxys 26 minutes ago
      I don’t think there’s a real answer for this. It depends on whatever number the accounting department wants to make up.

      Do you include research costs? Of all models or only specific ones? What percent of the R&D budget do you allocate to model serving? What about data center capacity? Do you count future commitments? All the circular financing deals? Employee equity grants?

    • wahnfrieden 26 minutes ago
      OpenAI didn't cut the price of Sol by 50% like they did with Luna's 80%. Sol was unchanged. This is just a limited promo for OpenRouter non-BYOK.
  • tartakovsky 35 minutes ago
    No ZDR. No dice.
  • OutOfHere 1 hour ago
    The title looks to be misleading, since this price cut is limited to OpenRouter. It does not apply for the native OpenAI price listed at https://developers.openai.com/api/docs/models/gpt-5.6-sol
    • matchagaucho 38 minutes ago
      Right? Should we switch from direct OpenAI API integration to OpenRouter?

      What's the incentive here?

      Open Responses API doesn't appear to support state management (yet)

    • paxys 46 minutes ago
      Which raises the question - who is subsidizing this, and why?
  • vorpalhex 1 hour ago
    Do other people find 5.6 to be worse at most simple tasks and frequently over complicate things?

    I asked it to write a user todo and it turned out a four page essay. I gave the same task to 5.4 and got the small list of checkboxes I expected.

    • drdexebtjl 11 minutes ago
      You would probably get better results with Luna for the real simple tasks, or Sol with low thinking effort.

      I find that I get exactly the effort that I asked for, which is pretty nice. The other side of that coin is that these are the least lazy models I’ve used so far. They will go on elaborate tangents to complete the task when I want them to.

    • infinite_spin 11 minutes ago
      I've found it's worse for simple tasks too, and I have to give it stricter guidelines, and sometimes it doesn't follow the same patterns I've grown to expect. I've found using 5.6 (sol) is good for diagnosing issues though, especially in terms of optimization of some given path
    • qup 35 minutes ago
      I've found it to be great for planning code changes (or new projects). I use the superpowers plug-in which I think guides the planning.

      Then I switch models (to luna) before implementation. I find this combo nearly always does what I want.

      I also use a skill called ponytail, its goal is to keep things terse and edits small. It may have contributed to the successes above.

      I like that skills are easy to try out, too.

    • dimgl 45 minutes ago
      Yep. I have not yet had a single good experience with Sol or the 5.6 models on a variety of harnesses and configurations. It overthinks, overcomplicates and often makes my code into an unmaintainable sludge. It'll usually take 5+ turns of steering to get it in the right direction.
    • OutOfHere 1 hour ago
      It's your responsibility to set an appropriate level of Thinking. For simple tasks, I use the instant model. As an approximation, the choice is proportional to the amount of time I want it spending on the task. Also, you can always ask it to respond succinctly.