lnenad 2 minutes ago

As a small background, I have a local server and I've been trying out different models with different inference engines, quants, configurations etc... I'm also using Opus and Sol at work consistently. I've used AI since the first wave, first as a toy, then as a highly specific tool, last 6+ months as the primary LoC generator.

This is the first time I've felt, and I use the word *felt* since I don't have a suite of benchmarks or any sort of material approach towards comparing models, that Opus has declined in quality compared to before. Primarily I think its powers of deduction and understanding, even on xhigh, have become much worse. Before, being vague and providing a simple prompt would be enough, it could deduce and expand the details it needed, plus ask you clarifying questions, now this is no longer the case. A concrete, personal example, for a personal project, I've asked it to setup ssl over local IP. I didn't go into too much detail in the prompt as there are many approaches it could take and I didn't care too much to choose. It did horrible. The first thing it did was say the best lightweight approach is to add a reverse proxy. I'm like ok, makes sense. Then after asking it to proceed, it went and added a bunch of config to my golang service and didn't even setup a reverse proxy even when it said that is the way to go. It even said it didn't set it up lol. Then after I told it to do so it failed building the config in a way it was asked of it (support LAN IP and tailscale IP). Etc etc...

When Fable came out it was huge, the benchmarks told the story, and the story mostly matched the experience. It felt, again, intentionally saying felt, like it was miles ahead. Now benchmarks say that there are many models that are close, but in actual use it still *feels* much better. I think benchmaxxing is ruining the value of these benchmarks, if they ever had any. The price for Fable is definitely too much for any personal use now that it's no longer included in the subscription, and GLM 5.2, Deepseek Flash and Qwen 3.8 served locally or via cloud provide a lot, requiring a bit more babysitting though. Considering the price of Fable, my 5k USD Epyc server would pay itself off in less than a year if I used Fable or Opus in the same manner so at least for me the decision seems easy. Probably the last month of my Claude subscription.

bentt 32 minutes ago

They've put themselves in a corner. Fable was too good and they gave it away with the $20 plan. It had to be a big step from Opus 4.8 to show progress, and Opus 4.8 is GREAT at coding in many different domains.

But they're getting killed on token cost. They have to get people paying more for tokens. So then they put Fable in the $200 plan and release Opus 5. I'm suspicious of Opus 5. It is mostly worse than 4.8. It _seems_ like they nerfed it to create more distance between it and Fable.

So what have most of us done? Stayed on Opus 4.8. The statistics bear this out. 4.8 still dominates.

Now they're stuck. If they take 4.8 away, everyone will riot. If they make Opus 5.x better than 4.8, they disincentivize everyone from moving to Fable and most importantly, paying more.

Really, all they can do is take the L for now and just let 4.8 be the apex of the $20 pro plan for the foreseeable future while they work like hell to make Fable THAT much better that it earns the $200 to $infinity that they really want everyone to pay.

jeffnash 21 minutes ago

Hasn't it always been the premise that intelligence would get cheaper? To me, on the enterprise side, it seems like firms are finally getting the memo that, whether you are locked into the Ant/OAI ecosystem or not, you don't need the smartest, most expensive model to do every single task. This is a good thing for overall adoption. Whether that trickles down into regular user behavior, especially with subscription pricing, remains to be seen; even though I intellectually know I don't need Sol for a simple refactor, I am sometimes hesitant to downgrade, as it's hard to accept using something positioned, even implicitly, as 'worse'. Remembering that the smaller models tend to be faster is what usually puts me over the edge.

Anthropic in particular is much more compute-constrained than OpenAI and SpaceXAI and has relied on partnerships to provide inference. This reality factors into their pricing and usage limits (they started 'adjusting' the 5-hour limits during peak hours, and it certainly wasn't an upward adjustment). Accordingly, this is presumably what Anthropic wants, given they develop and release the lower-end models, suggest users use them in various nudges within their product, position the bigger/more expensive models as "For the most complex tasks" in their UIs, and so on.

hbarka 13 minutes ago

Anthropic’s issue is churn because of the peak verbosity vomit coming out of Opus 5/Mythos/Fable. What the hell did they train it on. The sane model is still Opus 4.6.

ieie3366 an hour ago

Fable is not a tool for the average user. It’s a professional tool for highly complex work.

I would compare it to a extremely high end $15k PC, or an expensive pro-grade video camera, or a freight train, or a …

I would say at least 95% of the global population will not encounter a situation once in their life where it would be actually useful/warranted.

  • WinstonSmith84 38 minutes ago

    I wish Fable were as good as you make it sound. A plan created by Fable is good, but in my case, it always contain issues caught only when it's reviewed again (whether by itself, Opus, Sol etc.). That's (almost) not different from plans created by Sol, GLM 5.3 etc. The one thing where it's genuinely better is the front-end, but then again it's far from perfect, it just needs less iterations.

  • bentt 31 minutes ago

    Yeah this is a fair point. I only go to it when I have some big architectural problem I want its help in working out. Or a super nasty bug.

  • throwaway63467 35 minutes ago

    It works much better on regular software development e.g. for complex refactoring where cheaper models would produce a lot of garbage results.

  • tcp_handshaker an hour ago

    You have a $3 trillion bubble riding on this not being true.

    • jamiek88 an hour ago

      Right? If we’ve already reached ‘good enough’ then there’s rough waters ahead.

      • Yizahi 5 minutes ago

        I have a sneaking suspicion that someone at Google may be making the same bet, looking at the faster and faster Flash models which provide acceptable results to a lot of people (outside of coding).

      • deadbabe 3 minutes ago

        For Anthropic. But not for the AI industry at large.

        Cheaper, more powerful AI will continue to expand the bubble. Projects will get more ambitious. Everyone will build out their own custom little software. Code diversity expands and requires even more AI.

YuechenLi an hour ago

Fable is just way too expensive and limited compared to GPT 5.6 Sol, and the only task that requires that level of intelligence is frontier scientific research. I use GPT/Codex primarily for coding and usually keep Claude on Sonnet 5 most of the time as I use Claude primarily to debug/brainstorm/make frontends as a supplement to GPT.

  • hellisothers an hour ago

    I thought Sol was on par with Opus, so comparing it to Fable is apples and (very expensive) oranges?

    • YuechenLi 29 minutes ago

      Yeah, it's pretty much apples to oranges, and I don't consider GPT and Claude to be interchangeable at all. From my anecdotal experience, GPTs generally codes more creatively and verbosely but Claudes tend to code more carefully and precisely, so the result is that GPTs generally finds more creative solutions to problems but also writes buggier code, which is why I converged on the setup of GPT/Codex for implementation and Claude for debugging, which feels more like a force multiplier than using each model individually.

    • rybosworld 33 minutes ago

      I've used all three extensively.

      Most of the benchmarks have exceeded their usefulness. Opus 5 beats fable 5 on many of them. Anyone who has used both models will notice immediately that this doesn't translate to the real world. Opus 5 is nothing short of a regression from Opus 4.8. Fable is genuinely a great model so long as you don't trigger a guard rail and it downgrades.

      Sol in my experience isn't significantly different than fable ignoring that Sol burns usage 10x faster but the end result is hard to differentiate.

      GLM 5.3 is a hair behind these two.

      An anecdote but not an original one from the people I talk to.

throwaway63467 32 minutes ago

Yeah I mean if I run out of tokens every couple of hours and have to pause my work or shell out more money I’ll switch to other tools that don’t have this problem. Though they turned this down a bit it seems, I can work with Fable reasonably now and I enjoy it actually. I think they were just testing out how much they can raise the cost without users leaving when having the best model. I guess not much after all!

verdverm an hour ago

I've always wondered why everyone flocks to SV's latest darling company. Have we not learned from our history of glorifying these SV darlings that turn hostile?

  • Eufrat 40 minutes ago

    I think the glib answer is, “Greed blinds all”.

    Most of the people pushing this are just hoping that they can cash out before the hype pops and financial gravity crashes the party. Sam Altman recently claiming that the singularity is here is so stupid on its face he should just be treated as what he is, a huckster.

    None of this stuff ever made any sense on what it was being sold initially. It was always insulting that the media and business leaders tried to argue that the tech could replace entire call centers or vast swaths of entire industries.

    People keep arguing, but it will or it has based on extrapolating certain, reasonable use cases. Klarna has shut up about replacing call centers with bots because Markov chains with memory only can do so much.

  • tyleo an hour ago

    Are people flocking to them? I see people buy the products but if you ask I think they are just about as hated in the big techs.

sajithdilshan an hour ago

Every software engineer in my company uses Claude code heavily. However we’ve never enabled Fable and only use Opus, Sonnet and Haiku.

Nobody has complained and seems like for every use case we have Opus is more than powerful enough, especially with Opus 5

bellowsgulch 37 minutes ago

Reads like: brilliant Carnegie Mellon University computer science grad struggles to find job where he is not replaced by cheap, inferior Indian labor that still gets the job done, even if it takes marginally longer.

No shit we're all paying 清冲 Flash to do the grunt work. Turns out though, paying 清冲 Flash a few more cents does exactly what Ivy Wasp Pro Mythical does. Crazy how that works.

felixgallo an hour ago

[flagged]

  • colingauvin an hour ago

    I am because Claude has figured out I am a biochemist and therefore even asking Fable what the weather is gets me bumped back to Opus, sometimes even Opus 4.8 instead of 5. Kimi? GLM 5.3? DeepSeek? No such problem.

    I can literally open a new chat with just "Hello" and it gets bumped.

    • felixgallo 43 minutes ago

      [flagged]

      • dgellow 41 minutes ago

        Fable guardrails are insanely restrictive, are you denying that’s the case?

        • felixgallo 31 minutes ago

          You’re telling me that a biochemist opening up a chat and just seeing hello causes an immediate restriction? How exactly does the LLM know that the guy is biochemist? Do you guys even know how computers work?

          • kami23 20 minutes ago

            Yes that is exactly what's happening.

            Claude keeps track of memory if you've turned that on so all conversations are somehow tracked over time Claude randomly will mention that I'm a developer while I'm asking unrelated questions and say oh because you're a developer you might like this or because of my background and infrastructure you might find this interesting and I always get creeped out by it.

            They have the concept of incognito chats, but I can see those being worse without context from other conversations happening.

          • colingauvin 20 minutes ago

            It has figured it out through the software that I work on, data that I analyze, questions I ask, etc. I guess I can't prove it's because I'm a biochemist, but that's the only thing that makes any sense. Either way it won't let me use Fable.

          • dgellow 12 minutes ago

            Im starting to doubt you actually used fable

  • cmenge an hour ago

    Just cancelled my Claude subscription and used Ox Alpha Free through OpenCode (so that's one).

    In my view, from the testing I did since Thursday, it's better than Fable. I had just finished a rather large task that Fable completed, including a /review and an /ultrareview.

    Ox Alpha found bugs that Fable and Opus missed, and it continued to build things like a pro.

    It does have issues with availability - but it's on a free promo right now. That also means that I don't know how much it would have cost if I had to pay API prices for it, which might not be cheaper than the subsidized small company / consumer usage, but for large companies paying API prices for either offering, the difference will be considerable.

  • tyleo an hour ago

    [flagged]

    • dgellow an hour ago

      Doesn’t matter if your company pays for Claude, Anthropic and OpenAI valuations and expenditures commitment requires them to win the vast majority of the market to make economical sense. And the Chinese competition makes that very unlikely, to say the least. They likely won’t disappear fully but their “free” lunch as the AI darlings is done, on paper. Will be interesting to see how they adapt