OpenAI vs Anthropic
Video Overview & Insights
Get started with Higgsfield Supercomputer: https://tinyurl.com/4tbd99ut
Recursive self-improvement as the comeback narrative isโฆ optimistic. Even if models get more efficient, you still need compute to serve them at scale. Infrastructure is the real bottleneck.
Join My Newsletter for Regular AI Updates ๐๐ผ
https://forwardfuture.com
Are they really Open Source? I thought only Open Weights.
So source and training data are still secret..
Basically: Freeware- You get the .exe but NOT the sourcecode.
Will Codex reset Microsite: https://www.willcodexquotareset.com/
My Links ๐
When Iโm comparing AI tools, I always ask both models for the same task in a strict format first, then compare the results side by side. For
๐๐ป X: https://x.com/matthewberman
๐๐ป Forward Future X: https://x.com/forwardfuture
Nobody is winning the AI race. Winning would mean the others just stop and leave. That's never gonna happen. And as we have seen suddenly someone else comes around and has the best model for a while.
๐๐ป Instagram: https://www.instagram.com/matthewberman_ai
๐๐ป Discord: https://discord.gg/u7wTTGWhuJ
My problem with ChatGPT is it feels like it's constantly compacting the context - even more than if I were using claude opus and limiting myself to 256k token.
๐๐ป Spotify: https://open.spotify.com/show/6dBxDwxtHl1hpqHhfoXmy8
Media/Sponsorship Inquiries โ
Keep it up buddy ๐๐
https://bit.ly/44TC45V
Links:
Excellent video, Matt!
https://www.youtube.com/watch?v=n1E9IZfvGMA
https://x.com/thsottiaux/status/2075641131002700120
Yep, just waiting for them to pull fable and plan to cancel
https://www.reddit.com/r/ClaudeAI/comments/1u39y1c/subscription_plans_are_massively_subsidized/
https://x.com/thsottiaux/status/2075820987833274448
8:30 you could just learn how to prompt engineer for efficient token usage.. The rest of us dont get YT sponsorship money.
Claude Code has been great for my codebase and I dont have to worry about OpenAI's next lawsuit, like the one Apple just dropped accusing them of stealing trade secrets, affecting my ability to continue development.
You're looking at one snapshot at release where OpenAI is offering a slightly more generous deal instead of tracking across time, implying Anthropic cannot respond to this market pressure from OpenAI. Very bizarre take.
If we were using your logic, we'd be flipping between models on ever new flagship release. That's not how any serious production team operates. What are you even talking about?
https://artificialanalysis.ai/
https://x.com/claudeai/status/2076351399999557669
Anthropic is too expensive. I suspect it was a compromise to keep the chinese from mining it for vast amounts of data.
https://x.com/thsottiaux/status/2076365965915467978
https://x.com/trq212/status/2072814903170408784
One time, I got down to 80s. I had to take a screenshot so I could explain to friends on Claude.
https://x.com/sama/status/2076780425280954658
Great video. Thanksโค
More User Perspectives
I forget to mention this but why do you think that they squeeze the most out of the GPT 5 training run when GPT5.5 was literally a brand new pre-trained and they put it out there literally just to find the cracks and optimization needed for 5.6 so 5.6 is just literally the first major iteration to a brand new pre-trained Spud 5.5 kickstart, maybe they should have called it GPT 6 but I think that's why they waited because they wanted to optimize.
But I have heard they are doing another pre train for the GPT6 separate as well that part I cannot confirm.
I feel when they returned fable 5 it was significantly dumbed down... and burns credits like nothing... and yelds mediocre results... Just tried two prompts and got 20% of my credits... then I switched back to opus 4.8.... when my subscription expires on 25th... I switch to codex...
@QuantzAiGenuine question for everyone who switched: did you move your whole workflow over, or just the parts where cost matters more than reliability? I keep both open and route by task depending on the job. Still curious if anyone has actually found one that does everything well without needing to babysit it.
@magneticfoundersits for me sade you can talk with app in Polish like openai and create pictures. Right now people which I know mowving back to Openai because this simple thinks. People who code still have problem decide which team to chose
@jabcaqthings change fast see you when you change your mind
@SamTest-u5sTUNNEL VISION: Relying on a single AI model to design and write software code. (by KaiDev, me) THE METHOD- The Implementer: The primary AI that writes the core code.
The Auditor: A competing secondary AI model that stress-tests the code for flaws. Loop Sharpening: The process where the Auditor critiques the Implementer's logic, Auditor mandatory "must-haves" (security) and "nice-to-haves" (optimization). Be amazed at the gaps! Back to implementer: Tell implementer to sharpen it, Hint: Use Claude as your Auditor not implementer, make a bridge, automated process, ๐คซ (Keep this automation loop between us, do not be shocked at the results!)
And ChatGPT has an excellent dictation + listen mode and a poorly performing but best (among competitors) Live / Voice mode. Besides I want their leaders all in jail, except for Dario , a nice person.
@lacis9546GPT is great. I never get throttled. Use it every day, all day. Claude? I will do one task, start on another, then boom. My 5 hour limit is hit within 30 minutes.
@JohnDeNomChatGPT is the MySpace of AI ๐
@semperayeHere's the thing
@IN2FPAAll this BS until Anthropic pulls Opus 5 lol
@kiteemm9165You act like Altman didn't do any fear mongering.
@longtimelurker1A2B3CI wanted to hate open AI for a long time, because Sam. Idiot. But from a product perspective, Open AI has honestly killed it. So much nicer to use. I found the Claude Mac app so slow and clunky. The remote works so much better from my phone with Codex. And, at least in my last use, you don't have projects in claude code... which is mental. Sol 5.6 is amazing.
@scottaltham6189"go out and just build"? Build what? 'Go out and enjoy nature/life' would be my choice.
@alensiljakOnce Anthropic started locking in Claude code I was done with them. Switched to codex and never looked back.
@dutchdelite77If fable gets 5 times more efficient, anthropic will just make 5 times more per token.
@dutchdelite77Kimi 3 shot across the bow with all of this RSI stuff
@WILLinHDY'all quickly forgot all the sh1t that OpenAI did? Even the name itself is hilarious, there's nothing open in that company. Shady sh1t just smiled once to you and you already fall in love
@angryktulhuIt's interesting how we're reaching a point where the best model isn't automatically the easiest one to use. Reliability and generous limits matter a lot when you're using these tools every day.
@ABTalksOnAIIt is really depressing to hear that the evil Anthropic will win.
@anonofish576I was completely unable to use Fable due to their guardrails. On top of that Fable was an asshole towards me. I have none of those problems with GPT 5.6.
@anonofish576Now that Kimi k3 has released with same score as openai's sol and much cheaper, I feel that Antropic made the right move
@PushkarkvMatt on Anthropic:
"We loves it!"
"We HATES it"
What will uappen if its not needed to poke the bear anymore
@Joseph-e3l9vLet the models do the talking. MB's trash talking of Anthropic wont change anything. Nor will his back slapping of OpenAI. If you code for a living then you just pay for the best and get on with it. Right now - that's fable. And yeah they have been resetting once a week, sometimes twice since 5.6 sol came out. I dont think they can afford to cut fable from subs now.
@johnh6959for recursive learning to be effective, you need tons of data points. OpenAi is gaining way more datapoints, while investing into SOTA hardware, and TAKING market share all simultaneously. Anthropic is toast man. It doesn't matter if they have the smartest model anymore, the game changed, its economics people care about.
@FiveTwoSolutionsBut Open AI do treat you with contempt. What are you even talkng about?
They are not transparent at all. That's BS.
You compare Codex usage metering to any other serrvice in the world.
If any other service was so completely opaque about how much you actually pay and how much use you actually get for your money -- you would kick them in the teeth and laugh at them.
It's FKN nonense.
Discord link is expired
@AksharaTGdo a video about kimi k3
@gabrieltelles-e6sAnthropic banned my subscription. No warnings, no explanation. As it happens, I'm better off with Codex anyway.
@kiaranrKimi K3! Now Anthropic and OpenAI are in trouble! Gist: Ranks No.2 depending on the benchmark, fable/sol level. Open weights to be released! Didn't Musk "predict" that open source models would have fable level knowledge early 2027? Well, he's off. Not the first time that happened, lol.
@markg5891"Heres the thing" = AI wrote it, Matt. You helped us be smarter than this. Dont make us cringe, unless its on purpose.
@empathyworks8890Your ads are worse then Google
@PrepSchoolGangsteron a more personal note of reference on the whole "recursive self improvement, i would rather wager that openai likely have that efficiency being built, as that is essentially how we got to the 3 model thing of sol, terra and luna.
fable has been able to move from 5x opus in mythos preview to 2x in mythos/fable 5 so i wouldnt say its fully inefficient. but its still pretty expensive.
for all intents and purposes if i were to wager who reaches first in terms of recursive improvement i would say the difference between open ai and anthropic wouldn't widen per se but remain close.
its worth noting that anthropic only have gotten to building the best models by ignoring the efficiencies all together (haiku is a dead corpse, sonnet was mismanaged into being essentially a weaker but similarly priced compared to opus, the improvements in opus with 4.7 and 4.8 were at best minimal multimodality and honesty improvements.) while openai are tackling improvements on things that do feel like upgrades and hence closing the developing gap and for the first time the argument has been sol and fable (with all the pesky downgrade to opus stuff) are equivalent in a sense and pricing is more favourable with sol (also not to mention the cerebras partnership to make sol output even more soon).
...I think it will be fine... ๐
@StevanSrdicit's just so obvious that you're on OpenAI's teamโฆ
You backhand Anthropicโฆ You undercut Anthropicโฆ Your compliments always have a but.
But there is no need
@TimDavies1955Donโt want anthropoid
@TimDavies1955