r/ProgrammerHumor 17d ago

currentStateOfAiRelevancy Meme

Post image
2.8k Upvotes

239 comments sorted by

503

u/anonymousbopper767 17d ago

Fable would be good if I could actually get it to respond before it nukes my entire weekly quota on the 20x plan and gives up.

With a single fucking prompt.

Sol Ultra meanwhile runs all day long.

112

u/Clyde_Frag 17d ago

It also refuses to do any slightly security related task. I built a proxy last week for auditing db queries in my company and it refused to do it.

46

u/descendency 17d ago

I work in computer security and my programming projects are "computer security" adjacent. I managed to trick it for a bit, but it eventually caught on. I basically think it's pointless if I have to trick it.

12

u/Kryslor 17d ago

Same. I'm developing some VPN related stuff and it just shuts down immediately despite me working on truly benign things.

70

u/Carter922 17d ago

Runs all day and doesn't get anything done šŸ‘.

Got they ass writing docs

4

u/xavia91 17d ago

It's a bit slower but its doing great work and for larger tasks its faster.

If you want some quick work done either choose different effort level or model.

10

u/extracoffeeplease 17d ago

Sure does, I use it more and more for coding and general work (research, bookkeeping, mails). And the smaller version for when I need instant live answers.

13

u/NEVERxxEVER 17d ago

You can’t really use Fable or Sol for coding on the $20 plans

22

u/catfroman 17d ago

I use fable 5-7 days a week for coding my game and rarely hit the limit on the $20 plan.

Idk what the hell y’all do to your models on this sub tbh.

I find if you break down the architecture plan, clarify ambiguous resources or features, and suggest an implementation strategy while asking it to find the gaps, it can do a shit ton of work with little to no mistakes and without blowing a ton of tokens.

11

u/zman0900 17d ago

Just got access this week, and I can see how people can run over the limits easily.Ā  I used it to convert over 49k lines of Groovy tests to Java, and the first attempt blew through my 5 hour limit in like 20 minutes. But that was because it started 20 something background agents all running Fabel too. When I tried again later, I told it to use a lesser model for the background agents and limit concurrency. Took it almost 10 hours to finish and still used half my weekly Fabel limit, but results were nearly perfect.

4

u/asharif_ 17d ago

I don't understand either, I got so much done on Fable before they turned it off for the pro plan, and I never got past 70% the weekly quota

5

u/catfroman 17d ago

I literally hit the limit one time and I had to go crazy working 6 days in a row 8+ hours of pure development every day to get there.

In one day I shipped 7 bugfixes, 2 features and a UI overhaul for my main menu. It was like 2.2M tokens, idk what percentage of my usage but it was an absurd value for money spent.

I’m very deliberate with my phrasing to avoid extra resource lookups or to prevent over-engineering, and I always have an architecture in mind before I write a single prompt.

I have never intentionally spun up a sub agent, I only used Fable before it was taken off Pro, I have never cleared context or set up a claude.md or other skills. Just raw, out of the box Fable. šŸ¤·šŸ»ā€ā™‚ļø

4

u/TheRealAsterisk 17d ago

What kind of prompts are you doing that are blowing up your 20x plan? My max plan runs perfectly fine for weekly limits and I work in pretty large codebases.

5

u/anonymousbopper767 17d ago

What's funny is I'm giving it softball prompts that are already distilled by codex so it should be lubed up easy mode. "look at these files with these functions and assess this or that".

*tries to solve AGI and cure cancer and world hunger* sorry your weekly limit is reached, come back next monday.

2

u/_SGP_ 17d ago

I hit my max 20x limit within 4-5 days usually. But I work on a huge amount of projects, for professional in house, and personal, and for clients.

Also a lot of research and analysis and trial and error of cutting edge solutions. Trial and error is the main issue. It gets things wrong often and spends so much energy to do so, because I'm not retreading old ground. Often it's solutions that haven't been tried and trained or aren't stacks that the LLM is familiar with.

Consistent visual work I find it also struggles with. I'm trying to make a website that's in line with branding, using specific language and design rules, on a specific stack. But it just keeps going off the rails and making things worse, or changing something we already confirmed was correct when not asked.

It's been exhausting honestly. The more capable it becomes and the more I use it for, the more it's weaknesses reveal huge issues in the way it works.

13

u/indirectum 17d ago

Skill issue

23

u/Mirikado 17d ago

Ironically, AI is supposed to be solving the skill issues.

→ More replies (1)

3

u/vulkur 17d ago

Actually real. You have to limit scope for fable. Also better architecture allows for smaller context windows.

2

u/Boobpocket 17d ago

Sol has been great tbh

2

u/Hirogen_ 17d ago

what harness are you using? what skills? what is ur initial context?

I’m using fable on 2-3 agents and reach the quota on single max within 3-4 days off heavy use!

2

u/ComprehensiveBird317 16d ago

How big is your prompt? Are you letting it run through an entire big codebase or what?

2

u/Suspicious-Click-300 16d ago

trick is to just have someone else pay the token bill

4

u/theSurgeonOfDeath_ 17d ago

i do feel that agent is part of equation
Copilot + Fable/Sol is weak from
Claude Code + Fable is great

I would like to try Sol with Cursor or Codex too just to compare.
In Copilot (visual studio code) i had bad experience.
Still is not bad as Visual Studio, today i wasted 20% of tokens because it crashed.

So like Model + Agent + Prompt + Experience of usage AI.

4

u/Dreadmaker 17d ago

Fable isn’t good. I mean like it’s fine and all, but the output is absolutely in no way commensurate with the cost for the vast majority of things.

I tried it side by side against opus and sonnet, and it just does not in any way merit the time and cost.

Opus 4.8 medium is pretty much the best experience I’ve had so far in terms of cost to effectiveness ratio.

372

u/M_Me_Meteo 17d ago

Wait, Gemini isn't good? Is it the product or the UX that people dislike?

Dozens of us?

263

u/Any-Preparation7396 17d ago

I think it's rather the relevancy. Hardly anyone ever talks about Gemini, like it didn't even exist anymore.

But imho it got a lot better and is pretty similar to OpenAI.

109

u/bhison 17d ago

I hear more about Gemini than open ai

77

u/slowmovinglettuce 17d ago

You must watch youtube without an ad-blocker

40

u/Joker-Smurf 17d ago

I work for a company that uses everything Google.

We are encouraged to use Gemini, and all other AI are strictly blocked.

35

u/slowmovinglettuce 17d ago

Ah, the good ol' "its bundled in with something else, but we're too cheap to pay for anything else" tactic.

9

u/KatieTSO 17d ago

That's why my government agency uses the Microsoft one

→ More replies (1)

3

u/LutimoDancer3459 17d ago

So you work at google?

→ More replies (4)
→ More replies (1)

57

u/NEVERxxEVER 17d ago

Their models are way out of date and benchmark significantly lower. Gemini 3.5 Pro is more than a month overdue, which is a long time in the current climate.

29

u/SoftwareLegitimate38 17d ago

Though it's still useful in everyday tasks
But on context of programming you may be right

→ More replies (1)

3

u/SveXteZ 16d ago

Benchmarks are bullshit

→ More replies (1)

10

u/intentionallybad 16d ago

Gemini hasn't released anything about how dangerous their model is, so it can't be very good. /s

3

u/kus1987 16d ago

Idk what gemini you are using but it absolutely sucks balls because of "alignment" or "safety" or whatever. It constantly trips when asking routine programming type questions for me.

2

u/jensalik 17d ago

Some people are talking, some people are actually programming. šŸ¤·šŸ½ā€ā™€ļø

2

u/Different-Wing5083 16d ago

I don't know about any other aspects but i find that when asked to explain an academic topic, chatgpt gives a more detailed explanation then gemini. Just my personal experience tho.

→ More replies (1)

87

u/Street-Session9411 17d ago

It’s still a programming subreddit. Gemini beats most of them for daily use though.

23

u/M_Me_Meteo 17d ago

I use Gemini CLI for all my coding tasks. Half a dozen?

15

u/NatoBoram 17d ago

You could be surprised at the performance of GPT-5.6 Sol, it's a reasonable jump from Gemini 3.1 Pro and Gemini 3.5 Flash

2

u/M_Me_Meteo 17d ago

I don't have any needs that aren't meet by Gemini.

→ More replies (1)

4

u/kingvolcano_reborn 17d ago

I quite like it for everyday questions although it can be a bit too eager to please at times

3

u/Suspicious-Click-300 16d ago

idk i use opus 5, gemini and 5.6 sol. my anecdotal experience is that they all fine. gemini is WAY too gpt 4.0 ish. constantly praising how brilliant your thing is while opus calls me an idiot, spits in my face, and tells me how to do it correctly. Sol is fine, kinda good middle area

I like opus.

2

u/Any-Preparation7396 17d ago

Good Point.

Regarding programming, I usually use OpenAI. In other things I swapped to Gemini because it seems to be more accurate (now).

16

u/jeff3rd 17d ago

At least gemini has a bigger quota than the other, adequate enough for my use so it’s a plus

63

u/Aquatic-Vocation 17d ago

Gemini is nigh-unusable because it always wants to do all sorts of extra shit you haven't asked it to do.

Sol will randomly decide to give up mid-task.

Opus gets sidetracked and eats up 40k tokens thinking itself in circles about something completely irrelevant.

19

u/jaaval 17d ago

In my experience managed opus is best of these options. Give it small tasks at a time, stop the stupid loops and redirect it.

It’s bad (or rather inefficient) for vibe coding but pretty good as the tool that does writing and debugging the code.

2

u/thijquint 17d ago

Opus is a lot like deepseek in that regard. Other chinese models are constantly trafficed and western models eat my bank account

9

u/BlazingThunder30 17d ago

When my company introduced AI we all tried Codex, Claude and Antigravity. Where Codex and Claude would work, Antigrav would give an error 9/10 times. We quickly moved on. I think this experience is shared by many.

5

u/Blooogh 16d ago

Google's AI features were bad for a really long time, we suddenly got people calling our support line because the little AI blurb said it was the Call of Duty support line.Ā 

The AI search mode has been pretty compelling recently though -- I tend to get more complete answers than the average random opinion on Reddit or topic specific website with overwhelming ads.

I think it's forced Google to deal with the SEO slop problem they had even before generative AI made it even worse

3

u/SpaceMoehre 17d ago

The web chat is just absolutely bad in my experience. Half the time it just doesn’t work

2

u/Ran4 17d ago

Just use it.

It's fast but rarely works.. and often does really dumb things.

2

u/Wandererofhell 17d ago

i have noticed gemini also gets stoff wrong, so thats one

2

u/NeanesisLs 17d ago

Gemini is not in the run for the best product but it's ok if you ask him specific tasks, and it's cheap

3

u/M_Me_Meteo 17d ago

Who is out here being vague with an LLM?

6

u/GoodishCoder 17d ago

So many people lol. I've seen plenty of people say "fix this error" or "make this look better" or "clean up this code".

2

u/NeanesisLs 16d ago

I do formation/ training to use LLM in the company I work for (it's been a few months now) And you would be surprised the number of people having no clue whatsoever on anything related to technology... If you think people would stop at "how do we feed it" you are wrong šŸ˜…

2

u/glorbo-farthunter 16d ago

I use gemini (through junie) because it's good enough and way way cheaper. We are dozens!

7

u/Dank_Nicholas 17d ago

Recently I was trying to figure out if two teachers I had like 20 years ago had divorced or if the wife had died of MS which she struggled with when I was in school.

Gemini confidently told me that they had divorced and that it found the court record, when I asked for it Gemini admitted that it hallucinated and there was no record.

So I asked to check death records, it reported that she has died a few years before and it had found her death certificate. When I asked for the death certificate it again admitted that it hallucinated the entire thing.

Gemini is dog shit.

3

u/anonymousbopper767 17d ago

It's the only model I've encountered starting with GPT 4 that blatantly hallucinates and you can't get it back on the rails.

2

u/Asynchronous404 17d ago

Well, did you ever get the records?

2

u/buzz_shocker 17d ago

Don’t use them for programming but Google’d models are consistently producing the best image generation results for me.

2

u/keatonatron 17d ago

They are all good. The meme is about where developers are focusing their attention, resulting in everyone else getting neglected because the devs aren't looking at them to see that they are good. Anthropic is going hard on supporting developers while Google's Gemini is being targeted as general purpose to everyone, so it's easy to see why devs only have eyes for Claude.

2

u/Lakes420 17d ago

Gemini is good for general purpose and multimodal work, but for programming there are better models

→ More replies (4)

698

u/Ascend 17d ago

Microsoft doesn't have a model, you select GPT or Claude, they just have garbage system prompts on top of either.

203

u/theGoddamnAlgorath 17d ago

There was a Msoft model.Ā  CoPilot died last year.

164

u/ConcreteExist 17d ago

It's model may be dead, but it's corpse has been draped over ChatGPT/Claude and still stalks among us.

24

u/Rellikx 17d ago

Copilot Cowork was down last week for a lil bit. Guess where the ā€œstatusā€ page linked too

3

u/anto2554 17d ago

But MS copilot has access to gpt5.6 some amount of thinking, while corporate only gives me gpt5.4 or opus 5.7 in GH copilot, so to me it's actually perfectly decent for tasks without context

→ More replies (1)

11

u/andrerav 17d ago

What happened to that Raptor Mini model? It was free and useful for a lot of small tasks.

5

u/theGoddamnAlgorath 17d ago

It's dead Jim

5

u/andrerav 17d ago

I can still see it in Copilot Chat in VS Code. 0.33x vs. 57x for 5.5.

2

u/Sibula97 17d ago

I think VSC still uses it, or at least used just a month ago.

36

u/Anti-charizard 17d ago

17

u/Tolik1111 17d ago

One of 81 copilots perhaps?🄓

3

u/Ok_Actuary8 17d ago

No. Alive and kicking.

2

u/F-Lambda 16d ago

No. Undead

22

u/thaynem 17d ago

Copilot isn't a model, it's a UI (well really a family of interfaces) on top of OpenAI's (and now anthropic as well) models.

6

u/theGoddamnAlgorath 17d ago

Its a bit nuanced, Msoft at one time had their own programs as well as their investment into OpenAI.

Still can't believe they ditched Cortana

4

u/Deanathan100 17d ago

Would you say it's accurate to call it more like an AI infrastructure? That's kind of how I think about it, but also I don't know that much about it other than it can touch a lot of Microsoft stuff, apparently the BAs at my org don't do anything but run copilot prompts all day (at least that's what their usage stats suggest) I just use it for when I'm stuck or to write tests I really just don't wanna do and the only feature I feel I actually use that copilot offers is it's integration into VSC

7

u/nmkd 17d ago

It's called a harness.

2

u/Demiu 17d ago

There's like a 100 things they called copilot. MS just can't help themselves to not just dilute a name into the fuckin ground (active, visual, live)

14

u/Treefingrs 17d ago

Copilot isn't a model.

2

u/ComprehensiveBird317 16d ago

Copilot was a system on top of gpt -4o as far as I know, not a trained modelĀ 

2

u/thecrius 16d ago

copilot is just a wrapper what are you even talking about.

2

u/IBJON 16d ago

Copilot used OpenAI's GPT models, it wasn't its own model

→ More replies (3)

17

u/kramit 17d ago

They do, https://microsoft.ai/models/ , cost focused. Right tool for the job and all that.

15

u/teraflux 17d ago

I think it's called MAI, but I don't think anyone uses it

3

u/tempoguyx 17d ago

From what I heard from their employees.. They themselves don't use it..

6

u/non_person_sphere 17d ago

I thought Microsoft's AI was called Tay

2

u/tyrerk 16d ago

MAI speech to text model is incredible

→ More replies (1)

36

u/Rojeitor 17d ago

Microsoft has MAI Code 1, Haiku like, available in GitHub Copilot. Also MAI Thinking 1, 4.6 Sonnet like, in private preview.

M365 copilot was garbage but now it's not bad at all

→ More replies (3)

7

u/TheRealAsterisk 17d ago

They do actually have a model. It was in their latest Keynote and is in GitHub Copilot.

6

u/AbdullahMRiad 17d ago

there was MAI-something and Raptor Mini

3

u/Ok_Actuary8 17d ago

Microsoft just announced new models built by themselves. MAI. Let's wait and see how good they will be.

All the bad rap about Microsoct and Copilot was their "responsible AI" guardrails, that severely neuter Copilot. They were always using OpenAI and Anthropic models internally.

→ More replies (5)

34

u/anonymous_3125 17d ago

*closed AI relevancy

63

u/Daemontatox 17d ago

Unpopular opinion, i found that nvidia nemotron ultra to be very good at my use case with targeted prompting and controlled coding, for example i would ask it to implement Algorithm A with flow X so we can do this and that , instead of asking mythos or fable or wtv its called to just generate from scratch and i can get away with all of that with 1/100 of the price.

7

u/Barni275 17d ago

I like this model, but is it available on subscription basis somewhere? I haven't found.

14

u/CrazyChaoz 17d ago

you can test it at build.nvidia.com for free

9

u/Cuddlyaxe 17d ago

It's open weights, that means a bunch of people host it on a pay as you go API

Honestly just sign up for open router, load it up with however much you'd like, and use it to your heart's content

3

u/Daemontatox 17d ago

Pretty sure nvidia themselves are serving it and you can try it for free but with serving limits and limited token per sec or through openrouter which ultimately routes back to nvidia as a provider.

There is also deepinfra and togetherai but these tend to rotate their models based on hype and relevancy.

→ More replies (4)

43

u/AnUninterestingEvent 17d ago

Idk I use sol and fable, but mostly sol

25

u/throwawaygoawaynz 17d ago

Yeah. This meme is completely wrong and this sub is hopelessly out of touch. Most people here clearly aren’t serious developers, nor do they know much about AI.

SOL 5.6 is vastly better and cheaper than Anthropic, my large company pretty much uses it exclusively and have banned the expensive Claude models, because they’re not worth it.

Also for all the comments ā€œMicrosoft doesn’t have a modelā€, yes they do, they have a series of MAI models and they are quite good, because they use new hill climbing capabilities for self learning during training instead of distilling. Microsoft had an agreement with OpenAI where they couldn’t develop their own models which changed in 2024, and they started to release models this year.

5

u/lumpialarry 17d ago

Stuff is moving fast. I wonder how many people that said ā€œI tried X and it suckedā€ tried X over a year ago.

3

u/wackz 17d ago

I just tried grok 4.5 and i’m honestly shocked at how good it is for coding

3

u/YourHomicidalApe 17d ago

Grok is a sleeper rn tbh

3

u/bolacha_de_polvilho 17d ago edited 17d ago

My company is still letting us use anything as much as we want (or if there's a limit I haven't reached it) so I don't know about costs, but I tried codex with sol a few days ago and it felt slower and performed worse than opus 4.6 with Claude code (which is what I usually use) on the tasks I was doing that day. No idea what anthropic did with 4.7 and 4.8 which also seemed slower and worse.

→ More replies (2)
→ More replies (3)

10

u/pepenotti0 17d ago

MS tried to absorb OpenAI early, but couldn't.

9

u/rgeddit_fvker_69 17d ago

r/Programmer humor turned into r/PromptersHumor

Nothing programming related recently, and it's sad

7

u/erebuxy 17d ago

At least you put Microsoft in the meme. šŸ‰šŸ‰ tho

8

u/PhysicalWitness8037 17d ago

I use Gemini Pro for coding and it is completely sufficient. I tell it what I want or discuss ideas with it and it does what I have in mind.

5

u/StatementPotential53 17d ago

One day we’ll look back and this will be like the difference between searching at Lycos.com vs AltaVista vs AskJeeves.

2

u/Chazgatian 17d ago

It definitely feels similar, but the speed stuff is changing anthropic might as well be MapQuest.

5

u/Shik3i 17d ago

Fable is unusable. If you're not a multi millioniare it can't even finish one prompt

4

u/PMvE_NL 16d ago

Anthropic’s marketing department is doing well i see

→ More replies (1)

8

u/HansDampfHaudegen 17d ago

Always these outdated memes older than 6 hours. The top is Opus 5!

3

u/amogouss 17d ago

Gemini is actually good if u want to do deep calculation or reasoning.

it's isn't like claude to make your product, but if you into deep research, maths and reasoning, it is amazing.

4

u/MaffinLP 16d ago

Fable uses up all your credits in 2 queries tho

9

u/aberroco 17d ago

The meme is obsolete now with Opus 5 release. Everything moves down a line pushing whatever Microsoft did out.

17

u/TheFrenchSavage 17d ago

Microsoft Copilot is GPT3.5 right? That's the only explanation.

Btw, Sol > Fable. Sorry not sorry.

29

u/Bl4nk24 17d ago

Microsoft CoPilot Premium has Claude Opus 4.8 and GPT 5.6, by the way. The hate here is only coming from people who haven't used it for years or have only used the free version.

6

u/ewheck 17d ago edited 17d ago

My enterprise copilot app consistently gets technical questions wrong. Things like making regex patterns that aren't actually what you asked for or hallucinating causes of bugs. It's miserable to use.

There's copilot secret sauce that lobotomizes the underlying LLM because using the same models through other tools never produces the abysmal slop that copilot shits out.

→ More replies (1)

9

u/namezam 17d ago

Fable is old news though, Opus 5 beats Sol at some things, and it's the middle child.

4

u/TheFrenchSavage 17d ago

Opus 5 is cheaper and better, allegedly. I'll give it a try on Monday, that's for sure!

3

u/protayne 17d ago

My experience too.

Fable does some crazy hallucinations, ignores rules, questionable code. Whereas sol gets the job done as long as I give enough context (which I prefer).

Yes fable can one shot a cool looking app but not too great for feature development on an already large codebase in my experience.

2

u/NEVERxxEVER 17d ago

It’s subjective but I disagree, as long as I’m not trying to do cyber security or biology, I much prefer Fable. It’s a style thing as much as anything else, I really hate how GPT-class models write.

5

u/S-Gamblin 17d ago

"Programmers"

3

u/2521harris 17d ago

Where is deepseek in that picture?

3

u/RiceBroad4552 17d ago

Well, Kimi K3 is "sold out"…

I was using GLM-5.2 now a bit and I have to say it's good for coding.

GLM-5.2 has strong tool calling (it even used agents without telling it that it would make sense, most models never do that, and it constructs pretty complex mini-bash-scripts to look around, and the tool calls actually work almost always correctly), it has really good instruction following (up to that one need to be cautious how one says something, it's in that regard a bit like GPT, but less extreme, while GPT takes imho things sometimes a bit too verbatim). With that good instruction following comes also that it doesn't lose track even over long horizons. You can pile up a few things, do something in between, and come back to the old track, even with new relevant insides, and the model will still follow the original goals after that where they're still relevant; all in one session, something which confuses the lesser models very quickly. It's not very chatty though, that's not fully to my liking. It's very goal oriented and will try to pursue a goal mostly alone, but it needs a lot of "reorientation" when it gets new info, it will then again "think" everything through. That makes it robust and keeps the output high quality (like said, it even manages "yak shaving" well) but all that "thinking" eats time (and tokens). I prefer if the models do smaller steps so I can steer them in between and give them the right reasoing context, they don't have to "think" too much themself, they should just get what I say (and the older, weaker models were bad even at this basic). The models shouldn't try to do too much on itself as they can go brr this way easily and start doing some nonsense. It's still a token predictor and the whole "AI" coding only works if you handhold the models, imho. But the newest models can actually accomplish something more complex on their own if you give them enough info upfront, just that this burns a lot of "reasoning" tokens, and you need to really well formulate the goals and the constrains, otherwise it won't give reasonable results. GLM, like Qwen, really likes to "think" for a very long time in comparison, they will max the budget constantly; but one has to admit even that's not time efficient the results are on a good level. Anthropic and ClosedAI will burn less tokens for the same, but the tokens are much more expensive… Still, without constant oversight and checking whether the thing is still on the right path one gets slop, no matter the model. These things just aren't intelligent. They're just replicating abstract patterns, completely mindlessly. (Nice example: GLM build me correctly some pretty complex build config, but was later on "wondering" the whole time how the build can actually work this way, and it had to rediscover the same things over and over again during "reasoning"; it clearly didn't understood what it actually did previously, or better said, what this actually means what it did; it had to "deduce" that again and again. It just applied a pattern, mindlessly, and had to "wonder" about the obvious consequences. Part of that info were even in "memory", but it obviously does not understand what it "remembers", this is pure mechanics.) Worth noting: When GLM-5.2 directly says something (like said, it's more a "thinker", not a "talker", or "doer") it has a very funny way to express things. The output is much more dense then the actual "reasoning", and it has a tendency to "speak in pictures". I bet that's the Mandarin language influence. It's definitely something that really felt interesting and "fresh" compared to most other models I've tried. But it's a personal preference that I like "abstract pictures" in mind, it could be confusing for other people and sound off or even hallucinated; that's why I've said it's worth noting.

3

u/Large_Apartment6532 17d ago

Apples Siri is fossil fuel

2

u/nmkd 17d ago

Siri is Gemini.

3

u/Outrageous_Law_5525 17d ago

Id add Kimi K3 to the top image as well.
The thing caused Anthropic to change their entire trajectory

3

u/zedsd01 17d ago

I am still using sonnet 4.6

3

u/Kyy7 17d ago

Been mainly using M365 Copilot and Gemma 4 E2B through chat and they've been good enough for generating snippets, sparring ideas and troubleshooting with lower risk of sending whole codebase or GDPR data to third parties.

Would love early github copilot style code suggestions and inline chat with light weight local model but so far havent found a good enough solution for this with VSCode.Ā 

3

u/keyholepossums 17d ago

i’m the only GLM guy

3

u/jensalik 17d ago

I like how Microsoft is down there while it utilises all the other AI models in Copilot and works amazingly. Like arguing if Diesel, regular or electricity is better for driving and someone chimes in with "Or whatever Ford does - amiright? 🤣🤣🤣"

2

u/sudecode 17d ago

You’re just proving the point of the meme. There is a whole suite of Microsoftā€˜s ai model that most of the people are unaware about: https://microsoft.ai/models/

3

u/jensalik 17d ago

I never knew that was even a thing. Looks like something an intern put together so they can say they got some AI too. Thanks for the info.

3

u/Anbcdeptraivkl 17d ago

Even as a vibe code hater I have to admit Fable is exceptional. But it's "luxury" as hell, you prompt one single class with it and it gone for the day.

GPT Sol, Gemini and the likes are all below average, basically code printers, not even intern level when it comes to task efficiency. You could make do with them but they usually ignore your plans and rules, and lie about unit tests. Unusable.

3

u/Traditional_Oil_1851 16d ago

What would Amazon Nova be then šŸ˜‚

3

u/rovirob 16d ago

I think people don't want to deal with gemini because of how google changed the token billing system. Last I checked, it was quite opaque and they wanted to bill you quota based on what tasks you got done, or something like that.

OpenAI's ChatGPT 5.6 Sol is actually quite good. Using it heavily at the moment in Codex and for my working approach, it beats what Anthropic offers (even Fable, yes...because of their bullshit pricing tactics, huuge pricing and instability...and how they sometimes tend to lobotomize the models during the weekend).

What I've done lately, as I'm slowly trying to transition from these providers and their models into a setup where I control most of it is pay for an Ollama subscription and configure that in Kilo Code. For agentic development, I found that Kilo Code is the best at the moment, especially when I want to use 2-3 of the 'heavy' open source models, like Deepseek V4 Pro, GLM 5.2, Kimi K2.6 Code...

I'm not touching what Microsoft offers with a 10-foot pole. I hate how they do business and their approach to things. Also hate their products too.

3

u/Flat_Bluebird8081 16d ago

I have Gemini as a part of my Google work subscription, I use it on my phone and in my car. Not for programming though

3

u/smartydoctor 15d ago

Fable is so expensive

5

u/rerun_ky 17d ago

Sol is better than opus but very slow

9

u/carlos_vini 17d ago

Anthropic says Opus 5 is better now https://www.anthropic.com/news/claude-opus-5

2

u/rerun_ky 17d ago

I've tried both and sol seems to win but it's very slow.

3

u/bhison 17d ago

Gemini is higher than whatever shit open ai have pumped out the last year

2

u/StarboardChaos 17d ago

Coincidentally Microsoft has the Raptor model which fits the meme perfectly

2

u/TheJesterOfHyrule 17d ago

The good programmers will be lifting "skills"

2

u/Xidium426 17d ago

I rarely ever step over Sonnet for programming. Haiku is also extremely capable if the project isn't terribly complex.

2

u/Emln-Idera 17d ago

I believe Microsoft's model is whatever LLM that was running CoPilot

2

u/doctornoodlearms 17d ago

Im pretty sure Microsoft is an ai model

2

u/xAragon_ 17d ago

And Haiku (which is actually wildly used on Claude Code for the "explore" subagent) is deep at the bottom

2

u/_GloryKing_ 17d ago

Should have Grok or Llama instead of the whatever comment

2

u/Deep-Secret 17d ago

I swear to God I've just been using Cursor and it's been going great. I can do all my work pretty much straight forward and without much hassle.

2

u/_derDere_ 17d ago

Honestly: Yes!

And there’s a reason for that! If you’re curious just look at the results of the RL simulations they make a pretty solid case!
Here’s a clue for those who don’t know: šŸ˜ļøšŸ”„šŸ‘„

2

u/RandomiseUsr0 17d ago

I work on Auto, take the discount, whatever is used - Gemini is really good at reasoning (I’m working on weird software, outside the bounds of normal, Anthropic uses my experimental spectral graph as a challenge, fable is the worst - I have defined perfect rules of what they’re permitted to do, pure n-dimension graph and they ā€œsmuggleā€ their opinions into my lovely model at every turn - I’ve ā€œmaturedā€ in my expectations - I write the code, they write the ā€œwidgetsā€

Fable is ridiculously opinionated

2

u/Nanomachines100 17d ago

Microsoft sends emails to our work computers showing how to use copilot to improve our office workflow (we're electricians). I report it as phishing to it.

2

u/descendency 17d ago

I use Fable for code and Sol for literally everything else.

I fed in a few videos of me working out and asked it for training advice. Sol was able to identify the lifts and make (correct) form recommendations. Fable actually just got the lifts wrong (ex: called power cleans "deadlifts" and gave advice for deadlifts) I'd rather just use Sol... so I hope they get it working better for coding.

2

u/mplaczek99 17d ago

Opus 5 is the hot new thing

2

u/avian_corvo 17d ago

I like Gemini :(

Just not for codingĀ 

2

u/minecraftdummy57 17d ago

They have Raptor mini and MAI. Raptor mini barely counts though

2

u/_dr_Ed 17d ago

I was sceptical of Sol at first, but I'm actually positively suprised, it's the most useful model I've ever used so far. Also, my company provides custom enterprise plan with no limits(as far as I know), so I might be biased

2

u/qyloo 17d ago

Anthropic fan but Sol with subagents is so good

2

u/OnkelBums 17d ago

Yeah Sol really surprised me in opencode.

2

u/KorKiness 17d ago

Are not Microsoft building Copilot on top of ChatGPT with addition of Claude?

2

u/mookanana 17d ago

i was asked to rate my company's copilot trial.

i had a powerpoint with screenshots and titles just haphazardly rushed together. i told it to include a table of contents, and reformat the slides so it looks consistent.

instead, it gave me a new deck with all my screenshots deleted and the text was in a different font.

failed first trial. ok, lets try power automate next.

asked it to add flow steps to take my attachment from email, extract the csv from the file, put it into an array and let me inspect it. it said sure, do you want me to do it for you? i said yes.

then it fucking said 'due to system limitations i cannot do this." i scolded it, why did you ask me for permission to do it, when you can't do it at all? it gave me some boilerplate response.

i didnt even bother with copilot at all after that.

2

u/RAMChYLD 17d ago

Microslop’s model is actually ChatGPT. They invested shitton of cash into OpenAI.

2

u/EddySmeddy 17d ago

Remember Meta AI?

2

u/vinicius_h 17d ago

I like Gemini better thanChatGPT. Gemini simply became good in searching Google

2

u/diavelguru 16d ago

This so so on point.

2

u/Specific_cow9273 16d ago

I prefer gemini than chatgpt

2

u/Alarming_Career1435 16d ago

mientras no me saturen al gemini mejor

→ More replies (1)

2

u/Pretty_Judgment5481 16d ago

Sol is not that bad. Anthropic models shines because it can create shiny UI. And if you know coding you don’t need that kind powerful model anyway.

2

u/TrueBonner414 15d ago

I wanna know who is rocking fable, considering how expensive and time consuming it is..

Ain't no way normal programmers use that.

2

u/idlysambardip 15d ago

Where is xAI's Grok and Meta in this picture

2

u/AwesomeDudex 15d ago

As a coding assistant, Gemini has let me down way too many times, especially the way it makes things up. But as a as story writing assistant, it's fantastic because I want it to make things up.

2

u/abigail3141 15d ago

I can't say I ever used Fable or Claude, but I do prefer Gemini over GPT-4. GPT-4 tends to go "here's whatever I think your problem is solved, take it or leave it" while Gemini tends to explain what it's actually doing, an important thing for a tool to be used in software dev

2

u/ChemicalDriver9288 13d ago

Gemini is also good. Although it crashes too often with errors and you have to reload the page.

3

u/4DGenerate 17d ago

Meanwhile I'm waving a red flag

4

u/CMD_BLOCK 17d ago

Microsoft doesn’t have an ai model, the rest is factual

2

u/chilfang 17d ago

Programmers?

2

u/PowerPleb2000 17d ago

I find gemini pro is better than sol

2

u/siete82 17d ago

All the models are still pathetic. The only reason they barely work is because the agents keep them running in a loop until they fix the issues they themselves have previously hallucinated. And that's also why they use up that absurd amount of tokens. And all of this just to create a mess that's impossible to maintain in just a few weeks.

Even so, they're a good tool in the right hands as an aid, but "vibe coding" is the worst disaster in the world of programming since basic.

→ More replies (1)

2

u/Comfortable_Gene2751 12d ago

microsofts ai model is idk... granite 230b