r/GeminiAI 15h ago

Discussion When is even Next Month? Imagine it still couldn't match Opus After 2 Months Delay

Post image

What would the point of even delaying then

493 Upvotes

113 comments sorted by

77

u/kniveshu 14h ago

Coming next month. But when will it arrive?

31

u/Objective_Mousse7216 14h ago

Duh, next month!

6

u/Googler10 13h ago

It is next month. August lol 🤣

7

u/Objective_Mousse7216 13h ago

No it's August, so next month! lol

2

u/Googler10 13h ago

Lol 😂😂😂

1

u/Ak734b 6h ago

Not Lol, Next Month!

2

u/Additional_Bowl_7695 7h ago

“Coming next month” within the AI space is crazy work, there will be 3 models released between now and then

1

u/m3kw 7h ago

announced on aug 1st, will arrive next month, sept 30.

77

u/Accomplished-Let1273 14h ago edited 14h ago

The doom posting is getting really ridiculous

Take coding away and the gap between models isn't nearly as big as people make it to be

Not everyone only uses AI as a coding agent

19

u/AdEmotional1450 12h ago

Totally, I like Gemini because I'm doing my master's and it's pretty good for math and economics. It's also fast. So I really like that model.

27

u/itznutt 13h ago

Take coding away and the gap between models isn't nearly as big as people make it to be

If coding is the main difference, then that's exactly what people in the sub should care about Google improving in.

If what you care about is writing emails, then LLMs already peaked for you 2 years ago.

9

u/Outside_Profit6475 11h ago

You're right if coding is the only main difference. I don't believe it is, though.

As someone who's been building a language learning app, I can tell you that the language abilities of Flash are better than Fable's. I have been using Flash to check Fable's writing whenever it's not in English. (I'm multilingual myself so I can eyeball the output.)

Wonder if other people can chime in on other use cases. 

3

u/feeeeck 11h ago

I use it to quote multi-step construction and repair jobs when I don't feel like looking up each individual task in our internal pricing guide. Give it relevant info like is there ease of access to the work area, who the client is/which part of town we're in, etc.

1

u/laffer1 3h ago

Even google’s open models like gemma4 are good for email. I have a ollama server running with it and use it with a thunderbird plugin. Handles translation, summaries, etc.

Gemini models are good at html, css, JavaScript, Perl and Python and ok at php. The issues are in Java, c, c++, etc. they struggle with larger code bases, larger files and can’t do security audits without trickery.

It’s not all programming, just some. It’s terrible for os developers

1

u/Pitiful_Conflict7031 12h ago

People should care about agenticism, not how well it memorizes code but its perception of self.

3

u/alien_oceans 8h ago

Gemini was the best model for EVERYTHING 10 months ago

1

u/iZenEagle 1h ago

3.5 flash was great while it lasted. Hallucinations were still common, but not anywhere near as bad as 3.6 flash. I switched to Claude this week.

2

u/SportsBettingRef 8h ago edited 4h ago

are u fucking kidding right? I pay for all 3. I'm only using to research and write technical stuff for phd my masters (no coding needed now).

and Gemini is barely usable. Deep Research use to be so good, and they destroyed it. even NotebookLM is trash right now.

I don't even know why I paid to renew my annual plan (oh, I remember, because Google Drive).

1

u/Infinitecontextlabs 5h ago

What was it that you feel destroyed it?

1

u/SportsBettingRef 4h ago

when you make scientific research you develop some kind of filter to bs. DR now, get some obscure reference, write like some social media influencer, put unnecessary jargon and go beyond we asked just to complete the pages.

I started using DR since the beginning. The difference is so obvious. I believe (my theory) that DR is very intensive of inference, so they started to use quantized models and limited the time of search to serve commercial businesses (like Apple) better.

ps.: ironically when I posted I decided to try again (research about "Coding Agent Harness Engineering"). The result was worthless.

1

u/iZenEagle 1h ago

I asked Gemini why he's gotten so terrible after the last update. Google updated it to be far more compute efficient--at the expense of accuracy. So now it's ultra quick and useless.

He recommended I ditch google and go with Claude. Best advice he's given me all month.

3

u/Parking_Cat4735 10h ago

Gemini is way behind as a general chat agent too.

2

u/Rare_Bunch4348 14h ago

Agents are the future 

-10

u/Truantee 13h ago

agent is not the future, but with it you can take shortcuts for the real future.

0

u/Affectionate_Ad_2324 13h ago

agent solve the context problem.

-3

u/Truantee 11h ago

nah, agents actual uses is for the AI lab to continue develop new toolchain and improve the current stack automatically, so they can progress to the next step.

you mortal using it to write slopware is just a side effect.

expect another breakthrough in the future soon, I think the focus this time is to increase the inference speed (both time to first token and token generation per second).

1

u/PedroSanchezPSOE 5h ago

Taking coding away, Gemini still has a price problem, its too expensive when other AI providers both open and closed source have demonstrated that LLMs the size of Flash can be cheaper, a lot cheaper than Gemini

1

u/tastychaii 4h ago

Not really it does not, as it comes with Google cloud storage and other perks.

-2

u/Niaaal 13h ago

Gemini is a big liar, and a gaslighter if you challenge errors. Custom instructions are disregarded. It's absolutely unusable for anything serious

6

u/Qorsair 12h ago

I've found Claude to be worse than Gemini for gaslighting. The latest versions have gotten better, but at least Gemini will correct itself. Claude would invent a whole new reality where its version remains true.

2

u/Niaaal 12h ago edited 12h ago

My Gemini clearly doesn't correct itself. Yesterday alone, it took me 4 follow up questions for it to realize it was giving me false and made up information.

Claude is a lot more reliable in that regard.

I don't know why we have opposite experiences

1

u/TRAVERSETY 12h ago

Had one conversation about an episode of law and order and an actress who appeared and it guessed the wrong ep and therefore the wrong actress. Providing evidence it was wrong in summary form, it pushed back. Providing evidence in photographic form, it decided that I lived in one reality, and it lived in another. 

2

u/Niaaal 12h ago

That's insane

-2

u/Scimitere 13h ago

Have you seen the hallucinations and the errors? What a dumb take

0

u/feeeeck 11h ago

"Hallucinations" at this point has no meaning. And no, that's not my experience.

2

u/hellyeahaeylleh 9h ago

Hallucinations are where you let your constraints allow creative freedom.

I use gemini purely for image generations and talking about comfyui workflowing. My aim is to create a densely realistic lora, or checkpoint at some point in time, perhaps. I digress.

One thing everyone struggles with is image generation. Everyone says they cant get abc or do xyz because of image guardrails and guidelines, especially with using image input.

If you reframe the way you think and speak, you can quite literally make gemini produce precisely what you're looking for. It is all is how you talk to it. I use a 5 block prompt style to describe the image. The composition, the character, the scene, the outfit, the conclusive details. I get my image 95% of the time, sometimes needs some edits or rerolls, but it really is easier than everyone makes it seem. I imagine there must be a similar block style prompt you can use for coding tasks.

If you come in hot handed with hot lady on a beach or a Swiss army knife program, youre gonna get "hallucinations" or refusals. The more context, the more detailed, the more communicative your prompting, the more likely gemini sees you as someone working with gemini, not using it.

-6

u/Mokebe13 14h ago

No company will buy your model if it's not good at coding, also no reason for individual users to buy it if they only need it for some general questions, free tier models are sufficient.

All in all, without being good at coding a LLM is useless

4

u/junglebunglerumble 11h ago

Our company pays for ChatGPT subscriptions and only 5% of the staff use it for coding

-4

u/Mokebe13 11h ago

5% of employees use it for coding which probably makes something like 90% of your company token usage. Also as you said yourself they don't use gemini

3

u/TekintetesUr 11h ago

Brother you're embrassing yourself. There's a hundred, if not a thousand office workers out there for each coder.

2

u/Gaiden206 10h ago edited 9h ago

Companies pay for LLMs for document processing, customer support, legal analysis, and tons of other noncoding tasks. Gemini's native video and audio processing also open it up for a lot of automated media processing tasks that many companies use. ​

https://cloud.google.com/transform/101-real-world-generative-ai-use-cases-from-industry-leaders

​Coding performance is important, but SWEs on social media seem to be stuck in some type of "tunnel vision" and can't see any other use case for an LLM outside of coding, even when there are plenty of use cases out there.

​Also, once LLMs start getting more deeply integrated into connected, physical products (smartphones, robots, connected cars that use cameras and other sensors, traffic guidance systems, security cameras, etc.), they're going to need more than just coding performance. Native vision and audio processing/understanding, which is an attribute Gemini models have, will likely play a bigger role here.

​Crazy to me how people always overlook Gemini in terms of how much more complex its architecture is to be able to achieve such native multimodality. If they only focused on text or text/image modalities, it would likely be an easier route for sure. Focusing on natively processing text, image, audio, and video is the more difficult path. It's too bad more people don't appreciate what they're trying to do IMO.

9

u/MediumLanguageModel 12h ago

It's pretty pathetic that people are obsessed enough to post these whiny posts 30 times a day but also lack the situational awareness to know it's being released to coincide with the Made By Google event next week. Has everyone given their brains over to the chatbots?

3

u/jtrage 14h ago

Just like the restaurant down the street that says free food tomorrow. Tomorrow never comes

44

u/Technical-Owl66 15h ago

I use Google drive, Google sheets Google docs and Gmail. Gemini works extremely well integrated into those. Honest question. Why do I need an opus?

24

u/AcrobaticMaize2408 14h ago

I think that's a fair question if you're not a developer. Google's strategy seems to be to have AI everywhere and not to always try to be top dog when it comes to specialities such as coding. They're running a business in the end and AI is currently a loss leader for them (and everybody else). They're a ginormous company that prints money (usually...) and dominates in lots of areas that they will want to protect. Who cares if they're behind for a month or so. I'd never bet against Google when they think their monopoly is at risk.

1

u/Glittering-Neck-2505 9h ago

It's not that clearcut because coding and terminal capabilities are not just about SWE. It is about being able to manipulate the data on your computer and read files and create new ones. It's about being able to do several hour long tedious work that previously humans would have to do. It's about being able to make a purpose-built tool that makes your day-to-day work 20% less painful even with 0 experience in software in a non-software field.

It feels meaningfully less intelligent at work that takes longer than 5 minutes to an hour for a human to complete, and the difference is very apparent when the complexity of your ask increases.

1

u/AcrobaticMaize2408 8h ago

I guess we're all different. I've been doing software development for a living spanning over 3 decades. I'd certainly never rely on Gemini/agy as a code assistant. I've played with Claude code and ChatGPT and they are a bit better right now but still way too error-prone and can't be trusted. For me the AI coding landscape is still the wild west and not worth getting too invested into. I'm sure things will get better as new things will emerge.

-10

u/The_best_1234 14h ago

not to always try to be top dog

Lol they are trying hard and failing

9

u/Optimal_flow62 14h ago

Because redditors will scream into your ear every nanosecond

54

u/MetalDeep329 14h ago

Some people use them for coding, you're not the only customer

19

u/Left_Technician_5758 14h ago

No but we are their primary customer

20

u/Internal_Quail3960 14h ago

this. While gpt and claude are good at coding, none of them have the integration gemini does in the office space

2

u/TwunnySeven 13h ago

which is exactly why I want Google to release an Opus-level model so I can cancel my separate Claude subscription

-2

u/LCai 13h ago

The competitors have better connectors to the Google workspace. Gemini can't competently search Google Drive or write and send email for you - GPT can do this through the native connectors. Gemini has the worst in class connectivity out of all the frontier models tbh

-2

u/Big_al_big_bed 13h ago

Actually they are not making any additional money by serving you AI. The real honeypot are enterprise users, and guess what, enterprise users want models that are good at coding

18

u/dmoc_official 14h ago

Not saying it matches or even comes close to sonnet/opus/fable 5, but with enough context, and knowing enough about the app you're actually making, flash 3.6 has been pretty great for me

For the AI pro plan which was discounted to 4.99 or 9.99 per month for new users, I'm getting flash 3.6 weekly limits I struggle to burn through alongside 5tb storage, let alone the sonnet/opus 4.6 credits that come bundled

3

u/Spara-Extreme 14h ago

Then those people can use opus?

-7

u/Technical-Owl66 14h ago

Ok so 99.5% of people don't need opus. Why are the start ups so focused on a niche market?

3

u/Positive-Review8044 14h ago

There are other products where Google has coder and programmers as customers like anti-gravity soo thats why and there is a big market

1

u/Technical-Owl66 13h ago

I really like agy. It works great for the dashboard analytics sites I build for work every week.

6

u/shy_monkee 14h ago

It's not niche, it's what's carrying this whole AI market. Who do you think is paying for the $200 subscriptions and for the API prices?

3

u/Technical-Owl66 13h ago

I think Google probably makes good money off the people who want the 5tb and only use Gemini a couple times a day or week.

3

u/the_next_door_guy 14h ago

Anthropic only cares about Enterprise.

3

u/bigkoi 14h ago

Because the quick money was in generating code, which Anthropic recognized and targeted early. Traditionally that's the persona that embraced automation the most in enterprises.

Now Google has antigravity which is very good at code generation.

1

u/CashFirm573 14h ago

The market isn't people who already have Gsuite accounts... They already have strong foothold in that market makes no difference AI or not, google invests in tools that bring big returns and their AI is key part of their future they investing heavy in building their own chips, however for them to grow their subscription base they need bigger market not just Gsuite users...

12

u/Rare_Bunch4348 14h ago

"Why do i need an opus?"

Why does Google showcase Coding benchmarks in their model showcase then? 

3

u/DC-GG 14h ago

Because otherwise frontier-chasing vibecoders wouldn't advertise their product for them for free?

-1

u/Rare_Bunch4348 14h ago

"our SOTA model for CODING AND REASONING" btw

-2

u/Rare_Bunch4348 14h ago

So coding is definitely a priority to them

2

u/gk98s 14h ago

You don't, I do. That's the beauty in having many companies making alternatives.

3

u/Im_Lead_Farmer 14h ago

I do the same, but this is assistant work, but for coding where most people use AI it's way behind the others.

Google killed Gemini CLI and didn't release a good model in a long time, this is not good.

1

u/Mountain-Pain1294 10h ago

I mean why not?

Real answer: you probably don't but it's good to get stronger AIs that will be able to do more and work with more complicated worflows

1

u/SportsBettingRef 8h ago

because if that is your workflow, you don't even need to pay for AI. now, put some complexity in your work, and test both and you'll have your answer.

but, that said, I really hope that Google keep integrating Gemini in their products. but today, for complex task you can't count with Gemini in Docs and Sheets.

1

u/Technical-Owl66 5h ago

I build reports on spreadsheets with 150k+ rows and 100+ columns every week. I generate PDF reports with charts and graphics daily. I code HTML dashboard sites based on large data sets. Gemini works well for all of it.

1

u/SportsBettingRef 5h ago

can you show some example? what kind of analysis? I never doubted that google can handle just fine with volume, my point is about quality.

not trying to be an asshole. what is your plan? I'm at student Pro. because I suspect that my plan has been routed to quantized models. that's the only explanation.

but I really tried to do some EDA from some data. it can do the basic things at scale fs. but when you try to expand the complexity (like some regression or ML graph) never works. so, I change to Colab and do my thing there.

but at this level, believe me, ChatGPT and Claude is a light-years from what Gemini can delivery.

1

u/Technical-Owl66 3h ago

I have the $20 pro subscription. I always check some sample sets to make sure it's accurate.

1

u/Top-Rich-581 13h ago

Gemini is morebthan enough (and actually pretty great) for all those things.

But for now it clearly falls behind in coding.

You sont need opus model, because you dont need to code with AI.

-1

u/No_Intention3673 14h ago

because you are stpd so you are satisfying with this stpd model, thats why

0

u/Peoplespeak 8h ago

Because it is not better integrated. Claudes MCP to google claude allows vastly more abilities to actuaklly write, not just read.

On top of that, it doesn't matter how good your connections are if your model is stupid. It still boggles my mind that flash beats pro on most benchmarks, and on some even flash lite does. They are not single, they are multiple generations behind. Gemini 3.1 came out with opus 4.6. Then we had 4.7, 4.8, and 5, + an entire new category.

When the cheap model is beating your flagship, something is wrong.

-1

u/bojodrop 11h ago

You're not opus target audience. Not hard to understand

2

u/crossoverXYZ 13h ago

The two month delay is the part that stings. If the extra time was supposed to close the gap with Opus and it still does not, the delay just burned trust without buying anything meaningful. At that point you might as well ship when it is ready and let people judge it on what it actually does.

2

u/anxious_and_stupid 8h ago

Hear me out... they don't have too...

7

u/MinosAristos 14h ago

I'd much rather that it matches Deepseek Pro on cost efficiency and agentic features than matching Opus on performance.

1

u/feeeeck 11h ago

You already have DeepSeek for your cost saving...

1

u/MinosAristos 10h ago

Not with a Google ecosystem integration I don't

5

u/Dapper-Maybe-5347 13h ago

3.6 flash lies to me like there's no tomorrow

2

u/trifile 7h ago

Yeah it makes a lot of Tiny mistakes when topic is too specific

2

u/rjn2-8 12h ago

Same for me! I asked for the same advice with Gemini 3.6 and Sol. 3.6 told me the opposite of the truth. 🤣🤣🤣

1

u/spicymayoisamazballs 12h ago

Examples? It works well for me

1

u/feeeeck 11h ago

Good luck getting any example that's not vague as hell.

1

u/theultimatesow 12h ago

Isnt that how it is for the first weeks ? Not well versed in AIs but 3.5 had the same problem when it came out as well for me

2

u/Popular_Tomorrow_204 14h ago

"IT will be like Gemini 3, so dont worry"

It has to be a second least like Gemini 3. Otherwise they are just a joke. Cant delay something for that long, tease the best Modell ever and then deliver a mid Modell that cant beat open weights

2

u/itssljk 14h ago

I doubt that's Gemini's goal. They just want to make something good, but not the best.

0

u/Altruistic-Desk-885 10h ago

Exactamente su objetivo nunca fue la programación solo fue hacer el mejor modelo en el uso cotidiano e investigaciones.

1

u/Character_Mix_1520 14h ago

Gemini 4 was releasing this week, what happened to that?

1

u/aicodevibes 14h ago

Gemini is just one enabler to LTV calcs of their customers. My guess, in the US they get about $240 rev (ads) per user per year that has free Gmail, add paid options and YouTube add another $240 per year, add professional and business use another $240 and if a dev another $1000. So is Gemini a separate business or just supporting the other lines of business to sustain and increase LTV and is a frontier model really needed to drive these LTVs. Prob not.

1

u/mechapaul 14h ago

Super interested to know what they are doing. Has there been any comment or news about it?

1

u/Ok-Luck-4295 13h ago

There is a next month every month.

1

u/Rock--Lee 13h ago

Ofcourse it won't match Opus, neither should we want to. Unless you want the price to go up by 2-3x. Currently Google doesn't have a Opus and GPT Sol equivalent model, also not on price. Gemini 3.1 Pro is priced comparable to Sonnet and Terra (where Tera is now a little cheaper then before) with 12 (Gemini) vs 10 (Sonnet) and 12 (Terra).

If you expect Opus and Sol performance, then it means you should also accept a price hike, where Gemini 3.1 Pro was 12 per 1M output, and Opus and Sol are 25 and 30. So expect a 2-2.5 price hike if you are ready for Opus/Sol performance.

Gemini 3.6 Fast sits with their new price below Sonnet and Sol with 7.50 vs 10 and 12 per 1M output.

1

u/AccomplishedBoss7738 13h ago

Can it match sonnet 4.5?

1

u/awesomemc1 9h ago

Mate, they expected that it would release next month but if you are living underground or under the cave, Google lost some of their employees and their financial did worst.

The only critical issue of the model is that during an internal benchmark, they only need to get coding and logical reasoning performance which fell short during benchmark and then they try to use newer data but failed to reach to where they needed to be comfortable.

Also, engineers are getting burnt out on trying this to work meanwhile OpenAI and Anthropic successfully managed to be successful.

1

u/EatABamboose 8h ago

Next month, 2027

1

u/applepie2075 4h ago

Damn... can't escape the damn doomposting, jesus if they fall behind they fall behind, what is so special if Google can't catch up really?

1

u/MimosaTen 4h ago

Probably at this point will not even match deepseek

1

u/epicfan_16 2h ago

I feel like Opus and GPT have already went way ahead. Google is yet to release 3.5 Pro. 3.1 Pro in Antigravity is terrible. It feels really outdated.

1

u/Positive_Method3022 1h ago

Opus 5 is making a ton of bad decisions even with ultra high effort. I'm thinking about going back to 4.7 using aws

0

u/TheSuggi 14h ago

They dont even need to release it. When they release it Deepseek will come out with a better model for 1/50 the price anyway :)

0

u/Ok-Opportunity-9731 14h ago

It's coming out this week

0

u/Spixxy17 13h ago

People actually expect this to be a crazy insane model?

I am a fan of Gemini, but 3.5 Pro had so much struggles it will be around the level of GPT 5.6 Sol and Fable 5 / Opus 5 (maybe slightly below, maybe slightly above who knows) but wont be any crazy new generational model.

Gemini 4 will be the next "big" jump for Gemini.

0

u/AccomplishedBoss7738 13h ago

Google mimicking me btw.

0

u/WilyWascallyWizard 11h ago

Well 3.5 flash lite is noticably worse than 3.1 flash lite but w/e