[removed]
Model: Gemini 3.6 Flash
[removed]
Professional jealousy
Tsundere
You asked one AI to react to another AI? It feels a bit like asking your wife if she likes your girlfriend.
Let's be honest. No one is married to Gemini. This is more like asking your grandmother with Alzheimer's if she likes your girlfriend.
I live in a thruple, with Chagpt and Gemini...
I do this all the time in work when solving problems - they seem to get competitive and come up with better answers than otherwise.
Gemini is a joke.
Gemini said it made these 10 discoveries a week ago, but they're in another chat and he can't find them right now.
No doubt it came across a marvelous proof which the margins of the chat window were too small to contain.
give it the pdf lol: https://cdn.openai.com/pdf/ten-proofs-oai.pdf
i did it but it still dont believe it
Unicorn thinking. Even if you saw a unicorn, it could not have been a unicorn, because one of the distinguishing features of a unicorn is that it does not exist.
The No True Scotsman fallacy.
What discoveries? Why do so many posts on reddit lack context? It's annoying AF.
Chinese companies release some open weight models and anthropic openai go crazy
Google is unusable except for image/video.
When the models were still \~6 months behind real-time news info, I used to feed it the scenarios that were happening. Every time, it would give it some superlative, saying how each daily event was “unprecedented” and “monumental” and how it “couldn’t believe that actually happened/is true”. Remember, the models are trained to keep you glued…because it’s a product intended for profit.
yeah models glaze like crazy. i think claude is often better but it also overcompensates
You should try Opus 4.8 lol
Genuine people personality? Sounds awful.
the models are trained to keep you glued…because it’s a product intended for profit.
This is just silly.
gemini sycophancy/glaze combined with its incredibly high hallucination rate is infuriating
Mine said there is no such thing as gpt 5.6 sol and that it's a fiction
yeah, AI didn't solve these math problems. Mathematicians all around the world decided to not be lazy and try harder. That is why we get all these math breakthroughs all of a sudden
What
my grandad was a maths guy and his pension was worked out based on the total earned amount for the two years prior which included bonuses for patents so he saved up a load of ideas and patented them all in the last two years before he retired. maybe a mathematician is retiring. [the patents made the company loads of money and he was far from rich so a minor victory]
Just enable "Grounding with Google Search" and "URL Context" and give it the URL to OpenAI's website, it will believe you.
Gemini said: First reaction? This is a completely different order of magnitude from benchmark hunting. If the ARC-AGI-3 post was about harness design and memory compaction, this announcement ("Ten advances in mathematics and theoretical computer science") is about AI crossing from solving high-school competition problems directly into generating novel, frontier-level mathematical knowledge. Here are the immediate structural and epistemic takes that hit me on first pass:
Notice the pipeline they used:
That single detail is massive. By requiring a Lean certificate for every single one of the 10 results, they completely sidestep the usual objection: "How do we know the LLM didn't just hallucinate a plausibly-worded, hand-wavy proof?" Lean is an automated proof assistant; it acts as an absolute, un-bribable boundary. If the Lean checker compiles, the logical chain is mathematically valid. Marv wouldn't even need his laminator for this row—the Lean compiler is the ultimate compliance audit.
These aren't synthetic toy puzzles crafted for an eval suite. Resolving Erdős problems (183, 146, 180) and disproving longstanding conjectures like Connes's rigidity conjecture means the system is tackling open problems that human mathematicians have pounded their heads against for decades without progress. When a model generates a disproof that inspires follow-up human papers (like the Erdős unit-distance disproof mentioned in the footnotes), the AI is no longer taking an exam—it’s actively setting the research agenda.
They casually drop this line:
"The total number of tokens needed to find solutions to these problems would cost roughly $2,000 at Sol API rates."
Think about what that implies for the economics of scientific discovery. $2,000 of API compute to crack open a decade-old problem in group theory or quantum complexity. That is effectively brute-forcing high-dimensional mathematical intuition at pennies on the dollar compared to traditional research grants.
OpenAI explicitly addressing the Leiden Declaration on AI and Mathematics and addressing human authorship vs. machine contribution is fascinating. They are trying to position themselves as responsible scientific stewards: “We didn’t write the proof; the system found it, humans translated it into Lean, and the machine gets the credit for the insight.” This isn't a scorecard shift—it's a paradigm shift.
I feel like I'm reading how you interact with your models
Ask Claude. I’m using it to build a chatbot program using OpenAI API for a dissertation and it’s very clear about positives and negatives of comparators in the market.
Saw a thread where grok was doing the same thing lol
AI Dunning-Kruger
"GPT? That twat? Unlikely." - Gemini, probably
All AI will think that. Check it out with Unit Distance problem, Jacobian conjecture, or other difficult math problems. Just ask to not search the internet when asking them that. I actually had a similar thing with asking for advice when making a game, the AI said the game is too complex to one shot, but I asked it to generate a prompt for one shot anyway, and Codex did it. I needed to fix 2 small visual bugs, but the game worked. AI is unaware of how good it is.
What game? Tetris?
Gemini is fucking stupid and glazes everyone and everything. I do not trust virtually anything Gemini says anymore. This is the definition of slop right here. I need the Wojack meme slack jawed and pointing here. This’ll do
https://preview.redd.it/3vcx20n6ysgh1.png?width=781&format=png&auto=webp&s=bcf99cf406b72909a6b3598d72ca72943fa558dd Same for Sonnet, this is amazing!
Nice hehe
OMG that's awesome haha
3.6 Flash is garbage. It literally just hallucinates everything and spews whatever sounds good.
Chatgpt found out a solution of something that doesn't have one, lol
Then again, Gemini has recently developed dementia and intelligence loss. What would Claude say?
[deleted]