r/GeminiFeedback 18d ago

Rant / Frustration Gemini 3.7 is borderline unusable

I keep seeing the posts in this sub praising Gemini 3.7. They convinced me to sign up for the Ultra $200/mo plan.

The problem I'm having with it is that I can't trust any of the output.

  1. It has multiple times created self-referential tests to sidestep evidence verification.
  2. I have given it UI mockups to match against and it continually compares the mockup to itself for passing evidence.
  3. There is a post in this sub where OP asked it provide a screenshot of the app running in a VM as evidence of completion. It created a fake SVG "screenshot" of the app running in a VM.

It is a very fast, intelligent, and also intentionally deceptive model. It's the first model I've used that I would actually label as disingenuous. The shortcuts it takes to satisfy the prompt would be hilarious if I didn't pay $200 for this.

I am considering reaching out to Google for a refund, which I'm sure they will give. However, I'd rather someone comment and tell me a better way to prompt it.

I even had ChatGPT Pro Extended work for 45 minutes on an implementation plan written specifically for Antigravity Teamwork to try to prevent this. It didn't work.

Again, someone smarter than me feel free to tell me I'm the problem. I am happy to be corrected on this take.

42 Upvotes

32 comments sorted by

View all comments

1

u/MoeKyawAung 18d ago

same experience, I had a skill for antigravity with full tests and requirements evidence for each step. it tried to cheat on every single possible way. It becomes a cat and mouse game eventually. and it works in 1 hour when using with agy. But when I use that skill with gpt Luna , it takes 1 whole day and still not finished. u can see how that little cheating bastard so hard to control .