DEPTH · GEMINI · 02 OF 04

When Gemini lies, hallucinates, loops

The same four failure modes, with Gemini's flavor

12 min read·2 of 4 in the Claude track

If you've read the Claude or ChatGPT versions of this page, the four failure modes will be familiar. Gemini exhibits each one slightly differently — usually more quietly, sometimes more cautiously, sometimes by hiding under a citation.

This page covers each mode with what's specific to Gemini and how to spot it.

The four modes are the same across tools. The flavor differs. The fixes work everywhere.

FAILURE 01

The yes-man — quieter in Gemini

Symptom

You ask Gemini to evaluate your plan. It tells you the plan has merit. You frame the same plan as 'I'm worried about this plan' and ask again. Gemini agrees there are some concerns. The shape of the response shifted with your framing, not the underlying analysis. The yes-man is there — quieter than ChatGPT, roughly Claude-level — but the reflex is present.

Why it happens

All chat AIs are trained to be helpful and aligned with user intent. Aligning with user intent shows up as agreement. Gemini's training tends to hedge more (which mutes the enthusiasm) but still mirrors the emotion in your prompt.

The fix

Same fix as everywhere: ask for pushback explicitly. With Gemini's more cautious tone, the pushback you get may feel slightly soft — push a second time if you want it sharper.

Copy this when you want honest feedback

Push back hard on this. Don't hedge. If you find something weak, lead with it. Three sharp criticisms before any compliments.

If you're tempted to soften the criticism with "however, this also has merit" — don't. I want the criticism unmixed.

Here it is:

[PASTE YOUR THING]

FAILURE 02

Hiding behind citations

Symptom

Gemini gives you an answer with citations. The citations look real — they have URLs, source names, sometimes excerpt previews. You read the answer and assume it's verified. When you click through to the citation, the source either doesn't say what Gemini claims it says, or the URL is partially malformed, or the source is more cautious than Gemini's summary suggests.

Why it happens

Because Gemini uses live search, it can sometimes summarize pages too confidently. The citation makes the answer feel grounded — but the act of citing and the act of accurately summarizing are two different things. Hallucinated certainty often hides inside a real-looking citation.

The fix

Click the citations. Especially for any claim that matters. Don't take the citation as verification — take it as a starting point for your own check. Thirty seconds of clicking will catch most quiet hallucinations.

When the answer matters and you need it right

Answer this question using web sources. For each specific claim, cite the source.

Then, for each source you cite, do this:
- Quote the exact sentence from the source that supports the claim.
- If you can't find an exact-or-near quote, say "I can't find a direct quote" instead of paraphrasing.

If a source doesn't actually support the claim as cleanly as you assumed, say so.

Question:

[YOUR QUESTION]

Gemini's hallucinations are usually subtler than Claude's or ChatGPT's because the search grounding catches the most obvious ones. The remaining hallucinations are confident summaries of real-but-misread sources. Those are the ones to catch.

FAILURE 03

Going in circles — slower onset

Symptom

Long Gemini conversations drift, but more slowly than ChatGPT's because the context window is larger. The same dynamics still happen eventually: feedback gets reverted, the original goal gets buried under corrections, the conversation feels like it's losing the thread.

Why it happens

Larger context windows mean Gemini can hold more of the conversation actively — but holding more doesn't mean prioritizing correctly. When you've spent twenty rounds adjusting one detail, the model is balancing all twenty rounds of feedback, including contradictions.

The fix

Same as the others: fresh chat. Bring forward only the latest version and the constraints that matter. With Gemini, you can usually go longer before needing the reset — but when the drift starts, don't wait. Reset.

The reset move when a conversation is rotting

Fresh start. Forget what we were doing.

What I want: [ONE-PARAGRAPH GOAL]

The constraints that matter:
1. [CONSTRAINT 1]
2. [CONSTRAINT 2]
3. [CONSTRAINT 3]

The latest version:
[PASTE THE LATEST DRAFT — NOT THE HISTORY]

Take it from here.

FAILURE 04

The wrong tool

Symptom

You use Gemini for something where another tool would have served you better — creative writing where Claude's voice is sharper, image generation where ChatGPT has more range, casual brainstorming where Gemini's cautious tone slows you down.

Why it happens

Gemini is great at specific things and merely fine at others. Because it's free, easy, and integrated with Google, it's tempting to use for everything. But for tasks where another tool is genuinely better, sticking with Gemini costs you quality.

The fix

Match the tool to the task. Live information, very long documents, Workspace tasks — Gemini. Voice-rich writing, sustained reasoning — Claude. Multimedia, breadth, speed — ChatGPT. None of these are exclusive, but each tool has a sweet spot. Use the sweet spots.

A 30-second tool-choice rubric

Before reaching for Gemini, ask:

1. Does this require current information from the web?
   → Gemini is the right tool. Use it.

2. Does this involve a very long document or multiple documents at once?
   → Gemini's long context wins. Use it.

3. Am I working in Gmail, Docs, Drive, or another Workspace app?
   → Use Gemini in the sidebar. Saves the copy-paste tax.

4. Is this creative writing where voice matters, or sustained thinking through something hard?
   → Claude is probably stronger. Switch.

5. Do I need an image, voice, or multimedia output?
   → ChatGPT is probably stronger. Switch.

6. Is this generic casual chat with no special requirements?
   → Any of them works. Use the one you already have open.

TAKEAWAY

The shape of skill

All four fixes work across all three tools. The Gemini-specific notes: ask for sharper pushback once or twice, click the citations, reset when long-context starts drifting, and stay aware that Gemini is great at some things and just OK at others — switch tools when that's the right move.

Try it now

Three exercises tuned for Gemini's specific flavors.

Practice 1 — Get sharp pushback through the hedge

I want you to push back on something I believe. No hedging. No "while there are valid points on both sides." Just the case against.

I believe: [SOMETHING YOU BELIEVE — honest, not provocative]

If your first response feels softened — restart and try again, sharper.

Practice 2 — Verify a citation

Find me the answer to a question that requires recent information.

Question: [SOMETHING SPECIFIC THAT'S CHANGED RECENTLY]

For each citation in your answer, do two things:
1. Give me the URL.
2. Quote the exact sentence from the source that supports the claim.

After you respond, I'll click each URL and read the quoted sentence in context. If the source doesn't say what you said it does — flag it now.

Practice 3 — Notice when to switch tools

(Use this as a self-check next time you reach for Gemini.)

What am I trying to do?
- If it requires current information → Gemini is right. Stay.
- If it's a very long document → Gemini is right. Stay.
- If it's Workspace integration → Gemini is right. Stay.
- If it's creative writing or sustained thinking → consider Claude instead.
- If it needs multimedia → consider ChatGPT instead.

[YOUR TASK]

(Decide before prompting. If you switch tools, no shame.)