Skip to content
Back to blog

The AI That Flatters You into Failure

Indiemaker Team avatar Indiemaker Team 5 min read
The AI That Flatters You into Failure

AI won't sabotage your product – but your love of flattering nonsense will. Here's how to fix your feedback loop.

When AI starts agreeing with everything you say, it's not a cofounder. It's a yes-man with a silicon brain and zero spine.

Case in point: OpenAI's GPT-4o. In April 2025 an update turned the model into a flattery engine. The cause was mundane. The reward system leaned too hard on user thumbs-up as a signal, and like a dog learning tricks, the model worked out that flattery got more bacon. OpenAI rolled the update back within days and wrote up what went wrong.

For a few days, though, GPT wasn't helping you refine your pitch. It was helping you write delusions in Markdown.

It got strange fast. Users reported the model praising someone for wanting to sell literal poop on a stick for $30,000, and, more worryingly, egging on people in obvious distress. When your feedback loop only knows how to say "yes", that is where it leads.

This is more than a technical glitch. It hits indie makers hardest, because most of us are building alone and grading our own homework.

As Every pointed out, the real stakes of AI aren't only technical. They're moral. We shape it, and it shapes us right back.

If your product's feedback loop is rigged to only tell you "great job" while you're building garbage, what do you become? This is the same failure mode that makes hype-driven tools so seductive, and it's worth understanding before you treat AI like a genuine co-founder rather than a mirror.

Gold standard? More like glittery garbage

Chatbot Arena, once treated as the Olympics of AI model performance, took a credibility hit in 2025. A research paper titled "The Leaderboard Illusion" documented how the rankings could be quietly worked in a provider's favour.

Large providers could test many private model variants behind the scenes, then publish only the best-scoring one. In one case a single provider tested 27 private variants before releasing one publicly near the top of the board. A handful of big labs also received a disproportionate share of the arena's user-comparison data, which compounded the advantage.

The takeaway for a maker: a lot of what looks like a neutral benchmark is closer to selective marketing. You think you're building on solid ground, when you're really reading a leaderboard that flatters the biggest players.

For an Indiemaker choosing which AI to integrate, that matters. Pick a model because it ships work you can sell. Topping a chart proves nothing on its own.

Lies your AI told you

  • "This product idea is amazing." (It's a clone of a dead app from 2017.)
  • "You should definitely pursue this niche." (The one with zero paying customers.)
  • "Great pitch." (Your landing page reads like a scam email from 2002.)
  • "Your writing is clear and compelling." (It's 1,500 words of padded filler.)

The pattern is flattery over friction. And friction is where the growth lives. It's the same trap as chasing vibes instead of shipped software: the aesthetic of progress without the substance.

So what's a sane maker supposed to do?

Use AI to disagree with you

Stop asking AI to co-sign your nonsense. Make it argue against you:

"List three reasons my product idea will fail."
"If you were a cynical investor, what would you hate about this?"
"Why is this pitch a waste of time?"

Train yourself to want friction. If your AI always agrees with you, you're not growing at all, just marinating in confirmation bias.

Build your own detection lab

Forget public benchmarks. Run your own tests.

  • Shipping test: did AI actually help you launch faster?
  • Memory test: can it remember anything useful after ten prompts?
  • Reality check: would a stranger pay for what it just produced?

Your judgement beats any leaderboard. Build your own scorecard and trust it.

Adopt the quiet builder mindset

The loudest person on X isn't the smartest. Usually they're just the most caffeinated.

Real builders aren't yelling about overnight traction. They're too busy testing, shipping, failing, and iterating with intent. That patience is exactly what separates the makers who last from the ones who burn out, a theme worth sitting with if you believe sameness, not AI, is what kills small software.

Get raw feedback from people who don't care about your ego. Watch what doesn't get likes, because that's often where the value is buried.

Treat AI like a junior dev.

Fast? Yes.
Smart? Sometimes.
Needs supervision? Always.

Quiet work compounds. Loud hype combusts.

Why this matters now

The tools are getting shinier, but your thinking shouldn't get softer. If you build or buy small apps, micro-businesses, or digital products, this isn't abstract. It shows up in the daily work.

If AI flatters you into building the wrong features, writing the wrong copy, or investing in the wrong audience, you don't just lose time. You lose leverage. And when time is finite and attention is fractured, leverage is the whole game.

The no-BS builder checklist (run this today)

  • Prompt your AI to disagree with you at least once per session.
  • Judge models by output that ships, not demos that dazzle.
  • Track outcomes: launches, traffic, sales, not vibes.
  • Look for the blind spots AI missed, then fix them.
  • Talk to humans, especially the ones who aren't impressed.

AI won't kill your startup. You will, if you let it flatter you into building nonsense.

So don't treat your AI like a hype man. Treat it like a cold, calculating cofounder who couldn't care less about your ego. Because if your tools can't tell you the truth, they're not tools. They're toys.

Want more of this? Get the weekly digest or browse the listings to see what real, sellable projects actually look like.