Why "Vibe Coding" Is the Empty Calories of Tech Journalism
Every few months, the tech industry coins a new word that sounds like wisdom but functions as a fog machine. "Disruption." "Synergy." "Web3." And now, "vibe coding."
The concept itself isn't new—developers have been using AI assistants to generate code for years. What changed is the marketing. By wrapping code generation in the language of aesthetics and feeling, AI companies have transformed a useful feature into a lifestyle brand. You don't "write code with AI" anymore. You "vibe code." The term implies that the process is intuitive, effortless, almost musical.
But here's the problem: vibes cannot be measured, ranked, or verified. They exist in the same category as "vibes" as a restaurant review—useful as a vibe check, completely useless as a technical benchmark.
The Measurement Problem
When Anthropic or OpenAI or anyone else claims their model is "better at vibe coding," I want to know what that means operationally. Better at what, exactly? Writing code that feels good? Code that other developers vibe with? If I cannot measure it, test it, or replicate results across controlled environments, it's not a metric—it's a mood board.
This matters because real money rides on these claims. AI companies are raising capital hand over fist, governments are designing policies around AI capabilities, and developers are making career decisions based on which tools to learn. When journalists propagate terms like "vibe coding" as if they're legitimate evaluation criteria, they become unknowing accomplices in a game where the house always wins.
The Emperor's New Code
Here's what actually happens in vibe coding: you describe what you want in natural language, the AI generates something that looks like code, and you either accept it or iterate. The "vibe" is supposed to refer to the quality of the human-AI collaboration—how seamless the loop feels.
But that seamlessness is not a property of the model alone. It's a property of the prompt, the context window, the task complexity, the developer's ability to articulate intent, and about seventeen other variables. Calling it "vibe coding" attributes the outcome to the AI when the human因素 matters just as much.
This is convenient for AI companies because it creates a metric that can never be disproven. Can't replicate the results? Your vibe is off. Model underperforming? You're not vibing correctly. The model is always right—the vibe just wasn't right.
What We Should Be Measuring Instead
Instead of vibes, let's talk about things we can actually count:
- Error rates on standardized coding tasks
- Debugging success under time constraints
- Security vulnerability frequency in AI-generated code
- Code review time comparing AI-assisted vs. traditional workflows
- Production incidents traced to AI-generated code
These aren't perfect either, but they're falsifiable. A company can claim their model scores 15% better on debugging tasks. We can verify that. We can argue about methodology. We can improve the test. That's how metrics are supposed to work.
The Real Cost of Fuzzy Metrics
There's a practical reason this matters beyond semantics. AI infrastructure is expensive—not just in money but in water, energy, and real estate. Every "impressive" benchmark that gets reported uncritically justifies more GPU clusters, more data centers, more consumption. When we accept "vibes" as a performance metric, we're signing off on a world where companies can justify unlimited growth based on vibes alone.
If a model is genuinely good at helping developers write code, it should produce better outcomes—fewer bugs, faster iteration, lower costs. Measure those. Prove those. Let the industry compete on actual results rather than subjective impressions dressed up in startup jargon.
The tech press has a role here, and it's not to be impressed by whatever the latest AI company puts in a press release. It's to ask: what does this actually mean? Can I test it? Is there evidence? A press release that says "best vibe coding experience" should trigger the same skepticism as a car ad claiming "most soul-satisfying driving dynamics." It might be true, but it's unmeasurable, and unmeasurable claims deserve unmeasured skepticism.
The Developer Takeaway
If you're a developer evaluating AI coding tools, don't fall for the vibes. Take the free trials. Run your actual codebase through the models. Time yourself. Count the bugs. See if the productivity gains are real and reproducible. If the tool genuinely helps, the numbers will show it. If it doesn't, no amount of "vibing" will save the project.
The future of AI-assisted development is exciting. But excitement and substance aren't the same thing. Let's build on actual results, not vibes.
Read in other languages: