Усі статті

September 5, 2026 · 6 хв читання

Most AI Workout Generators Are a Wrapper With a Skin. Here's the Test

Most AI Workout Generators Are a Wrapper With a Skin. Here's the Test

Slap the letters AI on a fitness product and it sells. The people building low-effort apps figured that out fast.

Right now there's a whole wave of "AI workout plan generators" that are, underneath, one thing: a prompt sent to ChatGPT, with a nicer font on the answer. Some don't even hide it. You can buy prompt packs on Gumroad that are literally copy-paste text marketed as an AI coach. Free web generators spit out a one-shot plan and wave goodbye. Chat-first apps wrap a fitness themed conversation around the same general purpose model you could use yourself for free. The industry has a name for these products: wrappers. A thin layer of skin over someone else's brain.

I build a training engine for a living, so I want to make two arguments today. First, that a bare language model is genuinely bad at coaching you. Second, and this matters just as much, that the established algorithmic apps are NOT wrappers, and the real question you should ask is a different one.

What happens when a language model writes your program

This isn't just my opinion, it's been tested. A TIME journalist spent weeks using ChatGPT as a personal trainer and described the plans as underwhelming and at times incomprehensible, concluding that the model's skill is sounding human, not giving expert recommendations. Tom's Guide ran a similar experiment and warned that following a ChatGPT training program can be ineffective and a fast track to injury without expert guardrails. And in a peer-reviewed study, experienced coaches graded AI-generated training plans and ranked them below proper coaching, with the researchers advising against following them without an expert's review.

Why does a system that writes beautiful essays fail at sets and reps? Because coaching isn't a writing problem. A language model meets you as a stranger every single time. It doesn't remember that you stalled on bench three weeks running. It can't see that you halved your sleep this week. It has no model of your recovery, no record of your lifts, and no consequences when it's wrong. It generates a plausible looking plan the way it generates a plausible looking poem.

Plausible is not the same as right, and in training, the gap between them is where injuries live.

The apps that are not wrappers

Now the fairness part, because I've watched people lump everything with an AI label into one bucket, and it's wrong.

Fitbod is not a ChatGPT wrapper. It's a scoring and ranking system over hundreds of exercises plus a recovery estimate per muscle, built years before the chatbot era. JuggernautAI is an expert system, a powerlifting coach's actual rules encoded in software. Dr. Muscle and Alpha Progression are progression engines driven by the sets you log. RP Hypertrophy adjusts volume from structured fatigue feedback. I critique these apps elsewhere on this blog, mostly about how they handle recovery and how little of their reasoning they show you, but they are real, deterministic engines working from your real data. Calling them wrappers would be false, so I won't.

The market splits cleanly in two. Real engines that are often opaque. And wrappers that are often chatty and transparent-feeling but have no engine at all. Neither half gives you the full package, and knowing which half you're looking at is most of the battle.

How wrappers make money anyway

You might wonder how products this thin survive. The answer explains a lot about the fitness app market.

A wrapper's economics are beautiful, for the builder. There's no engine to develop, no exercise database to maintain, no years of tuning a recovery model against real training data. The cost is a subscription to someone else's model and a weekend of interface work. From there, every sale is nearly pure margin, which is why you see these products advertised so aggressively, and why new ones appear faster than anyone can review them. Some are even self-aware about the category: at least one fitness app now markets itself specifically as "not another GPT wrapper," which tells you how crowded the wrapper shelf has become.

The person paying the real cost is you, twice. Once in subscription money for capability you could get free by opening a chatbot yourself. And once in training time, because a plan with no memory and no model of your body doesn't just fail to help, it points your effort in plausible-sounding wrong directions. Effort is the one thing you can't refund.

None of this means the builders are villains. It means the AI label carries zero information about whether an engine exists behind it. Which is why you need a test.

My four question wrapper test

Here's the test I'd give any friend evaluating an AI fitness app. Four questions.

One: does it persist and progress the same lifts from session to session? Real progressive overload requires memory. If every session is a fresh plan floating free of the last one, you're looking at a generator, not a coach.

Two: does it adjust from your logged performance, or from re-prompting a chat? An engine reacts to what you actually lifted. A wrapper reacts to what you typed.

Three: can you inspect the reasoning? Ask where a number came from. An engine has an answer, a recovery state, a rule, a threshold. A wrapper has a shrug dressed in confident prose, and the research above shows those confident numbers are sometimes simply made up.

Four: does it know what it doesn't know? This is the subtle one. A real system models your recovery and admits when a muscle isn't ready. A wrapper will write you a leg day every time you ask for a leg day, forever, regardless of what your body has been through.

Where Workout With Me sits

You can guess my answers, but let me be concrete, because vague superiority claims are exactly the disease I'm describing.

Workout With Me's S.H.A.R.P engine is deterministic and recovery-first. Every workout you log updates a colour-coded body map across 6 muscle categories and 19 subcategories, each labeled CRUSHED, IN RECOVERY, or RECOVERED with a percentage and time remaining. The engine will not schedule a muscle below 80 percent recovered. That's a rule you can watch operating, not a vibe. When real life happens outside the app, you mark the affected muscles crushed yourself and the plan reroutes around them. Same lifts, progressed over time, from your logged sets, with the reasoning sitting on a map you can open whenever you like. That's the design philosophy behind our whole recovery-based training hub, and I've tried to make the thinking as public as the product.

Is a language model useless in fitness? No, and I'll be honest about that too: we use AI where it's actually good, like reading a photo of a written program and turning it into structured, loggable exercises. Pattern recognition is a real strength. Deciding what your body should do on Thursday is not.

The takeaway

The word AI on an app tells you nothing anymore. Some of the smartest training software ever built carries the label, and so does a prompt in a trench coat.

So run the test. Memory, real inputs, inspectable reasoning, and a model of your recovery. If an app has all four, it's an engine worth arguing with. If it has none, you're paying a subscription for a chatbot's first guess. And if you want to see what an engine looks like when recovery is the first input instead of the missing one, Workout With Me is free to try, and the body map will show you its reasoning before you've paid a cent.