Can you trust AI-generated flashcards?
Not blindly, and the failure is worse than being wrong.
A model that generates a hundred cards from your chapter will get most of them right, and the ones it gets wrong will look exactly like the ones it got right. There is no tell. Wrong cards do not arrive with hedged language or bad grammar — they arrive in the same clean question-and-answer shape as everything else, and spaced repetition then does what it is built to do: it drills them until you cannot forget them.
That is the actual risk. It is not that the AI wastes your time. It is that it is a very efficient machine for memorising something false.
What models actually get wrong
Three failure modes account for most of it, and they are worth being able to recognise.
Invention. The model fills a gap with something plausible — a date, a mechanism, a figure that was never in your source. This is the one people mean by "hallucination", and it is most likely where your material was thin or ambiguous.
Compression damage. The source said something is usually true, or true under conditions, and the card says it is true. Nothing was invented; a qualifier was dropped. This is far more common than invention and much harder to spot, because the card is almost right.
Lost context. The card is correct and unusable, because the thing that made it meaningful — which patient, which experiment, which chapter's assumptions — was in the paragraph and not in the sentence.
The rule that actually protects you
Verify at the moment of creation, not at the moment of review.
When a card is generated you still have the source open. You know what the chapter said. Ten seconds of reading the card against the paragraph it came from is cheap, and it is the only point at which the check is cheap. Three weeks later, in a review session, you have no source, no memory of the context, and no way to tell a wrong card from a right one — you will simply feel that you knew this, and press good.
So: read what was generated before you keep it. Delete more than feels comfortable. A generator that produces sixty cards from a lecture has not given you sixty cards, it has given you a first draft with maybe thirty survivors, and the editing is the part that makes them yours.
Keep the source reachable. Whatever tool you use, write down where each card came from — Anki has a Source field, and any app worth using should give you something equivalent. A card you can check is recoverable. A card you cannot trace is one you will eventually delete out of suspicion, and that is how good work gets thrown away along with bad.
Prefer small cards. One fact per card is not a style preference here; it is error containment. A card carrying four facts is wrong if any one of them is wrong, and you will not know which.
What this means for choosing a tool
The useful question is not "does it use AI" — nearly everything does now. It is:
- Can you see the source a card came from, later, from the card?
- Can you edit or delete a card easily, after it has been made?
- Does it show you what it generated before it commits it to a schedule?
A tool that generates straight into a review queue with no editable step and no path back to the source has optimised the fast part and left you the expensive part.
How Quat handles it
Quat is an iPhone and iPad app I make, so read this knowing who wrote it.
It uses gemini-2.5-flash to turn what you capture into a note, and the note into cards. It does not fact-check anything, and neither does anything else in the pipeline. I would rather say that plainly than let the word "AI" imply a verification step that does not exist.
What it does do is make checking possible. Every card carries a stamp saying how it came in, the date, and the note it came from — MIC · 28 JUL · CARDIO LECTU — so a card that looks wrong three weeks later can be traced to the recording or the page it came from instead of being deleted on suspicion. Cards are editable and deletable at any point. The note the cards were made from stays in the app, which means the source is still there, not just a citation of it.
The generation allowance is deliberately not enormous: thirty card generations a month free, five hundred on a subscription. That is enough to work with and not enough to encourage generating a thousand cards you will never read.
The limit worth repeating: the stamp tells you where to look. It does not tell you the card is right. Nothing in any of these tools does. That part is still reading, and it is still yours.
Free · iPhone and iPad · no account