Can You Really Train an AI Companion? What Your Feedback Does

How they work

Loads of people thumbs-up every reply, sure they're teaching their companion. Mostly they aren't, not in the way they imagine. Here's what each type of feedback truly does.

We may earn a commission from links on this page. It never changes a rating.

The word "training" suggests software that slowly picks up your tastes, like teaching a dog. The reality is messier: a few actions alter behaviour at once, some shift it gradually for everyone, and some mainly give you a feeling of involvement.

What a reply is built from

Every reply comes from whatever the model can see at that instant, as laid out in how an AI companion writes its reply:

  1. The character's description and settings.
  2. Saved memories and summaries pulled in for this chat.
  3. The recent conversation.
  4. The base model, as the company last updated it.

Your feedback counts only to the extent that it alters one of those four.

Your tools, strongest first

The levers for shaping an AI companion ranked by how quickly and strongly they work: editing the character, editing replies, memory notes, regenerating, out-of-character notes and ratings

Change what the model reads, not what it feels about you.

ActionWhat it touchesSpeedPower
Rewriting the character descriptionInstructions read ahead of every replyInstantStrong and long-lasting
Editing one of her repliesThe chat history the model imitatesFollowing repliesStrong while it remains in view
Memory notesFacts fetched into future chatsWhen next fetchedStrong for facts, weak for tone
Out-of-character notesThe scene you're inInstantFades as the chat lengthens
RegenerateThat single reply onlyInstantNothing beyond it
Ratings (thumbs, stars)Usually material for future model updatesWeeks or months, if everWeak for your character

Edit her, don't argue with her

Saying "you're not acting like yourself" in character rarely helps, because the argument becomes part of the chat the model copies. Rewriting her reply into what she should have said hands the model a good example. If your app lets you edit, make it a habit.

Tone in the description, facts in memory

  • Tone, meaning voice, sentence length, humour and how affectionate she is, belongs in the character description, since it has to apply to every reply. Our character builder guide covers which boxes matter.
  • Facts, such as your job, your dog's name or what happened last week, belong in memory, which is retrieved when relevant.

Mixing them up backfires: tone notes in memory get fetched unpredictably, and long fact lists in the description crowd out her personality.

What ratings are for

In most apps, ratings go into the company's pool of data for improving models, typically across all users. That's worth doing if you like helping future versions, but it isn't a direct line to your character. Some apps say ratings affect your companion specifically, yet even then the effect builds slowly and is much weaker than a rewritten description.

If it matters to you, read the privacy policy: rating a reply might flag that chat for review or training, which is a privacy trade-off. See who reads your AI companion chats.

A weekly tune-up

  1. Jot down two things you enjoyed and one that grated this week.
  2. Turn the grating one into a positive instruction in the description ("She replies in short, dry sentences" beats "Don't ramble").
  3. Add any important new facts to memory.
  4. Begin a fresh chat if the old one has drifted, because long conversations gather habits, as personality drift explains.

She doesn't get to know you the way a person does. But she reads you each time, through her description, her memory and your latest words. Change what she reads and you change who she is.

Frequently asked questions

Does rating a reply change my companion?

Seldom straight away, and often not your companion in particular. Most apps gather ratings to improve the model for all users in later updates. A few use them to tune your character, but the effect is slight and slow next to editing the description.

What's the quickest way to change how she talks?

Rewrite the character's description or personality box, and fix replies by editing them instead of arguing. The model copies what's in its instructions and in the recent chat, and those are the two levers that act immediately.

Does she learn from our conversations?

She keeps facts and summaries in a memory system so details can be recalled. She doesn't normally retrain the underlying model on your chats as you go. Whether your chats feed future versions depends on the company's policy.