How to "Train" an AI Companion - What Your Feedback Really Changes

How it works

Many users rate every reply, believing they are teaching their companion. Most of the time they are not - at least not in the way they think. Here is what each kind of feedback actually does.

We may earn a commission from links on this page. It never changes a rating.

"Training" an AI companion suggests the software gradually learns your preferences, like a pet. The reality has more parts: a few levers change behaviour immediately, some change it slowly for everyone, and some mostly make you feel involved.

What actually shapes a reply

Every reply is generated from what the model can see at that moment - the mechanism in how an AI companion writes its reply:

  1. The character's description and settings.
  2. Stored memories and summaries retrieved for this conversation.
  3. The recent conversation itself.
  4. The underlying model, as the company last updated it.

Your feedback matters to the extent that it changes one of those four.

The levers, from strongest to weakest

The levers for shaping an AI companion ranked by how quickly and strongly they work: editing the character, editing replies, memory notes, regenerating, out-of-character notes and ratings

Change what the model reads, not how it feels about you.

LeverWhat it changesHow fastHow strong
Editing the character descriptionInstructions read before every replyImmediatelyStrong and lasting
Editing a replyThe conversation the model copies fromNext repliesStrong while it stays in view
Memory notesFacts retrieved into future chatsFrom next retrievalStrong for facts, weak for style
Out-of-character notesThe current sceneImmediatelyFades as the chat grows
RegeneratingOnly this one replyImmediatelyNone beyond this reply
Ratings (thumbs, stars)Usually data for future model updatesWeeks or months, if at allWeak for your character

Edit, do not argue

Arguing in character - "you're not acting like yourself" - rarely works, because the argument itself becomes part of the conversation the model copies. Editing a reply into what the character should have said gives the model a correct example to follow. If your app allows editing, it is the most useful habit you can build.

Put style in the description, facts in memory

  • Style - tone, sentence length, humour, how affectionate she is - belongs in the character description, because it must apply to every reply. Our character builder guide covers which fields matter.
  • Facts - your job, your dog's name, what happened last week - belong in memory, which is retrieved when relevant.

Mixing them up causes problems: style notes in memory are retrieved inconsistently, and long lists of facts in the description crowd out personality.

What ratings are really for

In most apps, ratings feed the company's data for improving models, often across all users. That can be worth doing - you are helping future versions - but it is not a direct line to your character. Some apps say ratings influence your companion specifically; even there, the effect is gradual and much weaker than an edited description.

Check the privacy policy if you care: rating a reply may mark that conversation for review or training, which is a privacy trade-off. See who reads your AI companion chats.

A weekly tune-up

  1. Note two things you liked and one thing that annoyed you this week.
  2. Turn the annoyance into a positive instruction in the description ("She answers in short, dry sentences" rather than "Don't ramble").
  3. Add any new important facts to memory.
  4. Start a fresh chat if the old one has drifted - long conversations accumulate habits, as personality drift explains.

The companion does not learn you the way a person does. But it reads you every time - through its description, its memory and your recent words. Change what it reads, and you change who it is.

Frequently asked questions

Do thumbs up and down change my AI companion?

Rarely right away, and often not your companion specifically. In most apps, ratings are collected to improve the model for everyone in future updates. A few apps use them to adjust your character, but the effect is small and slow compared with editing the description.

What is the fastest way to change how my AI companion talks?

Edit the character's description or personality field, and correct replies by editing them rather than arguing. The model copies what is in its instructions and in the recent conversation - those are the two levers that work immediately.

Does my AI companion learn from our conversations?

It stores facts and summaries in a memory system, so it can recall details. It does not usually retrain its underlying model on your chats in real time. Whether your chats are used to train future versions depends on the company's policy.