18+

Three separate systems write every reply

How do AI girlfriends work? Most apps hide the pipeline — this page names every layer instead of guessing.

Six characters, six separate personality profiles

Each one runs on the same underlying models but a different personality profile and memory store — that's why they don't feel interchangeable after the first few messages.

Under the hood

What's actually running when you chat

Six things happening behind a normal-looking chat window: the model, the memory, the voice, the render, the latency, and the setting that controls all of it.

AI companion from ByteGF

A language model writes the reply

Every message goes through a language model tuned on the character's written personality, not a static script. That's why two chats with the same character diverge depending on what you say.

70+video modes

Tap one, she does it.

A separate layer stores what you said

The model itself doesn't remember earlier sessions — a memory layer outside it logs names, dates and running jokes, then feeds the relevant lines back in before she replies.

Where the seconds actually go

Tap one, she does it.

Flirt now

Synthesized, not a recording

Her voice comes from a voice model trained on a single actor's read of the character, run on whatever reply the language model just wrote. Nothing is pre-recorded.

Photos generated on request

Ask for a photo and an image model renders one inside the thread, using the character's established look and the scene you're both in — not a fixed set you're picking from.

The Lust Level is a setting, not a mood

Five steps control what the model is allowed to write and generate, from company to explicit. It's a switch you flip, not something she arrives at on her own.

The pipeline

What happens between your message and her reply

Four steps, four separate systems. None of them is instant, and none of them is quietly doing another one's job.

  1. 1

    You send a message

    Text or voice, it enters the same pipeline. Voice gets transcribed first; typed text skips that step entirely.

  2. 2

    The model generates a reply

    The language model writes the reply using the character's personality profile, the current Lust Level, and whatever the memory layer just fed it.

  3. 3

    Voice or photo gets rendered

    On a call, the reply gets synthesized into audio in her voice. If you asked for a photo, an image model renders one in the same beat.

  4. 4

    Memory gets written back

    Whatever mattered in that exchange — a name, a plan, a preference — gets written to the memory layer so the next session starts from where this one left off.

Against the typical app

One model doing everything, or several doing one thing each

  • Memory

    A dedicated layer, written to every session

  • Voice

    A voice model trained on the character

  • Latency

    Separate models for text and voice, running in parallel

  • Photos

    Rendered in the thread, on request

  • Explicit content

    A five-step setting you control directly

  • Transparency

    Every layer is something you can point at

70+ video modes

Tell her what to do. Watch it

70+ modes. Pick one, she does it in about a minute. Blurred here, not in your chat 😏

Flirt now· No card needed to start 😘

How AI girlfriends actually work

The mechanics