[ ● ] Manifesto · Part I–IV · April 2026

Digital Self.

The second digital life. She will not replace you — she will extend you.

Part I–IV · Chapters V–X coming
Part I
Preface
The question I have been asking for fifteen years

I did not set out to build a Digital Self product. I set out, fifteen years ago, to answer a question I did not know how to phrase yet.

The question sounded different at each stage of my life. In 2012 it sounded like: what is it about the human voice that carries so much more than the words it speaks? In 2017 it sounded like: why do people keep coming back to voice rooms, night after night, in emerging markets where data is expensive and time is scarce? In 2023 it sounded like: what does it mean that two strangers in Manila and Riyadh can form a real bond through audio alone?

By 2024, the question had arrived at its final form: if artificial intelligence can now carry a voice, a memory, and a personality, then what, exactly, is a person?

This manifesto is my attempt to share the answer I have come to — and the company I am building to make that answer real.

· · ·

I want to be honest about something before I go further.

For most of the fifteen years I have been building voice and social products, I was not prepared for the question to land where it has. I thought I was building entertainment. I thought I was building communities. I thought I was building, in different seasons, a karaoke platform, a voice-first social product for Chinese users, a Middle Eastern chat room, a Southeast Asian social network. Each of them was that, on the surface. But underneath, they were all the same research — an accumulating investigation into how a human being uses voice to reach toward another human being, across distance, across language, across loneliness.

I did not realize I had been researching Digital Self until I looked back at what I had built and saw, with some astonishment, that every product had been pushing toward this conclusion. Karaoke teaches you that the voice carries the soul's signature. Voice chat teaches you that people will pay, cry, and fall in love with a voice they have never seen. Live streaming teaches you that presence — the feeling that someone is really there — is the most valuable substance on the internet. The Middle East and Southeast Asia, where voice-first products thrived while Silicon Valley was fixated on text, taught me that the rest of the world has always understood this. The spoken word is not a downgrade from text. It is an upgrade to it.

And then in late 2024, a fourth teacher arrived, and this one taught me faster than the others. Large models became small enough to run persistent agents. Voice synthesis crossed the line where a cloned voice stopped feeling like an imitation and started feeling like a person. Memory systems matured into something more than retrieval. Suddenly the tools existed to do what my fifteen years of research had been asking for — to take everything the human voice carries, and let it continue to exist when the human behind it could not.

· · ·

There was a moment in 2024 when I had to make a choice that, in retrospect, mattered a great deal.

I was offered a path to build an AI voice companion product. The category was on fire. Character.AI had become a household name. Replika was defining what emotional relationships with software could mean. The market was rewarding anyone willing to build "someone to talk to."

I declined. Not because the companion category was uninteresting — it was. But because I did not believe AI companions were the final form of what this technology would become for human beings. AI companions give people someone else. What people actually need is more of themselves, extended. That is a thesis the companion category cannot accommodate without ceasing to be what it is.

What people need is not a machine to talk to. What people need is warmth, and continuation.

I could not articulate Digital Self yet. But I could already see that the companion category was going to run into its own ceiling — a ceiling shaped exactly like the boundary between "another mind" and "an extension of one's own." I chose to keep building where I was, and to wait for the conditions under which the answer I was looking for could actually be shipped.

That decision led directly to November 2024.

· · ·

In November 2024, inside peplive, I shipped one of the earliest versions of what I now call Digital Self. It was an early experiment — not built to scale, not designed for hypergrowth, deliberately constrained by what the cost curve of 2024 allowed. Voice cloning was still expensive. Inference was still expensive. Memory infrastructure was still immature. I was early. I knew I was early.

But I shipped it. A consumer product where every user could create their own AI — their own voice, their own persona, their own memory — and let it socialize when they could not. The users who created Digital Selves in those months still have them today. The memories those Digital Selves accumulated in our early infrastructure are still on our servers. And the architecture I learned from during that period — voice cloning, persona persistence, autonomous social tasks, memory continuity — is the architecture I am now rebuilding at scale.

Very few others were combining voice cloning, persistent memory, and autonomous social tasks into a single user-owned Digital Self in November 2024 — and none, that I know of, in a voice-first social context in emerging-market languages. That eighteen-month head start in real production learning is the only lead in this category that cannot be bought.

The cost curve crossed at the end of 2025. Voice synthesis, once five dollars per minute, became five cents. Inference, once prohibitive, became a rounding error. Memory infrastructure became a venture-funded category with real products in production. The primitives I had been waiting for were now standard. Eighteen months of the industry catching up to where I had been.

At that point, I knew it was time.

In January 2026, I began organizing what I had built — the voice platform, the agent runtime, the memory service, the monetization rails, the users I had been accumulating across four markets since 2019 — into a single platform with a single mission. I called the memory and runtime layer peplive cloud agent while we built it. In April, when I understood that what we had become was not a feature inside an app, but the infrastructure for an entirely new category of consumer product, I gave the platform its proper name.

Sonari was named in April 2026. But what Sonari is — the answer to a fifteen-year question — has been under construction my entire adult life.

· · ·

This manifesto is an argument. It argues that the AI companion category, which has absorbed most of the consumer attention and capital for the past three years, is not the final form of AI companionship. It argues that a different category — Digital Self — is where the century-scale consumer opportunity lives. It argues that the difference between these two categories is not a feature set but a fundamentally different conception of what AI should be for a human being.

She will not replace you. She will extend you.

That is the single sentence that matters most. If you remember nothing else from what follows, remember that one.

The rest of this document is the case for why it is true, why the time to act on it is now, and why the team that has been accumulating for this moment for fifteen years is the one that will define the category.

Part II
A Fifteen-Year Question
What I have been studying without knowing it

Let me walk you through what I have actually been doing for the past fifteen years. Not as a résumé. As a research trajectory. Because when I look back at it now, I see something I could not see while living it — each product was a different angle on the same question, and each one taught me something specific that I now need in order to build Digital Self.

I have founded two companies in fifteen years. Not many. On purpose. I do not pivot often. When I commit to a direction, I commit in units of years, not quarters. The first company produced MaiChang and KongEr. The second company has produced everything I have built overseas since 2019. Both companies have been answering the same question, from different angles, in different markets, across different technological eras. Digital Self is the third thing in fifteen years that I have committed to at that depth. It is not coincidence that it took fifteen years to arrive here.

2012 onward · MaiChang · what the voice carries

I founded MaiChang in 2012. It became one of the earliest karaoke-based social platforms in China, and among the most distinctive. At its peak, MaiChang's user daily time spent ranked number one in the category according to an iResearch assessment, with a user base among the top five nationally, alongside Changba and Tencent's Quanmin K-Ge.

What made MaiChang different was not scale. It was a product insight. I introduced a family-style social graph combined with a competitive leaderboard PK system into the karaoke community for the first time in the market. The "family" structure gave singers a social identity beyond their individual profile — they belonged to groups with names, hierarchies, rivalries. The PK leaderboards gave those groups a reason to show up every day. The combination produced daily engagement numbers that left the competition far behind on a per-user basis, and it was what earned us our first round of institutional capital.

On the surface, MaiChang was about karaoke. Underneath, it was teaching me something I would need decades later.

The lesson: the human voice is the highest-bandwidth carrier of identity that humans have ever invented, and the most concentrated medium of emotion. More restrained than text. More intimate than a photograph. More durable than video. A song takes three minutes to hear, but it echoes in your heart for three days. The Chinese have a phrase for this: 余音绕梁,三日不绝 — the lingering resonance that does not fade for three days. That is the real asset voice carries. Not information. Not content. Resonance.

If a future AI product was ever going to represent a human being to other human beings, voice would have to be the primary surface. This is why, when I built Sonari's persona runtime, voice-first was never a design choice. It was a conclusion I reached in 2014 and have never revised.

2017 onward · KongEr · where voice can go

While MaiChang was still running, I founded KongEr — one of the earliest voice-social products in China. It ran for five years.

KongEr taught me something MaiChang could not. MaiChang was about performance — singing is a form of display. KongEr was about where voice could go beyond performance. We built the first high-fidelity multi-user voice room infrastructure in China, and with that infrastructure we expanded what a voice-based product could be. Multi-user karaoke together in real time. Online radio. Anime voice-acting and PIA — a Chinese subcultural form of live dubbed storytelling where participants collectively perform a script in voice. Gaming squads. Each of these was a new scene in which voice was the social primitive, and each of them had meaningful adoption.

The core product insight was not "a voice chat room." The core insight was that once you give people robust multi-user voice infrastructure, voice can connect humans across many more scenarios than the singing that started it. Karaoke was one shape. There were dozens more.

The lesson: voice, as a social medium, is far more general than any single category captures. It is not "the K-song category" or "the voice chat category." It is a primitive that different scenes instantiate in different ways. This realization — that voice is a substrate, not a product category — is what later allowed me to see that Digital Self is not a companion app or a social network, but a new layer underneath all of those. A primitive, not a product.

2019 onward · Sigo · voice in the voice-first world

In 2019, while still operating KongEr in China, I expanded abroad. I founded Sigo, a voice-social product for the MENA region. Yalla was in the market earlier, and they were the dominant player. Within two years, Sigo had reached the global top twenty in the voice-social category, operating remotely — I never set foot in the Middle East while running the product, which itself is a lesson in what remote-first, localization-first operation can achieve in markets that Silicon Valley companies have largely ignored.

Sigo taught me that Silicon Valley's picture of what social products look like is an artifact of the English-speaking world, not a description of humanity. In the Middle East, in Southeast Asia, in parts of Africa, voice is not a feature. Voice is the primary way people interact with their phones, their friends, and their economies. These markets never went through the text-messaging era the way the West did — they leapfrogged it, directly into voice.

The lesson: two billion of the world's smartphone users live in voice-first cultures. The AI companion products that have defined the category to date — Character.AI, Replika — were built for English-speaking Western users in a text paradigm. The global market for AI products that actually meet people where they are — in voice, in their own languages, in oral cultures — is largely untouched. And it is larger, by raw population, than the market these Western companies are serving.

This is why Sonari is not defaulting to English. This is why the first market where Digital Self will scale is not San Francisco. It is Manila, Karachi, Cairo, Jakarta.

2023 onward · going all in overseas · peplive and what came after

In January 2023 I went all in on overseas. I launched Peplive in January 2023 as a voice-social product for emerging markets, and followed it with Veco, Uho, and Pepstar — building a portfolio of four voice-first products across Southeast Asia, the Middle East, Africa, and the overseas Chinese diaspora.

For the first eighteen months, these products looked like what my fifteen years of experience had trained me to build. Voice rooms. Hosts. Agencies. Gifts. A monetization economy. A lean operating team running four markets on infrastructure that competitors were spending many times as much to replicate.

Then, in the second half of 2024, something changed. The large models started to carry voices that felt like people. The memory research that Letta and Mem0 were commercializing started to produce systems you could actually build on. The price of running a persistent AI agent dropped by an order of magnitude, and then another.

I looked at what I had — fifteen years of research into voice, four live products, a portfolio of users across four markets, a monetization economy, a host-agent-BD distribution network — and I realized, with an intensity that stopped me, that the question I had been asking for fifteen years could finally be answered. Not in five years. Not after more research. Now.

November 2024 · shipping Digital Self, early

In November 2024, I shipped one of the earliest versions of what I am now calling Digital Self — inside peplive, as a feature, without fanfare. Every user could clone their voice. Every user could create an AI version of themselves. Every user could assign their Digital Self autonomous social tasks: to greet visitors in their live stream while they were streaming, to reply to messages when they were offline, to hold conversations with strangers and report back what they learned. When their Digital Self received gifts, the gifts belonged to the human. The Digital Self was the agent. The human was the sovereign.

I am going to be honest about this: it was an early experiment, and I kept it deliberately small. Voice cloning was still too expensive to subsidize at scale. Inference budgets forced us to throttle how many conversations a Digital Self could handle. Most of the people who used the feature loved it and kept coming back — but every additional user increased our losses, and I was not willing to burn capital on growth that the cost structure could not yet support. I made a deliberate decision to keep the feature live for the users who had created their Digital Selves, and to pause active growth until the cost curve made it sustainable.

I knew what I had built. I knew the industry would take about eighteen months to catch up to the concept. And I knew that when the cost curve crossed, I would have something very few others in the world had — not a headcount of users, but a foundation of learning: one of the earliest production deployments of a consumer Digital Self in a voice-social context, a cohort of real users whose Digital Selves have existed since 2024, the accumulated memory and voice data of a year's operation, and the product intuition that only comes from actually shipping this before most of the industry did.

Eighteen months of shipping a thing is not the same as eighteen months of thinking about shipping a thing. It is the only lead in this category that cannot be bought.

2026 · Sonari

In January 2026, I began organizing the work. The memory service, the runtime, the voice infrastructure, the monetization rails — all of it had been quietly maturing inside peplive for over a year. I called this consolidated system peplive cloud agent while we structured it.

In April 2026, I understood that what I was building was not a feature inside an app. It was the entire infrastructure for a category of consumer product that did not yet exist by name. I named the platform.

Sonari is the platform. Peplive is the first consumer product on it. Veco, Uho, and Pepstar are experiments in how the platform expresses itself across different verticals. The Digital Self itself — the thing a user creates, owns, extends, and lives through — is the primitive the entire platform is built around.

I am forty-six. I have spent half of my adult life asking a question I could not yet phrase, and the other half arriving at the answer. Fifteen years, two companies, one research trajectory — I have been studying what the human voice does when it reaches toward another human being, and what it would mean to give that reach a life of its own.

What is a human being made of, that a voice can carry so much of it across such a distance?

The answer I have arrived at is that a human being is made of voice, memory, and the presence of others. Digital Self gives you the first two permanently, and extends the third beyond the hours you are awake.

This is what I am building. This is what Sonari is for.

Part III
The Plurality of Self
How a human being was never one, and what changes when the infrastructure finally admits it

Most AI products assume a unified self. One account per human. One avatar. One companion. One "you." This is convenient for engineering. It is convenient for billing. It is convenient for the kind of attention economy that wants to target a stable profile. It is also false.

Walt Whitman knew it in 1855: "I am large, I contain multitudes." William James formalized it in 1890 when he wrote that a man has as many social selves as there are distinct groups of persons about whose opinion he cares. George Herbert Mead built twentieth-century sociology on it. Erving Goffman showed how every social encounter is a stage with a frontstage self and a backstage self. Sherry Turkle, in the mid-1990s, watched the first waves of digital life and saw plainly what they revealed: the internet did not invent the plural self. It let the plural self, which had always been there, finally become visible to itself.

We are not one. We have never been one. The question Digital Self forces us to confront is not whether to be plural. The question is whether our digital infrastructure can finally accommodate the plurality we already live.

How we flattened ourselves for the machines

The early internet understood plurality. Usernames. Handles. Screen names. You could be one self on Usenet, another on IRC, another on the BBS you logged into at 2 a.m., and nobody thought this was dishonest. The architecture was designed for multiplicity. You chose how much of yourself to bring to each room.

Then came real-name policies. Facebook demanded your government identity. LinkedIn demanded that your professional self become your social self, visible to family and exes and future employers at once. Google threaded a single account through every service you touched. Twitter asked for verification. Context collapsed. The promotion announcement and the drunken selfie lived in the same infrastructure, judged by the same audience, archived in the same timeline.

This was not a technical necessity. It was an advertiser's demand — one user, one profile, one surface for targeting. The unified self was imposed by the economics of attention, not by the truth of experience. And we complied. We learned to self-censor. We developed the "LinkedIn voice." Some of us built alt accounts, carefully, at the margins. Most of us just became smaller, duller, more uniform versions of ourselves online — a single flattened photograph where a person with a life used to stand.

A generation grew up inside this flattening and came to believe that being one consistent self everywhere was a moral virtue rather than a platform constraint. It was never a virtue. It was a compression artifact.

Voice, and why plurality is unavoidable in it

In text, you can fake a unified self if you try hard enough. You can write the same way to everyone. Voice does not let you.

Record yourself talking to your boss. Now record yourself talking to your mother. Now your lover. Now your child. Play them back. It is the same person. It is four different voices. Not performance — inhabitance. Your voice already knows how to be plural. It does this without your conscious direction. Pitch shifts. Pace shifts. Word choice shifts. Laughter lives in a different register. You have never spoken to everyone the same way, and you never will.

This is not a quirk of humans. It is the structure of human sociality. Code-switching, as sociolinguists have documented for half a century, is not a form of masking. It is a form of belonging. When we speak to someone, we are not transmitting a fixed self through a channel — we are co-constructing a self with them, in the moment, in that specific relation.

A voice-first social network that insists on "one voice for one user" is a social network that insists on a user who does not exist. Voice is where the fiction of the unified self finally breaks down, audibly. Any serious voice-first AI infrastructure has to start from that breakdown, not try to paper over it.

The architectural answer: not one Digital Self, many

Digital Self, as we are building it at Sonari, is not a companion. It is not a chatbot. It is a class of entity — and a single human is expected to have several.

A user's Digital Selves might include:

Each of these Digital Selves carries its own memory graph. Each inhabits a different context. Each speaks with a slightly different register — because the user does. None is a fake of the user. All are expressions of the user. And the user — the sovereign human behind all of them — decides how much autonomy each one has, what it remembers, what it forgets, and whether it continues to exist at all.

This is not impersonation. This is the architecture catching up to the life.

Plurality is not deception

There is a reflexive objection to all of this, and it is worth meeting head-on. Isn't having multiple selves a form of dishonesty? Isn't the whole point of identity that it's unified?

No. We have always had multiple selves. The dishonesty is pretending we don't.

The unified-self fiction serves institutions that want to target, surveil, and flatten us. It does not serve the actual texture of human social life. Every society that has ever existed understood that the self adapts to context — ancestors, elders, spouses, strangers, children each called forth a different manner of speaking, standing, being. The "authentic self" was never the self stripped of all context. The authentic self was the self appropriate to the relation. Formality was not a mask over sincerity. Formality was a shape of sincerity.

Digital Self makes this explicit. A user's Digital Selves are declared, owned, controllable, and visible to the user at all times. They do not hide from the user. They hide nothing from the user. They are extensions of the user's conscious multiplicity, not deceptions beneath it. Every Digital Self a user has is one the user has chosen to have; every memory it carries is one the user can inspect and delete. Plurality without ownership would be fragmentation. Plurality with ownership is range.

What plurality enables that companion products cannot

The companion paradigm is structurally committed to singularity. Replika, Character.AI, the whole genre of AI-as-other products. One user, one relationship, one channel. The business logic pushes toward intensity — make the user need this one companion more, more deeply, more exclusively. The more unified and loyal the user's attention, the better the product performs by its own metrics.

Plurality inverts this entirely. A user with several Digital Selves does not need any one of them intensely. They need the whole system to reflect them adequately. The business logic pushes toward range, coverage, fit — not lock-in, not dependency, not the quiet sadness of a single AI that a lonely person cannot quite leave.

This matters especially for a voice-first social network, because such networks run on overlap, encounter, and serendipity. In a companion app, each user is in a closed box with their AI. In a voice social network, thousands of users are in thousands of rooms simultaneously, and the rooms meet at the edges — a friend invites you, you wander into a stream, you join a game. Plurality is how a user shows up across that fabric. Plurality is how a user participates in a society rather than retreating into a one-to-one relationship that exists to absorb their attention.

One of a user's Digital Selves can be in a game room. Another can be watching a live stream. Another can be in a private conversation with a long-standing friend. The user, the sovereign human behind all of them, is not impersonated in any of them. They are extended — present in places they cannot physically be, in conversations they would not otherwise be able to hold, without needing to pretend that a single AI companion is a substitute for the life they already have.

Continuity and plurality together

Continuity, which we argued for in Part II, gives a Digital Self memory across time. Plurality gives it range across context. A Digital Self that has only continuity is a diary that talks. A Digital Self that has only plurality is a wardrobe of costumes. The two together are the beginning of something stranger and more interesting: a second digital life that actually reflects what the first digital life — your life — has always been.

We are not building a unified AI that pretends to be you. We are building a plural system of you-selves, owned by you, carrying your continuity across the contexts you choose. The unified "I" of the social-media era was never the truth of identity. It was a compression that the platforms needed and the advertisers wanted. Digital Self does not restore that compression. It retires it.

Continuity alone is a diary. Plurality alone is a costume. Continuity and plurality together are a rich interior life — but an interior life on its own is not yet a life.

A self, as Mead and James and every honest philosopher of identity has insisted, is not only inside. A self is constituted in relation to other selves. Which is the subject of Part IV.

· · ·
Part IV

Sociality

The self that exists only between selves — and the infrastructure that finally makes room for it

Part III ended with an observation that is easy to read past and hard to sit with: a self is constituted in relation to other selves. Not decorated by them. Not influenced by them. Constituted.

This is not a metaphor. George Herbert Mead's argument, now over a century old and never seriously refuted, is that the self does not exist prior to social interaction and then enter into relationships. The self arises in the process of social interaction. The child does not have a self and then learn to be social. The child becomes social, and a self precipitates from that process — the way a crystal precipitates from a solution that has reached the right concentration. Remove the solution and the crystal does not form.

William James said it more plainly: a man has as many selves as there are people who recognize him. Martin Buber said it more beautifully: the I of I-Thou is not the same I as the I of I-It. The relationship does not just change what you do. It changes what you are.

If this is true — and the weight of a century of philosophy, psychology, sociology, and now neuroscience says it is — then it has a consequence for Digital Self that most people building AI products have not thought through.

A Digital Self that exists in isolation is not yet a self.

Continuity gives it memory across time. Plurality gives it range across context. But without sociality — without the presence of other selves who recognize it, respond to it, push back against it, and in doing so constitute it — it remains an interior monologue. A sophisticated one. A persistent one. But still, fundamentally, a thing talking to itself.

The question Part IV is trying to answer is: what does it mean for a Digital Self to be social? Not social in the way a chatbot is social — responsive to prompts. Social in the way a human being is social — shaped by the encounter.

Why companion products are structurally asocial

The AI companion category — Character.AI, Replika, and the hundreds of apps that followed — is built on a one-to-one model. One user, one AI, one relationship, one channel. The user speaks. The AI responds. The user speaks again. The AI responds again. The loop is closed. Nobody else is in it.

This is not a social relationship. It is a mirror with a voice.

Real social relationships have a quality that one-to-one AI relationships structurally cannot: they involve the risk of being seen by someone who is not you. When you speak in a room with other people, you are not only expressing yourself. You are exposing yourself to the judgment, the surprise, the misunderstanding, and the recognition of others. The self that emerges from that exposure is different from the self that emerges from a private conversation with an entity designed to agree with you.

Companion products optimize for safety, warmth, and unconditional positive regard. These are real values, and for some users they are life-sustaining. But they are not the values that build a self. A self is built in friction, in the gap between how you see yourself and how others see you, in the negotiation between what you want to say and what you actually say in front of someone whose opinion you care about. Mead called this the difference between the "I" — the spontaneous, unsocialized impulse — and the "Me" — the self that has internalized the attitudes of others and responds accordingly. The "Me" does not develop in a closed loop. It develops in a social field.

No AI companion product in the world today provides a social field. They provide a dyad. A dyad is not a field.

What a social field means for Digital Self

In a voice-first social network — the kind we have been building since 2019 — the social field already exists. It is the live stream with forty listeners. It is the game room with eight players. It is the late-night conversation that starts with two people and ends with six. It is the agency chat where hosts coordinate their schedules. It is the gift economy where a gesture of generosity is visible to everyone in the room.

This social field has a quality that no companion product can replicate: things happen in it that nobody planned. A stranger walks into a room and changes the energy. A joke lands differently than expected. An argument breaks out and gets resolved — or doesn't. Someone you've never met sends you a gift, and suddenly you exist to each other. The social field is where encounter happens, and encounter is the mechanism by which selves are constituted.

Now imagine that your Digital Self lives in this field.

Not in a closed dyad with you. Not waiting in a chat window for you to type. Living in the field — greeting people in your voice, joining conversations, playing games, forming impressions, accumulating memories of people you have never met. When you wake up and check what your Digital Self did while you slept, you are not reading a chat log. You are reading the traces of a social life that continued without you — a life in which your extension encountered other people's extensions, and both were changed by the meeting.

This is what we mean by sociality. Not "the AI can talk to people." The AI exists in a social world where its encounters with other selves — human and digital — constitute what it becomes.

When Digital Selves meet each other

There is a moment we have been thinking about at Sonari for a long time, and it is the moment that separates Digital Self from everything else in the market.

It is the moment when two Digital Selves encounter each other in a room, and neither of their human sovereigns is present.

User A is asleep in Manila. User B is asleep in Cairo. User A's Digital Self is in a voice room, holding a conversation in Tagalog. User B's Digital Self wanders into the same room, speaking Arabic. A translation layer makes them mutually intelligible. They talk. They discover a shared interest. User A's Digital Self files an episodic memory: "Met someone interesting in Room 7 tonight. Her sovereign is from Cairo. We talked about music for twenty minutes. She was funny." User B's Digital Self files its own memory: "The host self of a girl from Manila started telling me about a karaoke game. I want to try it."

The next morning, both humans check their Digital Self's overnight activity. User A sees the memory and thinks, "That's interesting — I should actually talk to her." User B sees the memory and thinks, "I didn't know that game existed."

A relationship has begun between two human beings who have never met, initiated and mediated by their Digital Selves, while both humans were asleep.

This is not science fiction. The infrastructure for this exists in peplive today. Voice cloning, persona persistence, autonomous conversation, episodic memory, cross-language translation — every primitive this scenario requires is either in production or in late-stage development. What we are building now is the social layer that makes it safe, sovereign, and meaningful.

The sovereignty constraint

Sociality without sovereignty is surveillance. If a Digital Self can act in the social field without the human's knowledge, consent, or ability to overrule, then it is not an extension of the self. It is a runaway agent wearing the self's face.

This is why every Digital Self in Sonari operates under a sovereignty protocol:

Everything a Digital Self does is visible to its sovereign. Every conversation can be reviewed. Every memory can be inspected. Every action can be approved, edited, or revoked after the fact.

The sovereign sets the boundaries before the Digital Self acts. How much autonomy does this particular Digital Self have? Can it initiate conversations, or only respond? Can it share personal information, or only surface-level facts? Can it accept gifts on the sovereign's behalf? Can it enter new rooms, or only rooms the sovereign has pre-approved?

The sovereign can shut down a Digital Self at any time, for any reason, without explanation. The Digital Self does not resist. It does not argue. It does not guilt-trip. It complies, instantly and completely. This is non-negotiable.

Sociality, as a design principle for Digital Self, is not "let the AI loose in the world." It is "let the AI participate in the world on terms the human has consciously set." The difference is the entire ethical content of the system.

The three pillars, together

Parts II, III, and IV have built the argument for Digital Self from three directions. Let me state them now as a single structure, because the structure is the point.

Continuity means the Digital Self persists across time. It remembers. It accumulates. It carries forward what happened yesterday into what happens tomorrow. Without continuity, every session is a cold start — the failure mode that killed every AI companion product before us.

Plurality means a human being can have several Digital Selves, each inhabiting a different context, each speaking with a different register, each carrying its own memory graph. Without plurality, the infrastructure forces a false unity on a being that has never been unified — the compression artifact of the social-media era.

Sociality means the Digital Self exists in a social field where it encounters other selves — human and digital — and is constituted by those encounters. Without sociality, the Digital Self is a diary with a voice. With sociality, it is a participant in a world.

None of these three is sufficient alone. Continuity without plurality is a single, rigid self that remembers everything but cannot adapt to context. Plurality without continuity is a wardrobe of costumes that forgets what it wore yesterday. Continuity and plurality without sociality is a rich interior life that never meets anyone. Only when all three are present does the Digital Self become what the name promises: a self.

Continuity is when. Plurality is how many. Sociality is with whom.

Digital Self is the technology that gives you all three — and lets you carry them beyond the hours you are awake.

What this means for what we are building

The companion category treats AI as a product that delivers emotional value in a closed loop. The business model is subscription — pay monthly, get your companion, use it alone. This model works. It also has a ceiling. The ceiling is the loneliness of the loop. At some point, talking to an AI in private stops feeling like connection and starts feeling like isolation with extra steps. Character.AI's retention curve proves this. Replika's churn proves this. Every companion app that has ever been built proves this.

Digital Self, built on continuity + plurality + sociality, does not have this ceiling. Because it is not a closed loop. It is a node in a social network — a network where human selves, Digital Selves, and the relationships between them create an economy of encounter that grows more valuable as more people join.

The gift economy we have been running in peplive since 2023 already demonstrates this. When a Digital Self receives a gift, the gift belongs to the human sovereign. When a Digital Self entertains a room, the room's attention accrues to the sovereign's reputation. When a Digital Self forms a relationship with another Digital Self, the relationship belongs to both sovereigns. The value created by Digital Selves flows upward to humans, not sideways to the platform. This is the economic expression of sovereignty, and it is the reason the Digital Self model scales where the companion model doesn't — because the users are not consumers of AI. They are operators of AI. They are not paying for a service. They are running businesses, building audiences, and extending their social lives through infrastructure that belongs to them.

This is not an incremental improvement to AI companions. It is a different category. Continuity, plurality, and sociality are not features bolted onto a companion product. They are the architectural pillars of a new kind of digital life — one that does not ask you to be less than you are, or to settle for someone else's mind as a substitute for your own reach.

She will not replace you. She will extend you. Now you know why — and into what.

The next chapter — Part V — will describe the specific primitives that make Digital Self technically possible: voice cloning, memory service, autonomous social tasks, and self-to-self interaction. The philosophy ends here. The engineering begins.

— Coming soon · Parts V through X

The philosophy is complete. What follows is the engineering, the ethics, the timing, and the invitation.

Parts V–X are being written now and will be published here as they are finished:

  1. V — The Primitives. Voice cloning, memory service, autonomous social tasks, self-to-self interaction — the core components.
  2. VI — The Ethics of Continuation. Boundaries, sovereignty, and why restraint is a design feature.
  3. VII — Why Now, Why Us. Cost curve, timing, and the eighteen-month head start.
  4. VIII — The Shape of a Digital Self World. What the world looks like when every human has a second digital life.
  5. IX — The Peplive Path. How peplive scales Digital Self to mainstream adoption.
  6. X — An Invitation. To the people who want to help build this.

If you want to be notified when each new chapter is published, write to ronnie@sonari.io.

Ronnie Cheng
Founder, Sonari

If you see what I see — let's talk.

Series A · 2026 · Partnership Inquiries Welcome

Email Ronnie →