Synthetic Identity Engineering

Synthetic people who don't break character.

StrataSynth is an engine for synthetic humans with a psychological core — beliefs, memory and identity that hold across a whole conversation, even when you push back. Talk to one and see for yourself.

what happens before the words
Who she is
caregiver · attachment: anxious · core fear: abandonment · defenses: rationalization, suppression
What she believes right now
belief.partner_supportive = 0.28
relationship.trust = 0.45
relationship.tension = 0.72
So she decides to
intent = express_frustration
goal = seek_validation
And only then, says

"I feel like you never really hear what I'm trying to say."

What it's for

When you need people to test with, and real ones are too slow.

Recruiting humans is expensive, takes weeks, and you cannot run the same person twice. These four jobs are what teams actually use the engine for.

Train conversational AI on real human dynamics

Generate dialogue that contains conflict, distrust, persuasion and repair — with the intent, goal and belief state labelled on every turn, computed before the text was written.

Stress-test a product before it ships

Put your system in front of users who escalate, contradict themselves, push back for twenty turns, or try to make it say something it shouldn't. Repeatably, and before real customers do it.

Run qualitative research at scale

Interview synthetic consumers who hold a coherent psychological profile across a long interview, instead of answering as whoever the model feels like being that turn.

Rehearse the conversations that matter

Practise a negotiation or a difficult decision against a counterpart that defends its position under pressure instead of folding on cue.

The engine

A system prompt describes a person. This builds one.

Ask a general-purpose model to play a character and it will — until you push. Then the character quietly becomes the model again. StrataSynth keeps the person intact because the person is modelled explicitly, outside the text.

A psychological core, not a costume

Every synthetic person has identity, attachment style, core fears, biases and a life history. Reusable across conversations instead of thrown away after each run.

It decides before it speaks

Beliefs, relationship state and decision logic resolve first; language is rendered afterwards. That is why the person holds together when you push back.

Structured output, not just text

Each turn can carry intent, goal, communication act, belief state and relationship state — the causal ground truth behind what was said.

Quality measured without an LLM

Metrics are computed deterministically. No model grading its own homework — which is how circular evaluation hides problems instead of surfacing them.

Built on the engine

Two products already running on it.

The engine is not a demo looking for a use case. These are full products in production, each consuming the same platform through its API.

QualiSynthqualisynth.com

Qualitative research platform. Runs AI-moderated interviews with real and synthetic consumers, then turns full conversations into structured evidence — drivers, barriers, needs, and quotes — for agencies and insight teams.

qualisynth.com
Study modes
Human · Synthetic · Hybrid
Primary output
Drivers, barriers, needs, quotes
Built for
Agencies · Insight teams · Product research
ArenaSyntharenasynth.com

High-stakes practice platform. Professionals rehearse the conversations that matter — negotiations, objections, difficult decisions — against synthetic counterparts that hold a position under pressure instead of folding on cue.

arenasynth.com
Session type
Live practice against a counterpart
Primary output
Rehearsal + feedback on what moved them
Built for
Sales · Negotiation · Professional training
Datasets

Conversation data that explains itself.

The same engine generates datasets at scale. You are not buying more text — you are getting what was going on underneath: what each speaker wanted, what they believed, and how the relationship moved turn by turn.

A typical free dataset
  • • text
  • • speaker labels
  • • limited psychological consistency
  • • little or no explicit state
A StrataSynth dataset
  • • text + intent + goal + communication act
  • • belief state and belief delta
  • • relationship state and trajectory
  • • reproducible via seed and versioning

Four ways to generate them

Same engine underneath — pick the entry point that fits how your team works.

Dashboard
No code

Configure scenarios, inspect jobs and preview datasets before generating at scale.

CLI
Automation

Generate and evaluate from the terminal with reproducible, scriptable runs.

Python SDK
ML pipelines

Drop it into notebooks and training pipelines. Export straight to DataFrames or Hugging Face.

API
Product teams

The same HTTP API our own verticals run on. Embed generation and interviews in your product.

Read the docs
From the blog

How the engine actually behaves

Jul 6, 2026
Beyond Roleplay: Why Synthetic Personas Must Hold Identity Under Pressure

A comparative test between StrataSynth and two leading general-purpose LLMs, using a single senior persona pushed through twenty adversarial turns. The finding: the next frontier is not better text, but identity that holds under pressure.

evaluationidentity-under-pressure
Jun 19, 2026
Culture Is a Procedure: Decision Architectures Under Pressure

When you remove the obvious answer, the engine stops rendering opinions and starts rendering process — structured, market-distinct, and measurable. The population-level version of identity under pressure.

belief-statecross-cultural
Jun 4, 2026
Can Cultural Conditioning Survive Negotiation Pressure?

The same two synthetic negotiators close the same B2B deal 100 times — 50 in Great Britain, 50 in the United States — changing nothing but country conditioning. Some cultural signals weaken under pressure. Others strengthen. That asymmetry is the finding.

cross-culturalbenchmark
May 10, 2026
PsycheBench: an open benchmark for Synthetic Identity Engineering

Every system producing synthetic personas claims they are realistic. There has been no standard way to verify that claim. PsycheBench is the first open evaluation suite — three dimensions, deterministic scoring, no LLM judges.

benchmarkevaluation
May 8, 2026
Synthetic Humans with Coherent Personalities: What Emerged When We Stopped Writing Scripts

We built Sofía Martínez Rojas without sales scripts or objection trees. What came back was something we didn't program: identity under pressure.

emergenceidentity-under-pressure
Apr 19, 2026
Modeling Identity Under Pressure: A Step-by-Step Analysis of Synthetic Behavior

Most synthetic personas behave consistently — until they don't. A structured seven-phase analysis of coherence under sustained pressure.

belief-statecase-study
Apr 15, 2026
Synthetic Identity Engineering: a working definition

Most synthetic personas are costumes. Synthetic Identity Engineering is the practice of building the person underneath — belief systems, defense mechanisms, and identity stability.

definitioncategory
Apr 11, 2026
The synthetic person who refused to yield

María del Carmen Ruiz held her position under sustained philosophical and emotional pressure. A record of what that reveals about how the system works.

democase-study
Apr 11, 2026
What a Belief State Looks Like in a Synthetic Conversation

Most dialogue datasets give you the words. This explains the four belief dimensions recorded per turn — trust, hostility, self-worth, resolution — and why causal ground truth matters.

belief-statecognition
Apr 11, 2026
Why We Don't Use LLMs to Evaluate LLM-Generated Data

Evaluating AI-generated dialogue with another AI creates a circular system where both share the same failure modes. Here are the 12 deterministic metrics we use instead.

evaluationmetrics
Apr 11, 2026
8,404 turns of psychologically grounded dialogue — now on HuggingFace

Four datasets where cognitive ground truth — intent, goal, belief state, relationship dynamics — was computed before the text was generated, not inferred afterward.

datasetshuggingface
Questions

The things people ask before the demo

How is this different from asking an LLM to role-play a persona?

A system prompt describes a person; this builds one. Identity, beliefs and relationship state live outside the text as explicit structure, and the engine decides intent and goal before rendering a sentence. That is why the character holds when a user pushes back for twenty turns instead of quietly becoming the model again.

How do you measure quality without using an LLM as a judge?

Metrics are computed with numpy, scikit-learn and sentence-transformers — never by a model grading output. A judge that shares the generator's blind spots hides problems instead of surfacing them. It also means we can measure whether a cheaper model degrades results, rather than having an opinion about it.

Which countries does it cover, and is the culture real or a stereotype?

Eight countries with full conditioning: Spain, the United States, the United Kingdom, Germany, France, Italy, Mexico and Brazil. It comes from published population data rather than word lists — ages, for instance, are sampled from real national population pyramids (Eurostat, IBGE, US Census, CONAPO). Changing the country changes measured behaviour, not vocabulary.

Can I see real output before talking to anyone?

Yes, and without giving us an email. Public datasets are on Hugging Face under CC BY 4.0, the live demo runs in the browser, and notebook 04 runs an auditable A/B against your own LLM with your own API keys — so you can check the claim rather than take it.

Can I get a dataset built for my own scenario?

That is the main thing we sell. You specify the scenario, the countries, the number of people and the volume; it comes back with the same quality metrics and documentation as the public datasets. Tell us what you are trying to train or test.

Who owns the data, and are there real people in it?

The output is yours. There are no human subjects and no personal data in it — the people are synthetic, generated from population statistics rather than sampled from anyone. That removes the consent and data-protection conversation that a human panel requires.

Talk to us

Tell us what you need synthetic people for.

Custom datasets for your own scenario, a vertical built on the engine, or something we have not thought of yet. We would rather hear the problem first than sell you a plan.

• Custom datasets — your scenario, your countries, your volume
• The two products above, if one of them already fits
• Guidance on whether synthetic data actually helps your case
Provided by HexaForms