Commercial
Synthesia Roleplay Sessions: A Hands-On Review (2026)

Synthesia Roleplay Sessions launched on 22 July 2026. If you searched "Synthesia roleplay" you are probably trying to work out whether the biggest name in AI video has just built the thing you were about to buy from a specialist. Short answer: they have built a credible first version of it. Longer answer: it is a first version, and the launch says something interesting about where corporate training is going.
I founded Real Talk Studio. We have been building AI conversation practice since 2024, so read this with that in mind. The product facts below come from Synthesia's own pages, the launch coverage, and time spent inside their public demo, deliberately trying to break it. It scored me 13 per cent and failed me, which was fair. What the character did while I earned that score is the most useful finding in this article.
Here is the summary. Roleplay Sessions is 1-1 spoken practice with an avatar, scored against a rubric, with manager analytics. It is enterprise-only, sold through sales calls, with no published price. Practice runs in four languages. There is no team practice. If you are a large Synthesia video customer and you want practice bolted onto the videos you already make, it is a sensible add-on conversation to have with your account manager. If practice is the thing you are buying, compare it against tools that have been doing this for a while, because the gap shows. And whoever you buy from, one roleplay tacked onto the end of a course, done once per learner, has no return on investment. Practice pays back through repetition or not at all.
What Roleplay Sessions is
Roleplay Sessions is a live practice conversation with one of Synthesia's Interactive Avatars. A learner opens an assigned scenario, talks to the avatar, gets pushed back on, and is then scored against a skills rubric. An AI coach walks through the attempt afterwards. Managers get analytics: pass rates, per-skill scores, improvement across attempts, exportable to the LMS via SCORM.
The launch templates are cold call, discovery call, field sales, customer conversation, and manager feedback. You can build custom scenarios inside Synthesia Studio, the same editor their video customers already use. It is the first product in a broader platform they are calling Sessions, with job interviews and candidate screening reported as next.
The strategic read, which TechCrunch made explicitly, is that Synthesia is moving from generating training content to proving that training worked. That is the right instinct. It is the same instinct our whole company is built on.

One learner, one face on a screen. That is the shape they launched.
What Synthesia is good at
Credit where it is due, because a lot of it is due.
Synthesia is the most successful AI video company in the world. Over 50,000 companies, a hundred million dollars in annual revenue, logos like SAP, Merck, and Heineken. Their avatars are proprietary and their production quality is the benchmark everyone else gets measured against. If your brief is "turn this policy document into a training video in 40 languages by Friday", nobody does that job better.
The enterprise wrapper on Roleplay Sessions is also real: SOC 2 Type II, ISO 27001, 27701 and 42001, GDPR with EU data residency, SSO and SCIM. Their security page will sail through your procurement team. ISO 42001, the AI management standard, is still held by a small number of companies worldwide. We are one of them, so consider this the rare compliment paid across a certification both sides had to earn.
And they have distribution most practice specialists would give an arm for. Rolling practice out to a workforce that already logs into Synthesia for video is a shorter conversation than introducing a new vendor.
I tried the demo. Here is what happened
Their feature page has a public demo, the Coaching Check-in: you play a manager giving feedback to a direct report called Carol Liu. I ran it the way I test every practice tool, including ours. I played the worst learner in your organisation. I wandered off topic, repeatedly. I got personal instead of specific. I did the things real people do when a conversation gets uncomfortable.

Carol Liu, Direct Report. The avatar looks excellent. The conversation did not.
The good first, because there is some. The avatar looks excellent and the speech is natural, which is what you would expect from the best avatar company in the world. There is a Retry button sitting right next to Continue, which tells me the designers understand where the value lives. And when the conversation ends, a coach avatar debriefs you on your attempt before you see the score. Talking through your performance beats reading about it, and I can vouch for that design choice because we recently shipped the same thing.

A coach walks you through the attempt before the score. That design choice is the right one.
Now the test results. Carol did not react to any of it. I went off topic again and again and she carried on, pleasant and composed, as if the conversation were going fine. She did not challenge me, did not push back on the drift, and did not end the call, which is what an actual direct report does, one way or another, when their manager makes it personal. The system scored me 13 per cent and stamped it Fail, which was the correct verdict. But the verdict arrived afterwards, as a report. Inside the conversation there were no consequences at all. The character and the scoring engine were watching two different meetings.

13 per cent, Fail. Correct verdict. It arrived after the conversation had already stayed polite.
The feedback itself was thinner than the polish suggests. The rubric criteria are well chosen: being specific, setting the right tone, staying calm under pressure, agreeing on what happens next. But the output was two bullets on what worked, two on what to try, and a set of percentages that mostly read as demonstrated or not. My conversation was a 13 per cent disaster with a lot to say about it. The coach said a paragraph.

Two bullets each way. Retry sits next to Continue, which is the one thing I would keep.
And the conversation felt flat. The replies were generic, the exchange stunted, closer to a scripted branching exercise with a face than to the awkward, overlapping, half-finished way people talk when a conversation matters. Which points at something bigger than one demo.
The whole category has a perfection problem, and Synthesia's avatars are the most perfect of all. Real difficult conversations are messy. People interrupt, trail off, contradict themselves, lose their thread when they get emotional. That mess is precisely what learners freeze on in the real meeting, so we engineer it in deliberately: characters that interrupt, clip their sentences, get confused, lose composure under pressure. A flawlessly composed avatar is rehearsal for a conversation that never happens.
One design note in fairness: the avatar renders quite small on screen, and my guess is that is a deliberate fidelity trade, since a smaller window invites less uncanny-valley scrutiny. Sensible. It also quietly concedes that even the best avatars in the business are not ready for a close-up, which supports a view I have held for a while: the avatar is the least important part of a practice tool. The behaviour is the product.
The usual caveat applies. This was one public demo scenario, stress-tested on purpose, and enterprise builds with custom scenarios may go deeper. But a public demo is a product's best foot forward, and this one showed me a beautiful face on a conversation engine that is not yet listening.
Where the product sits today
Here is the honest assessment. Roleplay Sessions in August 2026 looks like the conversation practice market looked roughly eighteen months ago. That is not an insult. It is what version one of a hard product looks like, and everyone in this category shipped a version like it. But if you are evaluating, you should know what the specialists have moved on to.
It is 1-1 only. One learner, one avatar. The practice market has moved to team scenarios: a manager running their first meeting after a redundancy announcement while four characters react differently, a witness in front of a four-politician select committee, a rep handling a buying committee. Most difficult conversations at work are not duets. We shipped team mode this year for exactly that reason, with characters that interrupt, challenge, and talk over each other, and we were not the only ones. You can watch one run rather than take my word for it.

A redundancy meeting, a buying committee, a select committee: the room is the point.
Practice runs in four languages. English, German, Spanish, and French. The videos around the practice support 160+ languages, which tells you the constraint is the live conversation engine, not the content. For a global rollout, check the languages your frontline speaks before the demo, not after.
The characters hold their shape. As the demo above showed, the avatars respond and push back politely, but nothing the learner does changes their state. No losing composure, no getting flustered, no ending the call because trust broke. Pressure is where practice earns its keep. A prospect who always stays professionally sceptical is a kinder prospect than the ones your team meets.
Feedback arrives at the end. Score, coach, debrief. Useful. The next step the market took is feedback in the moment, mid-conversation, while the learner can still change course. Think of the difference between a driving instructor who talks to you at the lights and one who posts you a report.
You cannot try it without a form. The "free roleplay" on their feature page asks for your business email before it starts, and every CTA on the page is "Talk to sales". That is the enterprise motion, and it is coherent with their business. It also tells you who this is for. If you want to put a scenario in front of your team this week without a sales cycle, this is not that.
None of these are permanent. Synthesia ships fast and has $200 million of fresh Series E money to spend. But you buy the product that exists, not the roadmap, and the product that exists is a solid first release into a category where first releases happened in 2024.
The pricing question
There is no published price for Roleplay Sessions. The FAQ points to a sales call to "learn about pricing options". Coverage of the launch says it started with large enterprises, with plans to open up to smaller companies and individuals "as prices go down", which is an interesting phrase to sit with. Prices that need to go down are prices that started high.
If you take the call, the number to pin down is not the pilot. It is the fully loaded cost per learner per year once practice is actually happening at the volume the analytics dashboard is designed to show off. Which brings us to the more interesting question.
The paradox in the launch
Read the Roleplay Sessions page carefully and there is a quiet admission running through it. The "before" column describes training that people watch and forget. Skill gaps surfacing too late. Training that rarely made people ready. That "before" is a fair description of a lot of corporate video training. Including, awkwardly, the product Synthesia built its first hundred million dollars on.
I do not say that to score a point. It takes discipline for a company to market against its own core product, and the fact that Synthesia is willing to do it is the strongest signal in the whole launch. The world's biggest training-video company has looked at where budgets are going and concluded that watching is not enough, practice is the product, and proof of readiness is what buyers will pay for.
We agree. It is the premise of our entire company: watching a video about a hard conversation is the theory test, practising it is the driving test, and nobody gets a licence on theory alone. When the category leader in passive content starts selling active practice, the argument about whether practice matters is over. The remaining argument is who does practice best.
The economics problem
Here is the structural tension, and it is worth understanding before you negotiate.
Synthesia's video business has beautiful economics. You render a video once and it plays ten thousand times for nearly nothing. Every additional viewer is close to free, which is exactly why they can support 160+ languages and self-serve plans with public prices.
Live roleplay is the opposite shape. Every minute of practice is a live avatar, live speech, and live reasoning, and the reasoning layer is reported to run on OpenAI's models, so part of the cost per minute is a bill from someone else. There is no rendering once. The ten-thousandth practice session costs roughly what the first one did.
That is not a flaw in Synthesia. It is the physics of the category, and we live with the same physics. But it explains two things about the launch. It explains why there is no public price: metered practice at enterprise scale is a hard number to put on a page next to a video product sold on unlimited replication. And it explains why the first customers are large enterprises: the margin structure works better with big committed contracts than with a card payment.
Our answer to the same physics was to publish the number and let buyers do the maths: usage-based pricing at a flat rate per pooled practice hour, shared across the whole team. You can disagree with our rate. You can at least find it.
The mental model problem
Now the provocative part, and it is aimed at buyers, not at Synthesia.
Most organisations bought Synthesia to save money. That was the pitch and it was true: videos in minutes instead of weeks, no crews, no studios, no reshoots when the policy changes. Less time, less cost, same training. Efficiency was the product.
Roleplay is the opposite purchase. It costs more per learner, not less. It takes employee time instead of saving it. Nobody can run a roleplay at 1.5x speed while clearing their inbox. If you buy Roleplay Sessions with the video mental model still in your head, you will do the obvious thing: bolt one roleplay onto the end of the course, ask each learner to complete it once, and report the completion.

A live session, not a module. The next attempt is the point.
That version has no return on investment. It is a quiz with better production values. You have paid live-avatar prices for the same artefact the LMS already gave you, and one attempt under no pressure tells you very little about what the person will do on the real call. The uncomfortable truth underneath is one we have written about before: nobody practises, because most organisations have never set a bar of conversational performance high enough to require rehearsal. Adequacy does not need reps. Buying an avatar does not change the bar.
Practice pays back through repetition or not at all. Rep one is finding your feet. The skill shows up on reps two, three, and four, which is the number a classroom roleplay never reaches because the facilitator has to move on. It is why our platform puts no cap on practice volume, and why learners across our client base average around four hours of practice each, not four minutes.
The research name for this is deliberate practice: focused repetition with feedback, at the edge of your current ability, sustained over time. It is how surgeons, pilots, and athletes get good, and it is the design principle behind the tools in this category that work. Yoodli deserves credit here, and we give it freely in an article about a different competitor: their whole product is built for habitual use, low friction, come back tomorrow. Ours is built the same way, with pooled hours so the marginal rep costs a team nothing.
So the question to put to Synthesia is not "can learners retry?" Their copy says yes, as often as they want, with improvement tracked across attempts. The question is what the product design rewards. Three signals pull toward one-time use: the roleplay is built inside the training video the team already watches, which anchors it to the course. It is delivered through SCORM into the LMS, which is the natural habitat of the completion. And every minute of practice carries a live compute cost, at a price that is not published, which means additional reps are a cost the buyer cannot forecast. When reps cost money, the finance incentive and the learning outcome point in opposite directions. Deliberate practice needs the next rep to feel free.
Which surfaces the real question, and it is the one we would ask any L&D leader before they take the Synthesia sales call or ours: will you give your people the time? We have watched deployments stall for exactly one reason, and it was never the avatar. No practice time in the diary, no workflow to embed it in, usage dies by week six. A tool that needs repetition, bought by an organisation that will not fund repetition, fails at any price.
So the most interesting test of this launch sits with Synthesia's installed base: tens of thousands of companies that came for cost saving, now being asked to fund more time, more cost, and more reps. If they make that shift, the whole category wins, us included. If they treat roleplay as a fancier final quiz, the analytics will show a thousand single attempts, no uplift, and an invoice that is hard to defend at renewal.
What the analytics are for
One more difference worth naming, because it decides who each product is for.
Synthesia's analytics answer a performance question: who is ready to sell, who needs another attempt, where the skill gaps sit. Their CEO has talked about mapping the talent in a company. That is a sales enablement and L&D lens, and for those buyers it is the right one.
UK buyers increasingly need the analytics to answer a harder question: can we show a regulator, a tribunal, or a board that our people were competent before the conversation went wrong? The Worker Protection Act's "all reasonable steps" duty and the FCA's non-financial misconduct rules both land in autumn 2026, and a completion certificate answers neither. What answers it is a dated record of a person handling the situation under pressure, scored, exportable, and attributable to them.
That is the job we built for. When OVO Energy's Director of Customer Service described what she wanted from us, her phrase was evidence she could take to the board, not a completion certificate.

Competence, confidence, compliance — the board artefact, not a completion rate.
If your driver for buying practice is conduct risk rather than quota attainment, ask every vendor on this page, us included, to show you the artefact their analytics produce and whether your legal team would recognise it as evidence.
Synthesia roleplay alternatives at a glance
If your practice need is sales-shaped and US-based, look at Hyperbound and Second Nature. If it is presentation delivery, Yoodli. If it is the hard 1-1 and the hard room, that is our job. And if what you mostly need is video with some practice attached, Synthesia is the honest answer, which is the next section.
When Synthesia is the right buy
- You are already a large Synthesia video customer and practice is an expansion, not a new evaluation. One vendor, one login, one editor is a real advantage.
- Your practice needs are the launch templates: cold calls, discovery, customer conversations, manager feedback, in one of four languages.
- Your organisation buys through long enterprise cycles anyway, so "Talk to sales" is not friction, it is the process.
- Avatar production quality is near the top of your criteria and you want the same faces in the video and the roleplay.
Those are real conditions and plenty of buyers meet them. This page exists for the ones who do not, and who searched "Synthesia roleplay" assuming the biggest name must be the furthest ahead.
Skip us if you mostly need video. We do not make videos. Synthesia is the best in the world at that, and buying a conversation tool to solve a content problem is how you end up with two disappointed teams.
How to evaluate without a twelve-week theatre
Same advice we gave in the Attensi comparison, because it holds:
- Give every vendor the same three job moments. One boring-operational, one hard conversation, one team room. Watch what happens to the vendors on the third one.
- Put a real learner in front of it without a host. If the demo needs a ceremony, adoption will die in week six.
- Ask what happens when the learner is rude, evasive, or upset. A character that stays serene is not practice, it is theatre.
- Price year two at real practice volume, per learner, fully loaded. If the vendor cannot put that number in writing, budget for the surprise.
- Before any of it, answer the question that sinks these deployments: how many reps per learner, and where in the working week does that time come from? If the answer is "one, after the course", save everyone's money, including yours.
- Ask for the artefact: a dated, exportable record of an attempt that a lawyer, an auditor, or a board would recognise as evidence.
FAQ
Frequently asked questions
01What is Synthesia Roleplay Sessions?
A live 1-1 practice conversation with a Synthesia Interactive Avatar, scored against a skills rubric, with AI coaching and manager analytics. Launched July 2026 as the first product in Synthesia's Sessions platform.
02How much does Synthesia Roleplay Sessions cost?
There is no published price. It is sold to enterprises through sales conversations. Coverage of the launch reported plans to reach smaller companies as prices come down.
03Does Synthesia Roleplay Sessions support team practice?
No. It is one learner and one avatar. If you need multi-character practice, a meeting, a panel, a buying committee, look at tools built for team scenarios.
04What languages does Synthesia roleplay work in?
Practice runs in English, German, Spanish, and French. The surrounding video content supports 160+ languages.
05What are the best Synthesia Roleplay Sessions alternatives?
For hard 1-1 and team conversations with transparent pricing, Real Talk Studio. For presentation and speech coaching, Yoodli. For enterprise sales roleplay, Second Nature or Hyperbound. For video creation itself, stay with Synthesia.
06Do Synthesia Roleplay Sessions give a return on investment?
Not if you use them the way most organisations use training video: one roleplay at the end of a course, completed once per learner. That is a more expensive quiz. The return comes from repetition, learners practising a scenario several times and improving across attempts, which means allotting real employee time. Budget for the time, not just the licence.
07Can Synthesia Roleplay Sessions provide compliance evidence?
It produces scores, transcripts, and analytics, exportable via SCORM. Whether that satisfies a Worker Protection Act "all reasonable steps" review or an FCA conduct file is a question for your legal team. If audit-ready behavioural evidence is the reason you are buying, make it the first question in the demo, not the last.
08Synthesia vs Real Talk Studio, which is which?
Synthesia is a video company that now also does practice. We are a practice company that does not do video. There is overlap in the middle. They are not built for the same job.
---
Want to see the difference rather than read about it? [Try a practice conversation](/practice). No form, no sales call, it starts in your browser.