Machine translation has become extraordinarily good at the part of a message that can
be written down. It remains almost entirely indifferent to the part that cannot.
Pitch, timing, stress, hesitation, register, the small refusals and concessions that
a speaker never states outright. These are not decoration around the content. In most
human exchanges they are a substantial share of the content. A listener acts on
them before the sentence has finished. Strip them at the language boundary and two
parties can receive faithful translations of the same words and leave with materially
different understandings of what was agreed.
The stakes scale with the consequence of the conversation. In a negotiation, a
diagnosis, a deposition, a de-escalation, the discarded layer is often the one that
determined the outcome.
That gap is the company's research subject. Not translation
quality; the field has that well in hand. The transmission of everything translation
was never designed to carry.
Since 2025 the work has been concentrated in a single build, the RCG
Voice Model. What it is for is described by the four questions below. How it is
built is not published, and will not be. That boundary is set out on the
research philosophy page.