●  DEVOTION — A MULTILINGUAL FULL-DUPLEX VOICE API

Machines that listen while they speak.

Devotion is a full-duplex voice model you call over an API, priced like any other usage-based AI API. It hears, thinks and speaks at once, in every language — built on the best open duplex models and NVIDIA's open weights.

JOIN THE WAITLIST → HOW IT WORKS ↓
FIG. 1 — THE THREE-BODY PROBLEM t = 000.0 · E = −1.29 ● LIVE
■ m₁ — a voice ■ m₂ — a voice ■ m₃ — the model
CLICK TO PERTURB
FIG. 1 — A conversation is a three-body problem: two voices and a model in mutual orbit, plotted pixel by pixel. Simple rules, never repeating. Click near a body to perturb it.

THESIS

Voice AI today is a walkie-talkie. It waits for you to finish, transcribes, thinks, then performs a reply.

Human conversation doesn't take turns. It overlaps — interruptions, backchannels, silences that mean something.

Full-duplex models finally live inside that timing. Devotion is the API that puts it in every language, for anyone building on top of it.

01 — APPROACHENHANCE, DON'T REBUILD
A.1

Stand on open models

The duplex breakthrough already happened — PersonaPlex, Moshi and their kin listen and speak in one continuous stream. Rebuilding that would burn compute on a solved problem. Devotion starts from open weights and spends every GPU-hour on what's missing, then ships it as a single API endpoint.

A.2

Multilingual is not translation

Timing is cultural. How long a polite pause lasts, when overlap is warmth and when it's rudeness, what a backchannel sounds like — all of it differs by language. A multilingual duplex model has to learn each language's rhythm, not just its words.

A.3

Post-training as a craft

Multilingual conversational speech, synthetic overlapping dialogue, and post-training that adds languages without breaking the duplex behaviours already there — with yield latency and backchannel timing benchmarked per language.

FIG. 2 — A DUPLEX EXCHANGE, SIMULATED LAST YIELD · 180 MS
HUMAN MODEL
FIG. 2 — Both channels are open the whole time. The model backchannels while listening and yields in ~180 ms when you barge in. Press INTERRUPT to try it.
02 — APPLICATIONSWHERE DUPLEX BECOMES AN ECONOMY

Half the economy runs on conversation — most of it not in English. Full-duplex makes machine voice employable; multilingual makes it global.

B.1

Customer operations

Agents that can be interrupted, corrected and talked over — and still hold the thread, in the customer's own language.

NEAR-TERM
B.2

Simultaneous interpretation

Listening and speaking at once is the job description. Translation that keeps pace with the speaker — not after them.

NEAR-TERM
B.3

Accessibility

Real-time conversational support for people who speak, hear or process differently. Timing is dignity: no dead air, no being talked over.

PILOT
B.4

Healthcare front line

Intake and triage that listens the way a nurse does — probing, confirming, reassuring while the patient is still talking.

PILOT
B.5

Embodied systems

Robots, vehicles and machines that coordinate with people by voice, in places where hands and eyes are already busy.

RESEARCH

The Devotion API is opening to design and compute partners first. Join the waitlist to get access when it does.

03 — PRICINGAPI ACCESS, PAY AS YOU GO

Devotion is sold the way any usage-based AI API is sold: no seats, no minimums, pay for what you stream.

D.1

Pay per minute

Billed on streamed audio-minutes, both directions. No plan, no lock-in — usage stops, billing stops.

D.2

One API, every language

Same endpoint, same price, regardless of language — no surcharge for going multilingual.

D.3

Waitlist pricing

Published rate card lands at general availability. Waitlist members get early access and locked-in early-adopter pricing.

04 — NOTESBUILD NOTES, PUBLISHED IRREGULARLY
№ 001

Devotion — a multilingual voice API, the mission

SEP 2026 — FIRST NOTE
№ 002

Yield latency across languages

IN DRAFT
№ 003

The three-body problem of conversation

IN DRAFT
05 — CAREERSFOUNDING ROLES
C.1

Founding engineer, real-time audio

Ship the Devotion API. Streaming inference and infrastructure at conversational latency.

OPENING
C.2

Founding ML engineer

Post-train Devotion on multilingual conversational data. Own model quality end to end.

OPENING

No forms, no cover letters. Send the most interesting thing you've built to hello@instituteofthinking.com

JOIN THE WAITLIST

Get access to the Devotion API before it's public.

Early sign-ups get first access and locked-in early-adopter pricing when the API opens.