Roark for Hospitality — voice agent quality for reservations, concierge & guest services - Roark

Your concierge takes a booking in 45 languages.

Guests call from a loud lobby, in Gulf Arabic, Brazilian Portuguese, or South Indian English — and they want a king suite, late checkout, and to feel welcome. Roark scores every call on the audio: accent clarity, warmth, the booking detail, the upsell.

Backed by YC

Live · scoring every call 1,284 today

Here's how a guest call goes wrong.

Each one is a botched booking, a guest who felt unwelcome, or revenue left on the table — and most are inaudible to a tool that only reads the transcript.

The accent it could not parse

Caller: I want a room for the thirtieth of Ramadan.

Agent: I'm sorry, did you say the thirteenth?

A Gulf Arabic or South Indian English caller is misheard, the date flips, and the booking is wrong from the first turn. Accent breakdown is an audio failure a transcript-only tool reads as a clean exchange.

The wrong room, confidently confirmed

Caller: Two queens, non-smoking, two nights.

Agent: Confirmed — one king, smoking, for two nights.

The agent read back the wrong details with total confidence. Roark checks the confirmed booking against what the guest actually asked for, every call.

Warm request, robotic reply

Caller: It's our anniversary — can you do anything special?

Agent: Anniversary noted. Is there anything else.

The words were fine; the delivery was cold and flat on the one call that should have felt personal. The audio model scores warmth and emotion the text can never show.

The suite it never offered

Caller: Do you have anything with a view?

Agent: Yes, our standard rooms face the courtyard.

A clear buying signal, and the agent never offered the ocean-view suite or late checkout. Roark scores whether the agent surfaced the upgrade the guest was reaching for.

Drowned out by the lobby

Caller: ...checking in under Okafor, party of four...

Agent: Sorry, could you repeat the name three more times?

Background noise from a packed lobby buried the guest, and the agent collapsed into repeat-loops instead of recovering. Roark scores noise robustness and ASR accuracy on the real audio, not the cleaned-up transcript.

Roark catches every one of these — and proves the fix.

Each failure above is filed with its evidence, becomes a repeatable simulation until a candidate passes, and is verified on your next thousand live calls.

Catch

The ledger above — every failure filed live, evidence attached.

Simulate

Your fix, replayed against the exact failures above.

Review

Every change explicit and diffed — you apply it.

Verify

You ship — Roark confirms the metric moved on live calls.

Simulate before launch

Run your agent against hundreds of simulated callers (realistic personas, accents, background noise and edge cases) and get every conversation scored before a customer ever dials in.

Scenarios & personas

Hundreds of simulated callers (the angry one, the rambler, the interrupter) built from your real call types.

45 languages & accents

Native accents, code-switching and background noise, in every market your agent answers.

500+ metrics out of the box

Get started

First call scored in under a minute.

One click on any platform below and production calls stream in on their own, or send any recording with a few lines of code.

import Roark from '@roarkanalytics/sdk'

const client = new Roark({ bearerToken })
await client.call.create({
  recordingUrl, startedAt,
  interfaceType: 'PHONE',
  callDirection: 'INBOUND',
  agent: { customId: 'support_v2' },
}) // scored in seconds

Bring a recording. We’ll score it live.

See your own agent measured on the audio it actually produced, in the demo, in real time. Stop guessing whether your voice AI works.