All rolesApply
Open role
Founding Software Engineer — AI & Voice Systems
AI & infrastructure · Remote · US or Europe
Own the machinery that makes Roark honest: the evaluation pipelines that score every call, the audio-native metrics that hear what transcripts miss, and the simulation engine that dials agents like a real caller. This is the product — if it is wrong, everything above it is wrong.
What you'll do
- Build and scale evaluation pipelines that score thousands of production calls in near real time.
- Work on audio-native metrics — pronunciation, emotion, vocal stress — alongside LLM-based evaluators, and prove they agree with human judgment.
- Evolve the simulation engine: realistic personas, accents and noise, dialing real agents over the phone and WebRTC.
- Own analytics at scale: ClickHouse pipelines, DynamoDB streams, and the query layer that keeps dashboards honest.
- Orchestrate long-running work with Temporal — replays, backfills and load tests that converge instead of double-count.
What you'll bring
- 4+ years building backend or ML/audio systems in production, ideally in TypeScript or Python.
- Experience with at least one of: speech/audio processing, LLM evaluation pipelines, or high-volume data infrastructure.
- Rigour about correctness — you think in idempotency, replays and ground truth, not just happy paths.
- Comfort operating what you build: dashboards, alerts, and the 2am curiosity to find out why.
- Clear written communication — designs and post-mortems people actually want to read.
Nice to have
- Hands-on time with telephony/WebRTC (Twilio, LiveKit, SIP) or TTS/ASR systems.
- ClickHouse, Temporal, or stream-processing experience at scale.
Stack: TypeScript · Bun · AWS (SST) · ClickHouse · DynamoDB · Temporal · Postgres · WebRTC/telephony
Bring a recording.
We’ll score it live.
See your own agent measured on the audio it actually produced, in the demo, in real time. Stop guessing whether your voice AI works.
Or start free with $50 in credit · read the docs · support@roark.ai