Live opening · Posted 34 days ago

Member of Technical Staff, Cekura (San Francisco, In-Person)

Cekura · San Francisco, CA, US | 0.20% - 0.80%
Ycombinator No Any (new grads ok)
You are 34 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 34 days ago
CompanyCekura
LocationSan Francisco, CA, US | 0.20% - 0.80%
ExperienceAny (new grads ok)
Salary$150K - $200K
Work modeNo
SourceYcombinator
Listed34 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
6 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
71,405 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Read This First
We work with unusual intensity. In-person in San Francisco, six days a week, long days, most weekends. This is not a phase we'll grow out of. It's how we've chosen to build, because we're in a market where speed decides who wins.
We're telling you this in the first paragraph, not the last, because we only want people who read that and feel pulled in, not talked into it. If you want a 9-to-5 (genuinely, no judgment), this isn't your role, and we'd rather you know now.
Here's what you get in exchange:
Top-of-market cash. We don't pay average salaries and ask for extraordinary hours. The comp reflects the commitment.
Meaningful equity you can believe in. Significant grants, employee-friendly terms, and our intent to create liquidity opportunities as we raise.
Zero commute tax. We support you living close to the office, with dinner at the office every night. Your hours go into building, not commuting.
Founders in the trenches. We work the same schedule we ask of you. This is a shared war, not extraction.
Compression of a decade into two years. You'll ship more, own more, and grow faster here than anywhere paying you to coast.
About Cekura
Cekura (YC F24) is building the voice AI engineer. Teams use Cekura to test agents before going live, monitor real production calls, and self-improve continuously. Cekura doesn't just flag issues and suggest fixes: it reproduces failures in simulation, fixes them, tests the fix thoroughly, and raises PRs. The platform spans pre-production simulation, LLM-powered evaluation, adversarial red-teaming, production monitoring with live drift detection, and cross-provider benchmarking (Vapi, Retell, Pipecat, LiveKit, ElevenLabs, and more).
We're trusted where reliability is non-negotiable: customers include Five9, HighLevel, Twin Health, PwC, Deloitte, Jobber, and Jotform, with HIPAA, SOC 2, and GDPR as defaults, not checkboxes. We're growing fast and backed by top investors.
About the Role
You'll build the core of Cekura: the simulation engines, evaluation systems, self-improvement loops, and observability pipelines our customers rely on to ship voice agents with confidence.
We deliberately don't split this into "software engineer" vs. "AI engineer." The interesting problems live at the boundary: real-time voice infrastructure meets LLM-as-judge evaluation, distributed systems meet RL-style self-improvement loops, telephony meets audio and speech analysis (ASR quality, barge-in, latency, prosody), and classic NLP meets frontier agentic behavior. You'll work across that whole surface.
What You'll Do
Build the testing and simulation engine. Design and ship systems that simulate thousands of realistic conversations against customer agents across voice, chat, and phone, with control over personas, interruptions, background noise, and edge cases.
Push the frontier of agent evaluation. Build LLM-powered evaluators, metrics, and the closed self-improvement loop at the heart of the product: detect failure, reproduce in simulation, generate the fix, test it thoroughly, and raise the PR, all autonomously. This includes adversarial red-teaming for jailbreaks, PII leaks, and off-script behavior, plus production monitoring with live drift detection.
Do applied audio and speech research. Go beyond the transcript. Separate background noise from actual speech, and map the paralinguistic layer of conversation (emotion, tone, silences, hesitations, speaking rate, overlaps) into structured signals, so agents are evaluated on how something was said, not just what.
Own real-time voice infrastructure. SIP, WebRTC, WebSockets, STT/TTS pipelines, and providers like Twilio, Vapi, Retell, LiveKit, and Pipecat. Latency, barge-in, and audio quality are first-class problems here.
Ship end-to-end. Take features from design to production. You own the full stack of what you build: backend, infra, evals, and the product surface customers touch.
Shape how we build. We're a small, senior team. Your architectural decisions, code standards, and technical taste will compound as the team grows.
About You
A strong generalist engineer excited to work across systems and AI, not someone looking to stay in one lane.
You've built and operated production systems and care about reliability, latency, and correctness.
Fluent in Python and comfortable picking up whatever the problem needs.
Strong instincts about LLMs: where they fail, how to evaluate them, how to build reliable systems on unreliable models.
You like ambiguity, move fast, and raise the bar for the people around you.
You've read the first section of this JD twice and you're still here.
Bonus: you're an ex-founder or aspire to be one.
Minimum Qualifications
Strong coding ability in Python, TypeScript, or Go.
Experience with at least one of: distributed systems, real-time infrastructure, or LLM-based products.
Strong written and verbal communication.
Nice to Have
Hands-on experience with LLM evals, agent frameworks, or AI observability.
Voice AI experience: Twilio, SIP, WebRTC, Vapi, Retell, LiveKit, Pipecat, or STT/TTS pipelines.
Early-stage startup experience, or early engineer at a dev-tool, infra, or AI company.
Open-source contributions or public technical work.
This Is Not for You If
You want a narrowly scoped role or a fixed tech stack.
You prefer working only on models, or only on infrastructure, never both.
You need rigid processes or heavy structure.
You're optimizing for work-life balance right now. (Later in your career, maybe. Here, now, no.)
You don't want to work in-person in San Francisco.
Why Cekura
The hardest problems in AI agent reliability: simulation, voice evals, real-time voice at scale.
A 12-person engineering team post-seed: early enough to matter, funded enough to move fast.
Work directly with founders and a highly technical team.
Meaningful equity, top-of-market compensation, fast growth.
US visa sponsorship available.
Medical, dental, vision, daily lunches and dinners, relocation support to live near the office.

Experience
Any (new grads ok)

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App