Live opening · Posted 5 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
We’re looking for a Founding AI Engineer to build and scale production-grade conversational intelligence systems. You’ll work deeply at the model and infrastructure level—owning LLM pipelines, retrieval systems, and real-time inference optimization from scratch.
What You'll Do
Design, build, and optimize end-to-end LLM pipelines for conversational AI
Develop and deploy RAG systems using vector databases in production
Optimize model inference for latency, memory, and cost (GPU-level optimizations)
Own production ML systems including model serving, monitoring, and reliability
Collaborate closely with product and engineering to ship real-world AI features
Must have
3+ years of experience in AI / ML Engineering Strong hands-on experience with LLMs and conversational AI systems Experience building and deploying RAG pipelines in production Strong understanding of vector databases Experience with Transformers / modern NLP models Hands-on experience in model serving, inference, and production ML systems Ability to optimize latency, memory, and cost for inference workloads Experience building end-to-end AI systems from development to deployment
Good to have
Experience as a Founding Engineer or early startup engineer Exposure to health-tech / wearable tech domains Experience with real-time inference infrastructure Familiarity with GPU-level optimization Experience with monitoring, reliability, and ML system observability Strong product thinking and ability to ship real-world AI features Experience working in a 0→1 / seed-stage startup environment Ability to take full ownership of AI systems end-to-end
Work arrangement
Hybrid
More openings worth a look
Recently tracked roles with full details and direct application links.