Live opening · Posted 2 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
SynthMind AI is building next-generation LLM serving infrastructure. We need a Senior ML Infra Engineer to optimize inference latencies and build distributed training pipelines.
Build high-throughput GPU serving infrastructure for LLMs.
Reduce model inference latency and GPU memory overhead.
Collaborate with AI research scientists to deploy experimental models.
Python, C++, PyTorch, CUDA, and Triton inference server experience.
Hands-on experience scaling GPU cluster workloads.
M.S. or B.S. in Computer Science or related field.
Top-tier competitive salary + Generous AI equity grant.
Unlimited PTO policy.
Premium hardware selection (MacBook Pro M3 Max / RTX Workstations).
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.