Live opening · Posted 5 days ago
At a glance
The key details from the original listing.
Your early-applicant advantage
Live timing from JobBeeper.
About the role
Description supplied by the original job listing.
Own the layer between a trained model and something that actually runs—kernels, executors, and the runtime choices that make latency and cost real.
Nebulai is a Humans & AI Agents Marketplace. As vetted contract talent, you join enterprise teams hardening AI runtimes: execution engines, graph compilers, memory layouts, and the path from a packed artifact to stable inference in production. Not a notebook wrapper. A runtime people can ship on.
You’ll be a strong fit:
- Solid Python plus C++ or systems comfort around runtimes
- Hands-on with ONNX Runtime, TensorRT, TorchScript/Inductor, TVM, XLA, or similar
- Care about memory, kernels, concurrency, and reproducible performance
- Comfortable pairing with ML, platform, and infra partners in enterprise settings
Bonus if you’ve:
- Built custom ops, plugins, or graph passes for a serving stack
- Tuned GPU/CPU runtimes for LLMs, vision, or multimodal models
- Shipped runtime upgrades without breaking production traffic
Contract work via Nebulai’s marketplace. Apply at https://nebulai.app
Work arrangement
Yes
More openings worth a look
Recently tracked roles with full details and direct application links.