Live opening · Posted 5 days ago

LLM / GenAI Engineer

Evlo AI · Atlanta, GA (Remote)
Linkedin Yes
You are 5 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 5 days ago
CompanyEvlo AI
LocationAtlanta, GA (Remote)
Work modeYes
SourceLinkedin
Listed5 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
11 min from Linkedin publishing this role to us finding it
10 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
66,856 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

About The Role
The LLM / GenAI Engineer will design, build, and operate production AI systems spanning retrieval-augmented generation, tool-using agents, model adaptation, and automated evaluation. The role focuses on turning foundation models into reliable products with measurable gains in quality, latency, cost, and user experience.
Working alongside applied scientists, platform engineers, and product teams, this role will own critical components of the GenAI stack—from data and prompt pipelines to inference services, observability, and continuous improvement. The position is remote and aligned with the Atlanta, GA engineering organization.
Key Responsibilities
Design and implement production RAG systems using Python, LangChain, LlamaIndex, or custom orchestration frameworks
Build ingestion, chunking, embedding, retrieval, reranking, and grounding pipelines using vector stores such as Pinecone, Weaviate, Elasticsearch, or pgvector
Develop agentic workflows with tool calling, structured outputs, memory, human-in-the-loop controls, and robust failure handling
Create LLM evaluation frameworks covering offline benchmarks, LLM-as-judge workflows, golden datasets, hallucination detection, and regression testing
Fine-tune and adapt open-source models using supervised fine-tuning, LoRA, QLoRA, prompt optimization, or preference-based methods where appropriate
Deploy and optimize inference services on AWS, GCP, or Azure using Docker, Kubernetes, and APIs such as vLLM or Hugging Face TGI
Instrument production systems for latency, token usage, quality, safety, and cost; establish monitoring, alerting, rollback, and incident response practices
What We Are Looking For
3–8 years of software engineering, machine learning engineering, or applied research experience, including at least 1 year delivering LLM or GenAI systems to production
Strong Python skills with experience building maintainable services, asynchronous workflows, REST APIs, and automated tests
Hands-on expertise with LLM application patterns including RAG, embeddings, vector search, prompt engineering, function calling, and structured generation
Experience with PyTorch, Hugging Face Transformers, model fine-tuning, and inference optimization for open-source or hosted foundation models
Proficiency with cloud and production infrastructure, including AWS, GCP, or Azure; Docker, Kubernetes, CI/CD, and observability platforms
Bachelor’s or master’s degree in computer science, machine learning, data science, electrical engineering, or a related technical field
Bonus: Experience with Ray, vLLM, Triton, distributed training, multimodal models, privacy and safety controls, or enterprise-scale AI platform development

Work arrangement
Yes

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App