Live opening · Posted 5 days ago

Data Engineer - Python AIML Data Retrieval & Augmentation

Virtusa · Hyderabad, Telangana, India (On-site)
Linkedin No
You are 5 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 5 days ago
CompanyVirtusa
LocationHyderabad, Telangana, India (On-site)
Work modeNo
SourceLinkedin
Listed5 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
565 min from Linkedin publishing this role to us finding it
28 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
15,713 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Data Engineer
Working on top of worldwide research content, actions & knowledge, our Information Discovery,
Machine Learning & Data Engineering teams explore new ideas and develop novel, AI-powered
solutions which help publishers & researchers to know more, do more & achieve more. At the same
time, our goal is to highlight low-quality research and ensure research (data) integrity identifying
signals of potential research misconduct.
We are looking for a talented and self-directed Data Engineer to work on the following areas
Information Retrieval, Semantic Search, Recommendation services (on top of Elasticsearch
also leveraging LLMs/Embeddings and GenerativeAI)
Large scale distributed data processing (based on Apache Spark / Scala)
Streaming data processing (based on Google DataFlow/ApacheBeam)
Senior-level ML engineers are preferred, but all levels may apply.
Requirement Military obligations fulfilled for male candidates
How you will make an impact
Work cross-functionally collaborating with AI/ML engineers and product management to
develop AI-powered, highly efficient data & search solutions on all the above-mentioned
areas.
Design, build, orchestrate, and scale various streaming or batch processing data pipelines
utilizing optimal patterns, technologies and frameworks.
Be comfortable navigating the following technologies and programming languages Java,
Python, Scala, ElasticSearch, Apache Spark, Spring Boot, Apache Beam, Airflow etc.
What we look for
Minimum Qualifications
MSc with strong GPA in Computer Science/ Engineering
2+ years of work experience in data, backend and/or search engineering
Solid background and experience in CS and S/W development (especially in backend
development, databases etc)
Experience with Java (and/or Python) backend APIs / development frameworks also
supporting long running tasks (e.g., Spring Boot, FastAPI), microservices architecture, IoC/DI
patterns and containerization.
Strong Engineering Skills And Experience In Java
Background and experience with (at least one) of the following technologies
Lucene-based search Engines (e.g., ElasticSearch)
Distributed Data Processing using ApacheSpark (preferably in Scala)

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App