Live opening · Posted 3 days ago

Linux Automation Engineer

Netweb Technologies India Ltd. · Faridabad, Haryana, India (On-site)
Linkedin No
You are 3 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 3 days ago
CompanyNetweb Technologies India Ltd.
LocationFaridabad, Haryana, India (On-site)
Work modeNo
SourceLinkedin
Listed3 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
6 min from Linkedin publishing this role to us finding it
14 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
63,206 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Job Summary – Linux Automation Engineer
Experienced Linux Automation Engineer responsible for designing, developing, and managing enterprise-scale Linux infrastructure automation solutions, HPC cluster management platforms, and monitoring systems. The role focuses on automating server provisioning, operating system deployment, configuration management, patching, infrastructure lifecycle management, and cluster orchestration to improve operational efficiency and scalability. Based on the job description, the position also involves leading technical initiatives, collaborating with cross-functional teams, and driving best practices in software development, DevOps, and infrastructure automation.
The position requires strong expertise in Linux administration, Python development, infrastructure automation tools, and observability platforms. The engineer is expected to build and enhance scalable backend systems, REST APIs, centralized monitoring solutions, and management dashboards while ensuring high availability, security, and maintainability of infrastructure environments. Responsibilities also include implementing CI/CD pipelines, integrating enterprise platforms, and supporting HPC environments through cluster provisioning, resource monitoring, scheduler integration, and performance analytics.
Key Responsibilities:
HPC & Parallel Storage Design: Architect, tune, and maintain high-performance, multi-petabyte scale Parallel File Systems (PFS) like Lustre, IBM Spectrum Scale (GPFS), or BeeGFS. Optimize data-in-transit pipelines over low-latency InfiniBand or RoCE fabrics.
Private Cloud Infrastructure: Design, configure, and bootstrap custom on-premises private clouds using OpenStack, VMware NSX-T, or KVM-based hypervisors (Proxmox/RHEV).
Custom Tooling & Scripting: Act as a developer within operations. Write advanced, production-grade automation scripts and internal CLI tools in Python or Go to interface directly with cloud APIs and manage resources.
End-to-End Infrastructure Automation: Champion Infrastructure as Code (IaC) by creating modular Terraform plans and massive Ansible Playbooks/Roles to automate compute provisioning, job schedulers, and node validation.
Hands-on DevOps Pipelines: Construct and manage multi-stage Jenkins or GitLab CI pipelines to test and securely deploy system configuration changes across development, staging, and cluster nodes.
1. Technical Leadership:
· Lead and mentor a team of Full Stack Developers, Backend Developers, and Automation Engineers.
· Define technical architecture, coding standards, and development best practices.
· Conduct code reviews, design reviews, and technical feasibility assessments.
· Drive product roadmap discussions and technical decision-making.
· Collaborate with Product, Infrastructure, Validation, and Support teams.
2. Platform & Product Development:
Architect and develop infrastructure management platforms, including:
· HPC Cluster Management Tools
· Centralized Monitoring & Alerting Platforms
· Infrastructure Lifecycle Management Systems
· Asset & Inventory Management Solutions
· Capacity Planning & Analytics Platforms
· Infrastructure Automation & Orchestration Tools
3. Linux & Infrastructure Automation:
Design automation frameworks for:
· Server Provisioning
· OS Deployment & Configuration
· Firmware & Driver Compliance
· Cluster Deployment
· Patch Management
· Infrastructure Health Checks
· Automated Validation & Benchmarking
· Develop integrations with Linux-based infrastructure and enterprise platforms.
4. Monitoring & Observability:
Design centralized monitoring solutions for:
· Servers
· Storage
· Networking
· GPU Clusters
· HPC Infrastructure
· Data Center Environments
· Integrate monitoring tools such as Prometheus, Grafana, Open Telemetry, ELK, Redfish, SNMP, and IPMI.
· Build predictive monitoring, alerting, and analytics capabilities.
5. HPC & Cluster Management:
Develop and enhance cluster management capabilities including:
· Node Discovery & Registration
· Cluster Provisioning
· Resource Monitoring
· Job Monitoring
· Scheduler Integration (Slurm /Open HPC)
· Health & Performance Analytics
· Cluster Lifecycle Management
6. Software Engineering:
· Design scalable backend architectures and REST APIs.
· Guide development of web-based dashboards and management portals.
· Implement CI/CD pipelines and DevOps best practices.
· Ensure high availability, scalability, security, and maintainability of developed solutions.
Required Technical Skills
Linux Internals: Master-level knowledge of RHEL, Rocky Linux, or Ubuntu Server (advanced kernel tuning, memory management, and high-performance network configurations).
HPC Stack: Direct experience deploying and managing Slurm/PBS Pro alongside a verified track record handling parallel storage layers (Lustre, GPFS).
Languages: Advanced proficiency in Python or Go, and robust Bash shell scripting.
Automation Frameworks: Enterprise-level implementation of Terraform and Ansible/AWX.
Fabrics & Interconnects: Comprehensive understanding of InfiniBand routing, Subnet Managers, and network topology tuning for high-performance computing.
Linux Expertise
· RHEL, Rocky Linux, Ubuntu, Alma Linux
· Linux System Internals
· Performance Tuning & Troubleshooting
· Shell Scripting
Programming
· Python (Expert)
· Go Lang (Preferred)
· JavaScript/TypeScript
· REST API Development
Automation & Infrastructure
· Ansible
· Terraform
· Git
· Jenkins/GitLab CI/CD
Monitoring & Observability
· Prometheus
· Grafana
· ELK Stack
· Open Telemetry
· Zabbix/Nagios
· Redfish/IPMI
Databases
· PostgreSQL
· MySQL
· MongoDB
Front-End Awareness
· ReactJS / Angular (working knowledge to guide development teams)
· Dashboard & Visualization Frameworks
HPC Technologies (Preferred)
· Slurm
· Open HPC
· GPU Clusters
High

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App