Live opening · Posted 3 days ago

Principal Software Architect- High Performance Computing

Applied Materials · Chennai, Tamil Nadu, India (On-site)
Linkedin No
You are 3 days behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 3 days ago
CompanyApplied Materials
LocationChennai, Tamil Nadu, India (On-site)
Work modeNo
SourceLinkedin
Listed3 days ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
6 min from Linkedin publishing this role to us finding it
6 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
66,626 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Our Team
Our team is developing a high-performance computing solution for low-latency and high throughput image processing and deep-learning workload that enables our Chip Manufacturing process control equipment to offer differentiated value to our customers.
Your Opportunity
As an architect, you will get the opportunity to grow in the field of high-performance computing, GPU compute infra, complex system design and low-level optimizations for better cost of ownership.
Roles and Responsibility
As a Software Architect, you will be responsible for design and implementation of robust, scalable infrastructure solutions combining diverse processors (CPUs, GPUs, FPGAs).
You will analyze and partition workloads to the most appropriate compute unit, ensuring tasks like AI inference and parallel processing runs on specialized accelerators, while serial tasks run on CPUs.
You will work closely with cross-functional teams, including Algo engineers, product managers, and business stakeholders, to understand requirements and translate them into architectural/software designs that meet business needs.
You will be coding and developing quick prototypes to establish your design with real code and data.
You will be a subject Matter expert to unblock software engineers in the HPC domain.
You will be expected to profile entire cluster of nodes and each node with profilers to understand bottlenecks, optimize workflows and code and processes to improve cost of ownership.
Conduct performance tuning and capacity planning, monitoring GPU metrics (e.g., using NVIDIA DCGM) for reliability
Evaluate and recommend appropriate technologies and frameworks to meet project requirements.
Lead the design and implementation of complex software components and systems.
Ensure that software systems are scalable, reliable, and maintainable.
Your primary focus will be on ensuring that the software systems are scalable, reliable, maintainable and cost effective.
Our Ideal Candidate
Someone who is passionate about and has deep understanding and experience in design and development of cutting edge HPC systems and heterogenous computing infrastructure. He should have very good hands-on experience in parallel programming (CUDA) and AI inference infrastructure. He should be able to multi-task and switch contexts based on business needs.
Qualifications
12 to 18 years of experience in implementing robust, scalable, and secure infrastructure solutions combining diverse processors (CPUs, GPUs, FPGAs)
Working experience of GPU inference server like Nvidia Triton.
Very good knowledge C/C++, Data structure and Algorithms and complexity analysis.
Experience in developing Distributed High Performance Computing software using Parallel programming frameworks like MPI, UCX etc.
Experience in GPU programming using CUDA, OpenMP, OpenACC, OpenCL etc.
In depth experience in Multi-threading, Thread Synchronization, Inter process communication, and distributed computing fundamentals.
Experience in Inter Process communication using Shared memory and Pipes.
Experience in performance profiling at application and system level (e.g. vtune, Oprofiler, perf, Nividia Nsight etc.)
Experience in low level code optimization techniques using Vectorization and Intrinsics, cache-aware programming, lock free data structures etc.
Familiarity with microservices architecture and containerization technologies (docker/singularity) and low latency Message queues.
Excellent problem-solving and analytical skills.
Strong communication and collaboration abilities.
Ability to mentor and coach junior team members.
Experience in Agile development methodologies.
Additional Qualifications:
Experience in HPC Job-Scheduling and Cluster Management Software (SLURM, Torque, LSF etc.)
Good knowledge of Low-latency and high-throughput data transfer technologies (RDMA, RoCE, InfiniBand)
Good Knowledge of Parallel processing and DAG execution Frameworks like Intel TBB flowgraph, OpenCL/SYCL etc.

Work arrangement
No

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App