Live opening · Posted 7 hours ago

Senior Software Developer - Network and Collectives

Intel · 2 Locations
Workday
You are 7 hours behind. JobBeeper subscribers saw this role while it was still new.

At a glance

The key details from the original listing.

Posted 7 hours ago
CompanyIntel
Location2 Locations
SourceWorkday
Listed7 hours ago

Your early-applicant advantage

Live timing from JobBeeper.

Live data
13 min from Workday publishing this role to us finding it
14 min median time from a role going live to a subscriber being told
6 hours subscribers had this role before this page existed
71,666 roles found in the last 24 hours — the newest are not on this site yet
Start your free trial →

About the role

Description supplied by the original job listing.

Job Details:
Job Description:
About the team:
Network and Collectives owns the scale-out communication substrate for our AI accelerator - the NCCL equivalent for our hardware. We turn many chips into one machine: all-reduce, all-gather, reduce-scatter and point-to-point, made topology-aware and mapped onto the rack/pod interconnect.
We are measured on real fabric across a pod, not on a single card, and we pair tightly with the Tray/Rack/Pod/Cluster HW pillar and the Multi-Node Runtime team.
What you will do:
Design and implement collective algorithms tuned to our interconnect topology and bandwidth/latency profile.
Build a topology-aware transport layer over the tray/rack/pod.
Optimize end-to-end collective performance across multi-node pods; profile, find bottlenecks, and close the gap to fabric peak.
Co-design with HW on interconnect features, and with Multi-Node Runtime on partitioning, overlap, and scheduling.
Own correctness and numerical determinism of reductions across scale-out.
Qualifications:
5+ years AI/Systems/HPC software in C++; strong concurrency and lock-free design.
Hands-on with collective libraries (NCCL/MPI/etc) or scale-out communication.
Working knowledge of the communication patterns behind tensor, pipeline, and expert parallelism, and how they map onto collectives and the underlying fabric.
Performance engineering: profiling, bandwidth/latency tuning, roofline reasoning.
Nice-to-have
Topology/placement algorithms; congestion control; multi-tenant fabrics.
Large-scale distributed training/inference exposure.
Experience with large-scale inference serving stacks (vLLM, SGLang, TensorRT-LLM, DeepSeek).
RDMA/InfiniBand/RoCE, GPUDirect-style transfers, or comparable fabric experience.
Job Type:Experienced Hire
Shift:Shift 1 (Israel)
Primary Location: Israel, Haifa
Additional Locations:Israel, Petah-Tikva
Posting Statement:All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.Position of TrustN/A
Work Model for this Role
This role will be eligible for our hybrid work model which allows employees to split their time between working on-site at their assigned Intel site and off-site. * Job posting details (such as work model, location or time type) are subject to change.
*

Get JobBeeper Mobile App

Never miss a job opening! Get instant job alerts on your phone.

Subscribers see fresh openings within minutes. Download the JobBeeper App on Google Play to get real-time push notifications and apply before anyone else.

⚡ Instant Push Alerts 🎯 Tailored Filters 🚀 Direct Employer Links
GET IT ON Google Play

More openings worth a look

Recently tracked roles with full details and direct application links.

6 roles
Good roles move before most people even see them. Tell JobBeeper what you want and get fresh matches delivered in minutes.
Start your free trial →
⚡ Get fresh job alerts 📱 Get App