About the team:
Network and Collectives owns the scale-out communication substrate for our AI accelerator - the NCCL equivalent for our hardware. We turn many chips into one machine: all-reduce, all-gather, reduce-scatter and point-to-point, made topology-aware and mapped onto the rack/pod interconnect.
We are measured on real fabric across a pod, not on a single card, and we pair tightly with the Tray/Rack/Pod/Cluster HW pillar and the Multi-Node Runtime team.
What you will do:
- Design and implement collective algorithms tuned to our interconnect topology and bandwidth/latency profile.
- Build a topology-aware transport layer over the tray/rack/pod.
- Optimize end-to-end collective performance across multi-node pods; profile, find bottlenecks, and close the gap to fabric peak.
- Co-design with HW on interconnect features, and with Multi-Node Runtime on partitioning, overlap, and scheduling.
- Own correctness and numerical determinism of reductions across scale-out.
- 5+ years AI/Systems/HPC software in C++; strong concurrency and lock-free design.
- Hands-on with collective libraries (NCCL/MPI/etc) or scale-out communication.
- Working knowledge of the communication patterns behind tensor, pipeline, and expert parallelism, and how they map onto collectives and the underlying fabric.
- Performance engineering: profiling, bandwidth/latency tuning, roofline reasoning.
Nice-to-have
- Topology/placement algorithms; congestion control; multi-tenant fabrics.
- Large-scale distributed training/inference exposure.
- Experience with large-scale inference serving stacks (vLLM, SGLang, TensorRT-LLM, DeepSeek).
- RDMA/InfiniBand/RoCE, GPUDirect-style transfers, or comparable fabric experience.
Job Type:Experienced HireShift:Shift 1 (Israel)Primary Location: Israel, HaifaAdditional Locations:Israel, Petah-TikvaPosting Statement:All qualified applicants will receive consideration for employment without regard to race, color, religion, religious creed, sex, national origin, ancestry, age, physical or mental disability, medical condition, genetic information, military and veteran status, marital status, pregnancy, gender, gender expression, gender identity, sexual orientation, or any other characteristic protected by local law, regulation, or ordinance.Position of TrustN/A
Work Model for this Role
This role will be eligible for our hybrid work model which allows employees to split their time between working on-site at their assigned Intel site and off-site. * Job posting details (such as work model, location or time type) are subject to change.*
Skills Required
- 5+ years AI/Systems/HPC software development in C++; strong concurrency and lock-free design.
- Hands-on experience with collective libraries (NCCL, MPI) or scale-out communication.
- Working knowledge of communication patterns for tensor, pipeline, and expert parallelism and mapping to collectives/fabric.
- Performance engineering experience: profiling, bandwidth/latency tuning, and roofline reasoning.
- Topology/placement algorithms, congestion control, and multi-tenant fabric experience.
- Large-scale distributed training/inference exposure.
- Experience with large-scale inference serving stacks (vLLM, SGLang, TensorRT-LLM, DeepSeek).
- RDMA/InfiniBand/RoCE, GPUDirect-style transfers, or comparable fabric experience.
Intel Compensation & Benefits Highlights
The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Intel and has not been reviewed or approved by Intel.
-
Leave & Time Off Breadth — Sabbaticals and paid time off are highlighted as signature elements, with an established program offering four weeks after four years or eight weeks after seven years. This distinctive time off is positioned as a meaningful part of the overall package.
-
Parental & Family Support — Paid bonding leave of 12 weeks and a New Parent Reintegration program, plus fertility benefits around $40,000 and up to $15,000 adoption reimbursement with no lifetime cap, are clearly stated. These programs are presented as standout components alongside broader family support resources.
-
Healthcare Strength — Multiple medical plan options with 2026 updates, a shift to Spring Health for EAP, and in‑network virtual medical visits covered at 100% beginning in 2026 indicate a comprehensive offering. These features signal an emphasis on robust medical access and mental health support.
Intel Insights
What We Do
Our mission is to shape the future of technology to help create a better future for the entire world, that’s the power of Intel Inside. With more ingenuity and creativity inside, our work is at the heart of countless innovations. From major breakthroughs to things that make everyday life better— they’re all powered by Intel technology. With a career at Intel, you can help make the future more wonderful for everyone.






