Key Responsibilities
- Understand, analyze, profile, optimize, and provide guidance to the team on deep learning workloads on state-of-the-art hardware and software platforms to improve their efficiency with different levels of optimization
- Design and implement performance benchmarks and testing methodologies to evaluate application performance
- Build tools to automate workload analysis, workload optimization, and other critical workflows
- Triage system issues and identify bottleneck and inefficiencies by analyzing the sources of issues and the impact on hardware, network and propose solutions to enhance GPU utilization
- Support the team to develop appropriate kernels and systems for new model architectures and algorithms
- Participate in, or lead design reviews with peers and stakeholders to decide amongst available technologies.
- Review code developed by other developers and provide feedback to ensure best practices (e.g., style guidelines, checking code in, accuracy, testability, and efficiency).
- Contribute to existing documentation or educational content and adapt content based on product/program updates and user feedback.
- Represent MBZUAI at industry conferences and events, showcasing the institution’s cutting-edge HPC and deep learning capabilities and establishing MBZUAI as a global leader in AI research and innovation.
- Perform all other duties as reasonably directed by the line manager that are commensurate with these functional objectives.
Academic Qualifications
- Ph.D. in CS, EE or CSEE with 1+ years working experience, OR
- Masters in CS, EE or CSEE or equivalent experience with 2+ year working experience
Skills Required
- Ph.D. in CS, EE, or CSEE with 1+ years experience OR Master’s in CS, EE, or CSEE with 2+ years experience
- Strong background in parallel computing
- Hands-on experience in system-level coding and debugging methodologies
- Large-scale machine learning experience (training and inference stacks)
- Experience analyzing, profiling, and optimizing deep learning workloads on modern hardware and software platforms
- Ability to design and implement performance benchmarks and testing methodologies
- Experience building tools to automate workload analysis and optimization workflows
- Experience triaging system issues, identifying bottlenecks across hardware, network, and software to improve GPU utilization
- Experience supporting development of kernels and systems for new model architectures and algorithms
- Experience participating in or leading design reviews and performing code reviews to ensure best practices
- Ability to contribute to documentation and educational content
What We Do
First a passion, then an idea transformed into success – when it comes to pioneering automation and digitalisation technology, the ifm group is the ideal partner. Since its foundation in 1969, ifm has developed, produced and sold sensors, controllers, software and systems for industrial automation and for SAP-based solutions for supply chain management and shop floor integration worldwide. As one of the pioneers of Industry 4.0, ifm develops and implements consistent solutions to digitalise the entire value chain “from sensor to ERP”. Today, the second-generation family-run ifm group has more than 8,750 employees and is one of the worldwide market leaders. The group combines the internationality and innovative strength of a growing group of companies with the flexibility and close customer contact of a medium-sized company.









