About the AI Division:
The AI Division is a unique group within Ceva, driving innovation in Machine Learning and Generative AI architectures for edge and cloud inference.
Our R&D spans Neural Network Processors (NPUs), Vision DSPs, and advanced AI solutions.
About the Role:
You will be a key contributor to Ceva’s AI Graph Compiler software stack for NPUs, designing system-level execution flows for advanced neural networks, including LLMs.
The role focuses on L1/L2 memory management, data movement, and performance optimization, working closely with compiler and hardware architects.
What will you do:
· Design and own key components of Ceva’s AI Graph Compiler.
· Develop and optimize neural network execution flows and memory management.
· Enable complex AI workloads, including LLMs, and implement new NPU features.
· Analyze performance bottlenecks and drive system-level optimizations.
· Collaborate with compiler and hardware teams on HW–SW solutions.
RequirementsRequirements:
· 5 years of software development experience using C/C++.
· BSc/MSc in Computer Science, Electrical Engineering, or equivalent.
· Experience designing and developing complex software systems.
· Strong understanding of memory management and performance optimization.
· Strong problem-solving skills and technical ownership.
Advantages:
· Experience with AI accelerators, NPUs, GPUs, or DSPs.
· Experience with AI compilers, graph optimization, or neural network execution.
· Experience with LLMs or other large neural network workloads.
· Familiarity with PyTorch, TensorFlow, ONNX, or Python.
Skills Required
- At least 5 years of software development experience using C or C++
- Bachelor’s or master’s degree in Computer Science, Electrical Engineering, or equivalent
- Experience designing and developing complex software systems
- Strong understanding of memory management and performance optimization
- Strong problem-solving skills and technical ownership
- Experience with AI accelerators, NPUs, GPUs, or DSPs
- Experience with AI compilers, graph optimization, or neural network execution
- Experience with LLMs or other large neural network workloads
- Familiarity with PyTorch, TensorFlow, ONNX, or Python
What We Do
Ceva powers the Smart Edge, bridging the digital and physical worlds to bring AI-driven products to life. Our Ceva AI fabric portfolio of silicon and software IP enables devices to Connect, Sense, and Infer – the essential capabilities for the intelligent edge. From 5G, cellular IoT, Bluetooth, Wi-Fi, and UWB connectivity to scalable Edge AI NPUs, AI DSPs, sensor fusion processors and embedded software, Ceva provides the foundational IP for devices that connect, understand their environment, and act in real time. With more than 20 billion devices shipped and trusted by 400+ customers worldwide, Ceva is the backbone of today’s most advanced smart edge products – from AI-infused wearables and IoT devices to autonomous vehicles and 5G infrastructure. Our differentiated solutions deliver seamless integration into existing design flows, total flexibility to combine solutions based on design needs and ultra–low–power performance in minimal silicon footprint, helping customers accelerate development, reduce risk, and bring innovative products to market faster. As technology evolves toward Physical AI, Ceva’s IP portfolio lays the foundation for systems that are always connected, contextually aware, and capable of intelligent, real-time decision-making.







