Top Tech Jobs & Startup Jobs

3 Days AgoSaved
In-Office
Melbourne, Victoria, AUS
Mid level
Mid level
Artificial Intelligence • Hardware • Machine Learning • Software
Build and optimize an on-device AI inference stack, including custom inference engines, GPU kernels, serving runtimes, model routing, continuous batching, and decoding strategies. Develop C++ systems for low-latency inference across Apple, NVIDIA, AMD, Snapdragon, and other platforms. Profile performance, optimize memory and compute bottlenecks, and translate ML research into production infrastructure.
Top Skills: Amd GpusApple SiliconCC++CudaFp4Fp8Int8MetalNvidia GpusRocmRustSnapdragonTorch.CompileTriton
3 Days AgoSaved
In-Office
Berlin, DEU
Expert/Leader
Expert/Leader
Artificial Intelligence • Hardware • Machine Learning • Software
Lead frontier research in on-device AI inference, model routing, autonomous research pipelines, and real-world evaluations. Design and execute experiments involving speculative decoding, quantization, distillation, reinforcement learning, and other efficiency techniques. Own the research agenda, influence technical direction, document reproducible findings, and translate novel insights into practical performance improvements under device constraints.
Top Skills: Accelerator ArchitecturesAi InferenceArtificial IntelligenceCudaGpu ArchitecturesKernel OptimizationLarge Language ModelsMachine LearningMetalModel DistillationQuantizationReinforcement LearningRocmSpeculative DecodingTriton
3 Days AgoSaved
In-Office
Melbourne, Victoria, AUS
Expert/Leader
Expert/Leader
Artificial Intelligence • Hardware • Machine Learning • Software
Lead frontier research in on-device AI, focusing on inference efficiency, model routing, autonomous research pipelines, and real-world evaluations. The role owns significant parts of the research agenda and requires designing experiments, deriving actionable insights, and translating novel ideas into performance improvements. Work may include speculative decoding, quantization, distillation, reinforcement learning, accelerator optimization, and deployment under strict device constraints.
Top Skills: Accelerator ArchitecturesAi InferenceCudaGpu ArchitecturesLarge Language ModelsMachine LearningMetalModel DistillationQuantizationReinforcement LearningRocmSpeculative DecodingTriton
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account