Top Tech Jobs & Startup Jobs

Reposted 18 Days AgoSaved
Hybrid
Paris, Île-de-France, FRA
Mid level
Mid level
Gaming
Design and run experiments to measure how architecture decisions affect inference; create architecture variants optimized for inference speed; own post-training pipelines (fine-tuning, evaluation, adaptation); scale inference for large MoE models; publish research and build agent-driven research tooling.
Top Skills: Amd Mi300XDeepseek V4Delayed Tensor Parallelism (Dtp)Fine-TuningLadder ResidualLaneformerMixture Of Experts (Moe)Nvidia H100Nvidia H200Preference OptimizationPt-TransformerQuantizationQwen 3Transformers
Reposted 18 Days AgoSaved
Hybrid
Paris, Île-de-France, FRA
Senior level
Senior level
Gaming
Design, implement, and optimize extremely low-latency GPU kernels and a monokernel inference pipeline for LLMs across AMD and NVIDIA. Perform microarchitectural experiments, profile and instrument kernels, build in-kernel profiling, scale to large models (MoE), and develop autonomous AI agents for GPU engineering research.
Top Skills: Amd Mi300XCdna IsaCudaFp16HipNvidia H200PtxPyTorch
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account