Top Machine Learning Engineer Jobs

8 Days AgoSaved
Remote
2 Locations
Mid level
Mid level
Artificial Intelligence • Hardware • Software • Semiconductor
Drive end-to-end ML model inference performance: build kernel- and system-level performance models, optimize kernel microcode and compiler algorithms, debug runtime performance on system and cluster, and develop tooling to visualize and analyze performance data from the Wafer Scale Engine and compute cluster.
Top Skills: C++Cerebras Wafer Scale EngineCompilersCpu/Gpu SimulatorsHpcKernel MicrocodePerformance Profiling ToolsPython
Reposted 8 Days AgoSaved
Remote
United States
Mid level
Mid level
Artificial Intelligence • HR Tech • Software • Generative AI
Design and build reinforcement-learning training environments and diverse tasks to evaluate and improve LLM agents; iterate rapidly on task designs from customer feedback, deliver high-quality outputs with minimal supervision, and maintain PST overlap for collaboration.
Top Skills: Large Language Models (Llms)Machine LearningPythonReinforcement Learning
8 Days AgoSaved
In-Office
Milpitas, CA, USA
136K-232K Annually
Senior level
136K-232K Annually
Senior level
Hardware
Develop and maintain high-performance C++ machine control and data analysis software for advanced mask inspection systems. Integrate AI/ML capabilities, collaborate with multidisciplinary engineering teams, optimize and scale existing code, and support integration/testing and customer escalations. Work with APIs, containers, and observability tools to enable reliable production deployment.
Top Skills: AIAngularC++DockerGrafanaGtkKubernetesLinuxMachine LearningPostgresPrometheusQtReactRestRpcSingularityVue
Reposted 8 Days AgoSaved
Hybrid
Boston, MA, USA
170K-205K Annually
Senior level
170K-205K Annually
Senior level
Artificial Intelligence • Healthtech • Insurance • Software
Design, build, and deploy production ML systems that combine structured clinical/EHR data with computer vision outputs. Own model calibration, interpretability, monitoring, drift detection, retraining workflows, and regulatory documentation. Partner cross-functionally, mentor engineers, and translate clinical questions into reliable, production-grade ML solutions.
Top Skills: Ci/CdComputer VisionDrift DetectionEhrLongitudinal ModelingMachine Learning PipelinesModel MonitoringMultimodal ModelsPythonStructured Clinical DataSurvival AnalysisTestingTime-Series AnalysisTraining And Inference Infrastructure
Reposted 8 Days AgoSaved
In-Office
San Mateo, CA, USA
Senior level
Senior level
News + Entertainment
Design, build, deploy, and operate production-grade ML systems and pipelines for a massive consumer audience. Own full lifecycle from data processing and feature engineering through training, validation, deployment, and monitoring. Drive ML infrastructure and architecture decisions, partner with product and data teams, optimize models against business metrics, and mentor engineers to ensure security, reliability, and compliance.
Top Skills: Automated Model Lifecycle ManagementCloud-Native Ml ArchitecturesDistributed TrainingMlops PlatformsModel Monitoring/ObservabilityPython
9 Days AgoSaved
In-Office
Los Angeles, CA, USA
170K-300K Annually
Senior level
170K-300K Annually
Senior level
Aerospace • Hardware • Software • Defense • Manufacturing
Build and operate Hadrian's production ML platform: standardize MLflow/Dagster-based deployments, serve batch and online inference, maintain feature serving and lineage, detect drift, enable model CI/CD, observability, autoscaling, and developer tooling for reliable model production across factories.
Top Skills: Amazon EcrAmazon EksBentomlContainersDagsterFastapiFeastGoGpu InferenceGrpcKserveKubernetesMlflowPythonRay ServeRustSagemakerSQLTectonTritonVertex Ai
Reposted 9 Days AgoSaved
In-Office
San Francisco, CA, USA
Mid level
Mid level
Artificial Intelligence • Software
The Research Engineer, Infrastructure will build distributed training systems, optimize performance, manage data pipelines, and enhance research workflows, ensuring the infrastructure scales with AI advancements.
Top Skills: C++GpuPythonPyTorch
9 Days AgoSaved
In-Office or Remote
2 Locations
121K-219K Annually
Senior level
121K-219K Annually
Senior level
Cloud • Security • Software • Cybersecurity
Build and operate model validation, quantization, and safety systems for production ML. Develop pipelines for security scanning, optimization (quantization/pruning), routing, prompt management, and evaluation frameworks measuring accuracy, performance, and safety across model lifecycles.
Top Skills: AwqCi/CdCloud InfrastructureContainerizationGgufGptqJaxLlmsPythonPyTorchTensorFlowTransformers
Reposted 9 Days AgoSaved
Remote
Georgia, USA
90K-170K Annually
Junior
90K-170K Annually
Junior
Retail
Develop, deploy, and support secure, scalable production software and AI/ML solutions. Collaborate with product, UX, and engineering to deliver features, create tests and automation, instrument monitoring and dashboards, and participate in agile processes while continuously learning and improving team practices.
Top Skills: APIsAWSAzureAzure AiCi/CdCSSDevOpsGCPGitHTMLJavaJavaScriptLangchainMicroservicesMlopsNoSQLOpenaiPrompt EngineeringPythonPyTorchRelational DatabasesScikit-LearnSQLTensorFlowTypescript
9 Days AgoSaved
In-Office
Chantilly, VA, USA
Expert/Leader
Expert/Leader
Information Technology • Software • Analytics • Cybersecurity
Design, build, and maintain scalable MLOps pipelines and data workflows to train, fine-tune, deploy, and monitor multilingual text and speech translation models. Prepare and optimize multilingual datasets, support model evaluation and benchmarking, and enable low-resource language training while collaborating with data scientists to integrate models into production.
Top Skills: BedrockDockerEc2IamLambdaPythonS3SagemakerVpc
9 Days AgoSaved
Remote
United States
230K-270K Annually
Senior level
230K-270K Annually
Senior level
Software
Design, build, and operate core AI capabilities and reusable model integrations for an enterprise platform. Integrate foundation models into production, optimize performance and reliability, prototype new techniques, and support scalable backend services and distributed AI systems in collaboration with cross-functional teams.
Top Skills: Ai AgentsAPIsCloud-Native InfrastructureDistributed SystemsFoundation ModelsKubernetesLarge Language ModelsModel OrchestrationMultimodal AiObservabilityPythonRetrieval-Augmented Generation (Rag)Rust
9 Days AgoSaved
In-Office
San Jose, CA, USA
170K-244K Annually
Senior level
170K-244K Annually
Senior level
Fintech • Payments
Design, deploy, and productize scalable traditional and generative AI/ML solutions. Collaborate with data scientists and engineers, build LLM-based agents, ensure code quality and reliable CI/CD deployments, mentor team members, and drive adoption of state-of-the-art ML techniques.
Top Skills: Agent DevelopmentSparkAWSAzureCi/CdDockerGCPJavaKubernetesLarge Language ModelsMapreducePython
New

Track Smarter, Apply Better.

Ditch the spreadsheets. Organize your job search with our freeApplication Tracker.

Use For Free
Application Tracker Preview
9 Days AgoSaved
In-Office
San Diego, CA, USA
118K-162K Annually
Entry level
118K-162K Annually
Entry level
Automotive • Internet of Things • Mobile • Semiconductor • Industrial
Develop and prototype AI/ML-driven automation for ASIC/SoC design and implementation. Collaborate with academia and internal teams to apply GenAI, RL, GNNs, and CNNs to VLSI CAD flows, pilot solutions on real designs, and improve power, performance, area, and quality.
Top Skills: BitbucketC++Ci/CdConvolutional Neural NetworksDesignsyncEda ToolsGenaiGitGraph Neural NetworksJenkinsLlmsLsfMakePythonRecurrent Neural NetworksReinforcement LearningSplunkUnix/LinuxVlsi Cad
Reposted 9 Days AgoSaved
In-Office
Santa Clara, CA, USA
152K-288K Annually
Senior level
152K-288K Annually
Senior level
Artificial Intelligence • Computer Vision • Hardware • Robotics • Metaverse
As a Senior Software Engineer at NVIDIA, you will design and implement inference software optimizations for AI applications using TensorRT, collaborating with deep learning experts to influence hardware and software design.
Top Skills: C++CudaPython
Reposted 9 Days AgoSaved
In-Office or Remote
Evendale, OH, USA
112K-150K Annually
Mid level
112K-150K Annually
Mid level
Aerospace • Energy
This role involves designing, developing, and maintaining AI/ML products, collaborating with stakeholders, and ensuring model deployment on AWS. Responsibilities include creating APIs, establishing MLOps practices, and driving AI strategy for operational improvements.
Top Skills: AWSC#DatabricksFastapiFlaskJavaMlflowPythonTypescript
Reposted 9 Days AgoSaved
In-Office
Palo Alto, CA, USA
140K-280K Annually
Mid level
140K-280K Annually
Mid level
3D Printing • Consulting • Design • Manufacturing
Build scalable agentic frameworks and reproducible experimental pipelines that integrate frontier multimodal models with rich neuroscience data. Implement orchestration, MLOps, and cloud infrastructure, balancing research and production engineering on a small collaborative team.
Top Skills: Agent Orchestration Frameworks (Claude Code Agent TeamsAgentic FrameworksAzure)BeadsClaude Computer UseCodexCodex Multi-AgentsCrew Ai Agent Teams)Fdm-1GCPGemini-CliManusMicrosoft Agent FrameworkMlflowMultimodal Large Language ModelsOai OperatorOpenspecOrchestration Platforms (AwsPrompt EngineeringPythonPyTorchQwen Code)TransformersTui/Cli Coding Agents (Claude CodeWeights & Biases
Reposted 9 Days AgoSaved
In-Office
Palo Alto, CA, USA
200K-280K Annually
Mid level
200K-280K Annually
Mid level
3D Printing • Consulting • Design • Manufacturing
Optimize training and inference performance of foundation models by writing and tuning CUDA/Triton GPU kernels, profiling and removing bottlenecks, implementing low-precision and mixed-precision strategies, optimizing MoE routing and expert dispatch, integrating high-performance libraries, and collaborating closely with researchers to translate model ideas into scalable, efficient implementations.
Top Skills: CudaCudnnCutlassFlashattentionGpu KernelsKv-Cache ManagementLow-Precision TrainingMixed-Precision TrainingMoe (Mixture Of Experts)Moe RoutingNcclNsight ComputeNsight SystemsPost-Training QuantizationPyTorchPytorch ProfilerQuackSpeculative DecodingTensorrt-LlmTransformer ArchitecturesTritonVllm
Reposted 9 Days AgoSaved
In-Office
Palo Alto, CA, USA
175K-250K Annually
Mid level
175K-250K Annually
Mid level
3D Printing • Consulting • Design • Manufacturing
Design and build an end-to-end, high-throughput dataloading stack for massive multimodal datasets: formatting, preprocessing, filtering, sharding, caching, and streaming data to distributed GPU training with observability, reliability, and performance benchmarking.
Top Skills: AirflowC++CudaDagsterDockerGpuKubernetesMlflowPrefectPythonPyTorchRustW&B
Reposted 9 Days AgoSaved
In-Office
Palo Alto, CA, USA
200K-280K Annually
Senior level
200K-280K Annually
Senior level
3D Printing • Consulting • Design • Manufacturing
Design, build, and optimize distributed training systems for foundation models across thousands of GPUs. Implement advanced parallelism, fault-tolerance, profiling, and tooling to enable large-scale, reproducible ML experiments in cloud/HPC environments.
Top Skills: AutogradCloudCudaDeepspeed ZeroHpcJob OrchestrationMegatronMixed-Precision Training (Mxfp8/Nvfp4)NcclNvlinkNvswitchProfiling ToolsPytorch FsdpTorch.CompileTorch.DistributedTorchtitan
Reposted 9 Days AgoSaved
In-Office
Palo Alto, CA, USA
175K-250K Annually
Senior level
175K-250K Annually
Senior level
3D Printing • Consulting • Design • Manufacturing
Develop, scale, and productionize state-of-the-art computer vision foundation models for multimodal data pipelines. Move research code to production, optimize inference and memory, build reliable ML workflows, and collaborate closely with researchers and engineers on large-scale AI research.
Top Skills: AirflowDagsterDvcPythonPyTorchRay
Reposted 9 Days AgoSaved
In-Office
San Francisco, CA, USA
Mid level
Mid level
Artificial Intelligence • Logistics • Robotics • Transportation
Design, build, and scale ML infrastructure: ingest vehicle sensor data, create batch pipelines for dataset curation, run distributed GPU training, and ensure performance, observability, efficiency, and security across the ML pipeline while partnering with ML teams.
Top Skills: AnsibleCi/CdCloud InfrastructureCluster Scheduling SystemsCryptographyGpu ClustersLinuxNetwork SecurityTerraform
Reposted 9 Days AgoSaved
Remote
US
Senior level
Senior level
Artificial Intelligence • Healthtech • Software • Biotech
Lead design, development, and production deployment of ML models and GenAI/LLM solutions to optimize clinical trial workflows. Build scalable pipelines, monitoring, and CI/CD for robust production use. Collaborate cross-functionally, mentor junior engineers, and communicate technical results to technical and non-technical stakeholders.
Top Skills: Ci/CdCloud InfrastructureFine-TuningGenaiLarge Language Models (Llms)Ml PipelinesModel DeploymentModel MonitoringPythonSQL
Reposted 9 Days AgoSaved
In-Office
College Park, MD, USA
80K-110K Annually
Mid level
80K-110K Annually
Mid level
Edtech • Information Technology • Professional Services
The role involves developing end-to-end machine learning solutions, focusing on NLP and computer vision, collaborating with teams, and preparing technical briefings for stakeholders.
Top Skills: Ci/CdElasticsearchFastapiMachine LearningMongoDBPostgresPyTorchReactScikit-LearnScipy
Reposted 9 Days AgoSaved
In-Office
Redwood City, CA, USA
Senior level
Senior level
Artificial Intelligence • Software • Conversational AI • Generative AI
As an Applied ML Software Engineer, you will design and implement ML models, manage cloud infrastructure, and optimize backend systems for AI-driven content discovery and recommendations.
Top Skills: AWSAzureGCPGrpcPythonPyTorchRestfulTensorFlow
Reposted 9 Days AgoSaved
Remote
Texas, USA
Mid level
Mid level
Information Technology • Robotics
The ML Platform Engineer will build scalable architecture for ML training workloads, optimize system performance, and collaborate with teams for enhanced efficiency on Kubernetes.
Top Skills: Argo WorkflowsC++GoKubernetesMlflowPythonRay
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account