ML Engineer

Posted 2 Days Ago
San Francisco, CA, USA
In-Office
165K-210K Annually
Senior level
Artificial Intelligence • Information Technology • Software
The Role
Architect, build, and optimize production multimodal generative AI systems across model serving, post-training, and agentic frameworks. Develop multi-agent orchestration, retrieval, memory, tool-use, evaluation, benchmarking, profiling, and MLOps capabilities. Integrate open-weight models into production runtimes, improve performance against hardware limits, contribute to open-source projects, and communicate technical innovations through documentation, blogs, and whitepapers.
Summary Generated by Built In

Sciforium is an AI infrastructure company developing next-generation multimodal AI models and a proprietary, high-efficiency serving platform. Backed by multi-million-dollar funding and direct sponsorship from AMD with hands-on support from AMD engineers the team is scaling rapidly to build the full stack powering frontier AI models and real-time applications.

About the role

As an ML Engineer at Sciforium, you will operate at the intersection of production software engineering and Core AI/ML to architect, scale, and optimize end-to-end multimodal GenAI systems. In this role, you will build production-grade solutions across Serving, Post-Training and Agentic frameworks. You will also be responsible for driving deep technical optimizations and MLOps process improvements.

What You’ll Be Doing

  • Build and scale Agentic AI Systems:
    Design and implement intelligent systems that can reason, plan, and execute complex multi-step workflows. Develop architectures that combine LLMs, retrieval systems, memory, tools, and feedback loops.Build orchestration frameworks for multi-agent and tool-based systems. Develop evaluation frameworks that measure accuracy, reliability, latency, and task completion.

  • New Model Enablements, Automated Benchmarking, Profiling & Roofline Analysis: Rapidly benchmark, adapt, and integrate state-of-the-art open-weights models into production runtimes. Build automated MLOps tooling to profile deep learning workloads against theoretical hardware limits to drive optimization.

  • Open-Source Leadership & Knowledge Sharing: Drive technical evangelism and elevate Sciforium’s presence in the global AI ecosystem through high-impact community engagement. Actively contribute code, features, and optimizations to high-visibility open-source repositories. Author and publish deep-dive technical blogs, whitepapers, and architecture breakdowns showcasing the novel innovations and complex problem-solving happening at Sciforium.

Must-Haves
  • Experience: 5+ years of professional ML/AI software engineering experience with a proven track record of architecting and shipping performance-critical systems. Proven experience maintaining and developing model libraries or reusable ML components.

  • Education: BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field (or equivalent practical experience).

  • ML Systems: Strong knowledge of generative AI systems including Large Language Models, Transformers, Reinforcement Learning, RAG, and agentic patterns such as Chain-of-Thought, Tool Use, and Multi-Agent orchestration

  • Machine Learning Expertise: Experience with one or more distributed ML training frameworks such as PyTorch, TensorFlow, or JAX, or Ray and inference engines like TensorRT, vLLM or SGLang. Good understanding of deep learning architectures across multiple domains (e.g., NLP, vision, speech, generative models).

  • Communication: Ability to articulate complex technical trade-offs, write clear documentation, and collaborate smoothly across multidisciplinary engineering teams.

Nice-to-have
  • Experience building production AI agents or autonomous systems. Experience with reasoning frameworks, planning systems, memory architectures, and tool-use ecosystems. Track record of reducing operational complexity while increasing scalability and maintainability.

  • Experience with vector databases, retrieval systems, knowledge graphs, or semantic search. Experience with AI evaluation, benchmarking, and observability platforms.

  • Familiarity with distributed serving or large-scale inference frameworks (e.g., vLLM, TensorRT, FasterTransformer).

  • Experience with model performance optimization and profiling.

  • Familiarity with low-level performance considerations when running models on GPUs/TPUs.

  • Contributions to open-source model repositories or ML frameworks.

Benefits include
  • Medical, dental, and vision insurance

  • 401k plan

  • Daily lunch, snacks, and beverages

  • Flexible time off

  • Competitive salary and equity

Equal opportunity

Sciforium is an equal opportunity employer. All applicants will be considered for employment without attention to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran or disability status.

Skills Required

  • 5+ years of professional ML/AI software engineering experience
  • Track record of architecting and shipping performance-critical systems
  • Experience maintaining and developing model libraries or reusable ML components
  • BS, MS, or PhD in Computer Science, Computer Engineering, or a related technical field, or equivalent practical experience
  • Strong knowledge of generative AI systems, including LLMs, Transformers, reinforcement learning, RAG, chain-of-thought, tool use, and multi-agent orchestration
  • Experience with at least one distributed ML training framework, such as PyTorch, TensorFlow, JAX, or Ray
  • Experience with inference engines such as TensorRT, vLLM, or SGLang
  • Understanding of deep learning architectures across multiple domains, including NLP, vision, speech, or generative models
  • Ability to articulate complex technical trade-offs, write clear documentation, and collaborate across multidisciplinary engineering teams
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
7 Employees
Year Founded: 2024

What We Do

Sciforium is pioneering the future of AI infrastructure and research. Backed by AMD and SignalFire, we're developing byte-native multimodal foundation models while delivering serverless LLM serving at a fraction of traditional costs.

Similar Jobs

Hybrid
4 Locations
289097 Employees

AMP Logo AMP

Machine Learning Engineer

Artificial Intelligence • Computer Vision • Greentech • Machine Learning • Robotics • Industrial • Automation
Easy Apply
Remote or Hybrid
United States
175 Employees
162K-170K Annually

ServiceNow Logo ServiceNow

Machine Learning Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Santa Clara, CA, USA
29000 Employees
176K-308K Annually

ServiceNow Logo ServiceNow

Senior Machine Learning Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Mountain View, CA, USA
29000 Employees
161K-274K Annually

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account