AI Engineer, Data Infra

Reposted 16 Hours Ago
Be an Early Applicant
Singapore, SGP
In-Office
Mid level
Artificial Intelligence • Healthtech • Information Technology • Biotech
The Role
Design, build, and scale data infrastructure for LLM training, fine-tuning, evaluation, and RAG. Implement and maintain data architectures, pipelines (batch and streaming), org-level IAM and governance, cost/performance optimizations, analytics/dashboarding, AI-assisted ops and internal automation tools, and support multi-cloud production workloads.
Summary Generated by Built In

AI Singapore (AISG) is a national AI programme launched by the National Research Foundation (NRF), Singapore, to build and anchor deep national capabilities in AI.  AISG is supported through a government-wide partnership including the NRF, Ministry of Digital Development and Information (MDDI), Infocomm Media Development Authority (IMDA), Economic Development Board (EDB) and Enterprise Singapore (ESG). We bring together research institutions and the vibrant ecosystem of AI start-ups and companies to support impactful research, develop talent, and power Singapore's AI efforts.

This position will be hosted at Nanyang Technological University (NTU) under VP (Artificial Intelligence & Digital Economy)’s office and we welcome you to join our community.

We're looking for an AI Engineer to join the Platform team within AI Products at AISG. In this role, you will design, build, and scale the data infrastructure that powers large language model (LLM) training, fine-tuning, evaluation, and retrieval-augmented generation (RAG) across the organisation. Your work will directly contribute to the architecture and reliability of data storage systems, data pipelines, and the compute infrastructure that supports large-scale AI workloads.

Responsibilities:

Data infrastructure and architecture

  • Develop and maintain the overall data architecture, ensuring it scales to support AI training, fine-tuning, evaluation, and RAG workloads.

  • Define and manage the data technology stack, evaluating and adopting tools that best fit evolving data needs.

  • Build automation for data transfer and backup, and support data cataloging/discovery tools.

  • Design and maintain data pipelines (batch and streaming) using tools like AWS Glue and Amazon EMR.

Data governance and management

  • Architect and govern org-level IAM policy across AWS Organizations including SCPs, cross-account roles, and permission boundaries to enforce least-privilege access at scale.

  • Review and enhance data access patterns for performance and cost optimization.

  • Develop and manage data analytics and dashboarding capabilities using tools like Amazon Athena and QuickSight to give stakeholders visibility into platform usage and cost.

AI-assisted ops and continuous improvement

  • Use AI tools (e.g. Claude, Copilot, Cursor) appropriately in your daily work responsibilities.

  • Build internal tools leveraging AI to reduce manual effort in day-to-day operations.

Requirements:

You should be a hands-on engineer who enjoys both building robust infrastructure and designing tools that other engineers and researchers want to use. You should be comfortable balancing platform ownership (architecture, cost, governance) with product thinking (usability, self-service, adoption).

  • A degree in Computer Science, Information Technology, or equivalent.

  • At least 2–4 years of data infrastructure, platform or systems engineering experience, with a track record of operating production systems at scale.

  • Strong knowledge of distributed data systems and storage technologies (object storage, data lakes, distributed file systems, vector databases) and data pipelining tools (e.g. Apache Spark, Apache Airflow, Ray, Dagster).

  • Working knowledge of data access control and data orchestration.

  • Familiarity with cloud organisational structures (e.g. AWS Organizations, GCP folder/project hierarchy), including multi-account/multi-project setups, org-level IAM, and billing/cost allocation.

  • Hands-on experience operating workloads on different cloud providers including IaC (e.g. Terraform), containers and orchestration (e.g. Docker, Kubernetes), and managed services for compute, storage, and networking.

  • Demonstrated use of AI tools (e.g. Claude, Copilot, Cursor) in your day-to-day engineering — for code generation, review, debugging, and documentation — with a clear sense of where they help and where they don't.

  • Solid scripting/programming skills (e.g. Python, SQL) and comfortable reading other people's code across the stack.

  • Strong communication skills and a team collaborator.

Good to Have:

  • Experience with LLM training dataset (e.g. Common Crawl).

  • C/C++/Rust/Go or other relevant programming languages.

  • Contributions to open-source AI/ML projects.

We regret that only shortlisted candidates will be notified.

Hiring Institution: NTU

Skills Required

  • Degree in Computer Science, Information Technology, or equivalent.
  • 2-4 years of data infrastructure, platform or systems engineering experience operating production systems at scale.
  • Strong knowledge of distributed data systems and storage technologies (object storage, data lakes, distributed file systems, vector databases).
  • Experience with data pipelining tools (Apache Spark, Apache Airflow, Ray, Dagster).
  • Working knowledge of data access control and data orchestration.
  • Familiarity with cloud organizational structures (AWS Organizations, GCP folders/projects), org-level IAM, multi-account/project setups, and billing/cost allocation.
  • Hands-on experience operating workloads on cloud providers including IaC (Terraform), containers and orchestration (Docker, Kubernetes), and managed services for compute, storage, and networking.
  • Demonstrated use of AI tools (e.g., Claude, Copilot, Cursor) in engineering work.
  • Solid scripting/programming skills (Python, SQL) and ability to read code across the stack.
  • Strong communication skills and ability to collaborate in teams.
  • Experience with LLM training datasets (e.g., Common Crawl).
  • Experience with C, C++, Rust, or Go.
  • Contributions to open-source AI/ML projects.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Singapore
10 Employees
Year Founded: 2020

What We Do

The Lee Kong Chian School of Medicine (LKCMedicine) trains doctors with a focus on patient-centered care, integrating precision medicine, Artificial Intelligence (AI) in healthcare, and medical humanities into its undergraduate medical degree program.

Similar Jobs

Airwallex Logo Airwallex

Senior Devops Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
Singapore, SGP
2300 Employees

Pfizer Logo Pfizer

Senior Health Representative (Vaccines)

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Hybrid
Singapore, SGP
121990 Employees

UL Solutions Logo UL Solutions

Senior Learning & Development Specialist

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Hybrid
Singapore, SGP
15000 Employees

Airwallex Logo Airwallex

Senior Software Engineer

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
In-Office
Singapore, SGP
2300 Employees

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account