Platform Engineer (Infrastructure & DevOps)

Posted 4 Days Ago
Be an Early Applicant
Hiring Remotely in Osaka, JPN
Remote or Hybrid
6M-10M Annually
Mid level
Artificial Intelligence • Computer Vision • Machine Learning • Software
The Role
Design, deploy, and operate infrastructure for AI and software products. Own CI/CD pipelines, Kubernetes and containerized services, backend and worker systems, monitoring, secrets, networking, backups, and recovery. Support hybrid cloud and on-premise environments, ML workloads, and developer tooling. Collaborate with engineering teams to improve reliability, deployment, debugging, documentation, and scalability.
Summary Generated by Built In
Platform Engineer (Infrastructure & DevOps)Company Description

Rokken is a software development company that provides customized technical solutions for clients. Our work combines machine learning, real-time 3D visualization, medical software, and AI-powered web applications.

Website: https://rokken.tech/

Main office:
S-Cube108, 130-42 Nagasonecho, Kita Ward, Sakai, Osaka 591-8025, Japan

Job Offer Description

We are looking for a Platform Engineer to help design, deploy, and operate the infrastructure behind our AI and software products.

This role sits between infrastructure, DevOps, backend systems, and machine learning operations. You will help make our development and deployment environments reliable: CI/CD, worker machines, backend services, monitoring, secrets, backups, and developer workflows.

We are a small engineering team, so this is a high-ownership role. We need someone who can build practical systems, document them clearly, and work closely with ML engineers, backend developers, and project leads.

Annual Compensation

6,000,000 to 9,500,000 yen

Language Requirements
  • English: Business level required

  • Japanese: Not required, but nice to have

Minimum Experience

Mid-level or above.

This is an ownership-heavy role, but we are not necessarily looking for a senior specialist. We care more about solid fundamentals, good judgment, and the ability to learn and take responsibility for systems over time.

In This Role, You Will
  • Design, deploy, and maintain infrastructure for AI and software products.

  • Own CI/CD pipelines, self-hosted runners, deployment workflows, and release reliability.

  • Operate backend services, worker services, databases, and supporting infrastructure.

  • Improve observability through logs, metrics, dashboards, alerts, and incident investigation.

  • Manage secrets, service accounts, network access, VPNs, and security-sensitive configuration.

  • Support hybrid infrastructure: cloud or VPS services plus local or on-premise machines when needed.

  • Help manage computational resources for ML training, batch processing, and AI workflows.

  • Build and maintain backup, recovery, and maintenance procedures.

  • Collaborate with ML and software engineers to make systems easier to deploy, debug, and scale.

  • Document infrastructure decisions, operational procedures, and best practices.

Minimum Qualifications
  • 3+ years of experience in infrastructure, platform engineering, DevOps, SRE, or a similar role.

  • Strong Linux administration skills.

  • Practical experience with Docker and containerized services.

  • Ability to work with Kubernetes-based systems.

  • Experience with CI/CD systems and Git-based development workflows.

  • Ability to write scripts or small tools, preferably in Python or shell.

  • Understanding of networking, security, secrets management, and service configuration.

  • Experience operating production or production-like services.

  • Strong problem-solving, documentation, and communication skills.

Preferred Qualifications
  • Deeper experience with Kubernetes, Helm, GitOps, or other container orchestration/deployment workflows.

  • Experience with GPU servers, ML workloads, or distributed computing.

  • Experience with Python backend services, PostgreSQL, Redis, or similar systems.

  • Familiarity with workflow orchestration systems.

  • Experience with monitoring and logging stacks such as Grafana, Prometheus, Loki, Tempo, or similar tools.

  • Experience with self-hosted GitHub runners or internal developer tooling.

  • Experience with distributed or object storage such as Ceph, MinIO, JuiceFS, S3-compatible storage, or NAS systems.

  • Experience with VPN or private-network tools.

  • Experience with hybrid cloud/on-premise deployments.

  • Backend development experience, especially around ML or data-processing workloads.

Working Style

This is not a pure IT support position and not a pure cloud administrator position. We are looking for someone who enjoys building practical infrastructure for real engineering teams: deploying services, making worker systems reliable, improving CI, debugging failures, and helping developers move faster without making systems fragile.

You do not need to be an expert in every tool listed above. We care more about strong fundamentals, good judgment, clear communication, and the ability to own systems over time.

We are an AI-forward engineering team, and we especially value people who already use modern AI agents in their daily development and operations work.

We currently use Kubernetes internally. We expect the person in this role to be comfortable working with it, while also being able to recommend better approaches when they are genuinely more appropriate for our scale and needs.

Support For Foreign Candidates
  • Assistance with visa renewal or visa application.

  • Help with finding housing.

  • Support for daily life in Japan, such as city office registration and opening a bank account.

Trial Period

3 months.

Remote Work Policy

Because this role may involve local servers, GPU machines, and office infrastructure, regular on-site presence near our Osaka/Sakai office is expected.

  • Initial phase: on-site work preferred.

  • After initial assessment: flexible work arrangements may be possible, with regular on-site presence when infrastructure work requires it.

Skills Required

  • 3+ years of experience in infrastructure, platform engineering, DevOps, SRE, or a similar role
  • Strong Linux administration skills
  • Practical experience with Docker and containerized services
  • Ability to work with Kubernetes-based systems
  • Experience with CI/CD systems and Git-based development workflows
  • Ability to write scripts or small tools, preferably in Python or shell
  • Understanding of networking, security, secrets management, and service configuration
  • Experience operating production or production-like services
  • Strong problem-solving, documentation, and communication skills
  • Business-level English
  • Deeper experience with Kubernetes, Helm, GitOps, or other container orchestration and deployment workflows
  • Experience with GPU servers, ML workloads, or distributed computing
  • Experience with Python backend services, PostgreSQL, Redis, or similar systems
  • Familiarity with workflow orchestration systems
  • Experience with monitoring and logging stacks such as Grafana, Prometheus, Loki, or Tempo
  • Experience with self-hosted GitHub runners or internal developer tooling
  • Experience with distributed or object storage such as Ceph, MinIO, JuiceFS, S3-compatible storage, or NAS systems
  • Experience with VPN or private-network tools
  • Experience with hybrid cloud and on-premise deployments
  • Backend development experience, especially around ML or data-processing workloads
  • Japanese language ability
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
Year Founded: 2016

What We Do

Rokken is a Japan-based software development company that creates customized technical solutions, particularly for medical-device manufacturers and related clients. Its capabilities include medical-device software, artificial intelligence and machine-learning technologies, real-time 3D/4D rendering, and AI-powered web applications. The company also provides consulting on medical-device regulatory compliance, combining engineering expertise with domain knowledge to deliver quality projects on schedule and within budget.

Similar Jobs

Accuris Logo Accuris

Account Manager

Information Technology • Machine Learning • Software • Conversational AI • Generative AI • Manufacturing
Remote
Japan
1000 Employees

Dropbox Logo Dropbox

Business Development Representative

Artificial Intelligence • Cloud • Consumer Web • Productivity • Software • App development • Data Privacy
Remote
Japan
2500 Employees

UL Solutions Logo UL Solutions

Senior Business Manager

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
日本
15000 Employees

UL Solutions Logo UL Solutions

Field Engineer

Automotive • Professional Services • Software • Consulting • Energy • Chemical • Renewable Energy
Remote or Hybrid
Japan
15000 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account