Rokken is a software development company that provides customized technical solutions for clients. Our work combines machine learning, real-time 3D visualization, medical software, and AI-powered web applications.
Website: https://rokken.tech/
Main office:
S-Cube108, 130-42 Nagasonecho, Kita Ward, Sakai, Osaka 591-8025, Japan
We are looking for a Platform Engineer to help design, deploy, and operate the infrastructure behind our AI and software products.
This role sits between infrastructure, DevOps, backend systems, and machine learning operations. You will help make our development and deployment environments reliable: CI/CD, worker machines, backend services, monitoring, secrets, backups, and developer workflows.
We are a small engineering team, so this is a high-ownership role. We need someone who can build practical systems, document them clearly, and work closely with ML engineers, backend developers, and project leads.
Annual Compensation6,000,000 to 9,500,000 yen
Language RequirementsEnglish: Business level required
Japanese: Not required, but nice to have
Mid-level or above.
This is an ownership-heavy role, but we are not necessarily looking for a senior specialist. We care more about solid fundamentals, good judgment, and the ability to learn and take responsibility for systems over time.
In This Role, You WillDesign, deploy, and maintain infrastructure for AI and software products.
Own CI/CD pipelines, self-hosted runners, deployment workflows, and release reliability.
Operate backend services, worker services, databases, and supporting infrastructure.
Improve observability through logs, metrics, dashboards, alerts, and incident investigation.
Manage secrets, service accounts, network access, VPNs, and security-sensitive configuration.
Support hybrid infrastructure: cloud or VPS services plus local or on-premise machines when needed.
Help manage computational resources for ML training, batch processing, and AI workflows.
Build and maintain backup, recovery, and maintenance procedures.
Collaborate with ML and software engineers to make systems easier to deploy, debug, and scale.
Document infrastructure decisions, operational procedures, and best practices.
3+ years of experience in infrastructure, platform engineering, DevOps, SRE, or a similar role.
Strong Linux administration skills.
Practical experience with Docker and containerized services.
Ability to work with Kubernetes-based systems.
Experience with CI/CD systems and Git-based development workflows.
Ability to write scripts or small tools, preferably in Python or shell.
Understanding of networking, security, secrets management, and service configuration.
Experience operating production or production-like services.
Strong problem-solving, documentation, and communication skills.
Deeper experience with Kubernetes, Helm, GitOps, or other container orchestration/deployment workflows.
Experience with GPU servers, ML workloads, or distributed computing.
Experience with Python backend services, PostgreSQL, Redis, or similar systems.
Familiarity with workflow orchestration systems.
Experience with monitoring and logging stacks such as Grafana, Prometheus, Loki, Tempo, or similar tools.
Experience with self-hosted GitHub runners or internal developer tooling.
Experience with distributed or object storage such as Ceph, MinIO, JuiceFS, S3-compatible storage, or NAS systems.
Experience with VPN or private-network tools.
Experience with hybrid cloud/on-premise deployments.
Backend development experience, especially around ML or data-processing workloads.
This is not a pure IT support position and not a pure cloud administrator position. We are looking for someone who enjoys building practical infrastructure for real engineering teams: deploying services, making worker systems reliable, improving CI, debugging failures, and helping developers move faster without making systems fragile.
You do not need to be an expert in every tool listed above. We care more about strong fundamentals, good judgment, clear communication, and the ability to own systems over time.
We are an AI-forward engineering team, and we especially value people who already use modern AI agents in their daily development and operations work.
We currently use Kubernetes internally. We expect the person in this role to be comfortable working with it, while also being able to recommend better approaches when they are genuinely more appropriate for our scale and needs.
Support For Foreign CandidatesAssistance with visa renewal or visa application.
Help with finding housing.
Support for daily life in Japan, such as city office registration and opening a bank account.
3 months.
Remote Work PolicyBecause this role may involve local servers, GPU machines, and office infrastructure, regular on-site presence near our Osaka/Sakai office is expected.
Initial phase: on-site work preferred.
After initial assessment: flexible work arrangements may be possible, with regular on-site presence when infrastructure work requires it.
Skills Required
- 3+ years of experience in infrastructure, platform engineering, DevOps, SRE, or a similar role
- Strong Linux administration skills
- Practical experience with Docker and containerized services
- Ability to work with Kubernetes-based systems
- Experience with CI/CD systems and Git-based development workflows
- Ability to write scripts or small tools, preferably in Python or shell
- Understanding of networking, security, secrets management, and service configuration
- Experience operating production or production-like services
- Strong problem-solving, documentation, and communication skills
- Business-level English
- Deeper experience with Kubernetes, Helm, GitOps, or other container orchestration and deployment workflows
- Experience with GPU servers, ML workloads, or distributed computing
- Experience with Python backend services, PostgreSQL, Redis, or similar systems
- Familiarity with workflow orchestration systems
- Experience with monitoring and logging stacks such as Grafana, Prometheus, Loki, or Tempo
- Experience with self-hosted GitHub runners or internal developer tooling
- Experience with distributed or object storage such as Ceph, MinIO, JuiceFS, S3-compatible storage, or NAS systems
- Experience with VPN or private-network tools
- Experience with hybrid cloud and on-premise deployments
- Backend development experience, especially around ML or data-processing workloads
- Japanese language ability
What We Do
Rokken is a Japan-based software development company that creates customized technical solutions, particularly for medical-device manufacturers and related clients. Its capabilities include medical-device software, artificial intelligence and machine-learning technologies, real-time 3D/4D rendering, and AI-powered web applications. The company also provides consulting on medical-device regulatory compliance, combining engineering expertise with domain knowledge to deliver quality projects on schedule and within budget.








