Cloud Engineer II

Posted Yesterday
Be an Early Applicant
3 Locations
Remote
Mid level
Software
The Role
The Cloud Engineer II will join the SRE team to improve service mesh and Kubernetes architecture across AWS and Azure environments. Responsibilities include designing and maintaining reliability tools, monitoring systems, analyzing metrics, troubleshooting performance issues, improving availability, reducing operational toil, and optimizing cloud costs. The role requires hands-on experience with Kubernetes, Terraform, Helm, cloud deployments, container diagnostics, Python or Shell scripting, and mature development and deployment practices.
Summary Generated by Built In

Company overview:

TraceLink is the world’s largest Agentic Business Network, enabling life sciences and healthcare companies to build and manage a scalable digital workforce of governed, no-code AI agents that execute and coordinate mission-critical supply chain operations alongside human teams. Powered by the Integrate-Once™ OPUS platform, TraceLink links more than 300,000 network participants, enabling multi-enterprise processes at global scale.

Founded in 2009 with the simple mission of protecting patients, today Tracelink has 5 global offices, over 800 employees and more than 1700 customers in over 60 countries around the world. Our expanding product suite continues to protect patients and now also enhances multi-enterprise collaboration through innovative new applications such as MINT.

Tracelink is recognized as an industry leader by Gartner and IDC, and for having a great company culture by Comparably.

Role summary

We're looking for an experienced, driven and passionate engineering team member with backgrounds in programming, distributed systems and Kubernetes to help our SRE team improve its Service Mesh and Kubernetes architecture. The SRE group is building and expanding on the critical need to maintain visibility and provide scalability of the TraceLink global platform. Within SRE, you'll have plenty of opportunities to share your strengths, guide us on how to build a scalable platform and collaborate closely with various engineering stakeholders. 

You will work in a global team, in an inclusive environment with AWS cloud-based deployments and focus on ensuring services are running smoothly, continuously assess opportunities to reduce toil and help improve service availability and reliability, optimise AWS resources usage across multiple environments to deliver cost effective services to the engineering organisation.

Responsibilities

  • As a member of the SRE core team, ensure high availability, performance and reliability expected by our customers and delivery to defined OKRs

  • Design, build, document, test new tools and technologies as part of an Agile development team. Maintain and improve these to eliminate bugs, increase performance/efficiency, or extend capabilities

  • Play an active role in the development process, deliver on commitments, communicate issues, work with others both in the team and in other teams

  • Gather and analyse metrics from systems to assist in performance tuning and fault finding.

  • Motivated, self-organized and have good time & work management skills.

  • Implementing and monitoring systems to proactively detect and address issues

Qualifications

  • 3+ years of experience with increasing responsibility as an SRE/Cloud Engineer

  • Strong understanding of cloud deployment and management practices in Multi Cloud Environments (AWS and Azure)

  • Hands-on experience with Terraform Helm and Kubernetes

  • Hands-on experience with tools and techniques to diagnose and uncover container and overall system performance

  • Proactive approach to identifying problems, performance bottlenecks, and areas for improvement

  • Skilled in AWS services both from technology and cost optimisation perspectives

  • Multi Cloud skills with AWS and Azure are preferred.

  • Experience working with mature development practices and tools for source control, security, and deployment

  • Hands on experience with Python/Shell

  • Excellent communication skills, written and verbal

  • Strong analytical and problem-solving skills

  • Nice to have: Experience in Observability, Istio and Agentic AI capabilities. 

Please see the Tracelink Privacy Policy for more information on how Tracelink processes your personal information during the recruitment process and, if applicable based on your location, how you can exercise your privacy rights. If you have questions about this privacy notice or need to contact us in connection with your personal data, including any requests to exercise your legal rights referred to at the end of this notice, please contact [email protected].  


Skills Required

  • 3+ years of experience with increasing responsibility as an SRE or Cloud Engineer
  • Strong understanding of cloud deployment and management practices in multi-cloud environments, including AWS and Azure
  • Hands-on experience with Terraform, Helm, and Kubernetes
  • Hands-on experience diagnosing container and overall system performance
  • Experience with AWS services and cloud cost optimization
  • Experience with mature development practices and tools for source control, security, and deployment
  • Hands-on experience with Python or Shell
  • Excellent written and verbal communication skills
  • Strong analytical and problem-solving skills
  • Multi-cloud skills with AWS and Azure
  • Experience with observability, Istio, and agentic AI capabilities
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Wilmington, Massachusetts
942 Employees
Year Founded: 2009

What We Do

TraceLink is the only network creation platform company that builds integrated business ecosystems with multienterprise applications - the true foundation for digitalization - delivering customer-centric agility and resiliency for end-to-end supply networks and leveraging the collective intelligence of entire industries. Delivering end-to-end supply chain solutions, TraceLink's Opus Platform enables speed of innovation and implementation with an open partner model for no-code and low-code development of solutions and applications. At TraceLink, we blend decades of knowledge in SaaS technology and supply chain business processes with a clear vision for advancing manufacturing industries through disruptive, unconventional software solutions. With headquarters in Massachusetts, TraceLink has six global offices through North America, South America, Europe, and Asia.

Similar Jobs

Micron Technology Logo Micron Technology

IE AMHS 自動搬送設備管理・保全スタッフ(シフト勤務)

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Hiroshima, JPN
45000 Employees

Micron Technology Logo Micron Technology

F15 ADTJ PI Device Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Hiroshima, JPN
45000 Employees

Micron Technology Logo Micron Technology

Fab15 CMP Process & Equipment Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Hiroshima, JPN
45000 Employees

Micron Technology Logo Micron Technology

Process Control System (PCS) R2R Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
Remote
Hiroshima, JPN
45000 Employees

Similar Companies Hiring

Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account