Latin America | 100% Remote
About the RoleWe're looking for a senior engineer who can design, build, and operate production systems end-to-end. You'll work across our core stack building features and services, while also owning the reliability, observability, and performance of what you ship. You won't be handing your code off to a separate ops team — you'll be on-call for it, instrumenting it, and improving it based on how it behaves in production.
What You'll Do
- Design and build backend/fullstack features across Java 21, Spring/Spring boot, JPA/Hibernate, NodeJS, React/React Native
- Instrument services with logging, metrics, and tracing (e.g., Splunk, Prometheus, Grafana, Datadog, OpenTelemetry) as a standard part of development, not an afterthought
- Define and monitor SLIs/SLOs for the services you own; use error budgets to guide prioritization between feature work and reliability work
- Participate in an on-call rotation; respond to and resolve production incidents affecting your services
- Write and maintain postmortems/RCAs, and drive follow-up fixes to prevent recurrence
- Contribute to infrastructure-as-code (Terraform, CloudFormation, etc.) for the systems you build
- Perform capacity planning and load testing for services ahead of scale events
- Collaborate with platform/infra teams on shared tooling, but take primary ownership of your service's health
- Participate in code reviews, architecture discussions, and mentor junior engineers
- 5+ years of professional software engineering experience, with deep expertise in Java 21, Spring/Spring boot, JPA/Hibernate, NodeJS, React/React Native
- Demonstrated experience owning services in production — not just writing code, but debugging, scaling, and maintaining it live
- Comfort with observability tooling and reading dashboards/logs/traces to diagnose issues under pressure
- Experience with incident response processes (on-call, paging, postmortems)
- Working knowledge of cloud infrastructure with AWS and containerization (Docker, Kubernetes) — enough to reason about deployment and scaling, even if you're not a dedicated platform engineer
- Strong communication skills — you can explain a production issue to both engineers and stakeholders
- Bonus: experience with CI/CD pipelines, infrastructure-as-code, or chaos engineering practices
- This is not a dedicated SRE/DevOps/Platform Engineering role. You won't be building the observability platform itself or managing infrastructure for other teams — you'll be a strong practitioner of reliability engineering within your own feature work.
100% Remote
Holidays off
Paid Time Off
Health insurance assistance
Competitive USD compensation
Growth opportunities
Skills Required
- 5+ years of professional software engineering experience
- Deep expertise in Java 21, Spring or Spring Boot, JPA or Hibernate, Node.js, React, and React Native
- Experience owning, debugging, scaling, and maintaining production services
- Experience with observability tooling and diagnosing issues using dashboards, logs, and traces
- Experience with incident response processes, including on-call, paging, and postmortems
- Working knowledge of AWS cloud infrastructure and Docker and Kubernetes containerization
- Strong communication skills with engineers and stakeholders
- Experience with CI/CD pipelines, infrastructure as code, or chaos engineering practices
What We Do
CodeRoad is a nearshore software development and technology-execution partner that helps organizations modernize applications, migrate to the cloud, build digital products, and implement enterprise AI. Its services span the full software development lifecycle, including dedicated engineering teams, platform engineering, data and AI infrastructure, legacy modernization, cloud architecture, quality automation, and DevOps. The company serves startups, enterprises, and private-equity portfolio businesses.







