Site Reliability & Infrastructure Engineer

Reposted Yesterday
Be an Early Applicant
Budapest, HUN
Hybrid
Senior level
Edtech • Information Technology
The Role
Build and lead SRE practice for custom integrations: design observability, define SLOs/SLIs, architect for reliability, troubleshoot incidents, own cloud infrastructure, drive IaC and CI/CD automation, lead post-mortems, document runbooks and operational handoffs.
Summary Generated by Built In

At Instructure, we believe in the power of people to grow and succeed throughout their lives. Our goal is to amplify that power by creating intuitive products that simplify learning and personal development, facilitate meaningful relationships, and inspire people to go further in their education and careers.
We do this by giving smart, creative, passionate people opportunities to create awesome. And that's where you come in:

This is a foundational role for an experienced Site Reliability & Infrastructure Engineer to build and lead the reliability practice for our entire portfolio of custom integrations.

You will be the first SRE dedicated to the post-deployment success of our custom solutions, moving beyond project-based delivery to a model of continuous, proactive operational excellence.

This is a high-impact position to design and implement a modern SRE practice that integrates into our existing applications and workstreams, establishing a foundation for continued success of our custom development organization.

Key Responsibilities

1. Reliability & Performance Engineering:

  • Design & Implement Observability: Build and manage a centralized monitoring, logging, and alerting strategy for all custom integrations.

  • Define Service Level Objectives (SLOs): Work with stakeholders to define SLOs and Service Level Indicators (SLIs) that align with customer expectations and business impact.

  • Architect for Reliability: Partner with Solution Architects and Developers to establish and enforce best practices for integration architecture, ensuring solutions are built for scalability, resiliency, and performance from day one.

  • Troubleshoot & Remediate: Serve as the primary escalation point for critical incidents related to custom integrations. Lead troubleshooting efforts and perform hands-on development work to resolve complex, high-stakes issues.

2. Infrastructure & Automation:

  • Manage Integration Infrastructure: Own the cloud infrastructure (AWS, Azure, etc.) that hosts our custom solutions, focusing on security, cost-optimization, and scalability.

  • Champion Infrastructure as Code (IaC): Partner with our core engineering team to align on best practices and systems to manage deployments and infrastructure.

  • Automate Everything: Develop and manage CI/CD pipelines for the safe and efficient deployment of integration code and infrastructure changes.

3. Cross-Functional Collaboration & Knowledge Sharing:

  • Bridge Team Gaps: Create and document clear operational handoffs and processes between teams to ensure a seamless flow from development to production support.

  • Lead Post-Mortems: Foster a blameless post-mortem culture to analyze incidents, identify root causes, and drive actionable improvements to prevent recurrence.

  • Create a Knowledge Hub: Develop runbooks, architectural diagrams, and best-practice guides to empower all of Professional Services with the knowledge to better support our solutions.

Qualifications & Experience

Required:

  • You have experience in a Site Reliability Engineering (SRE), DevOps, or Cloud Infrastructure role.

  • Expertise with AWS.

  • Hands-on with Infrastructure as Code (Terraform, CloudFormation).

  • Experience with observability platforms (Datadog, New Relic, Prometheus, Grafana).

  • Proficient in at least one scripting or programming language, preferably Ruby.

  • Experience building and managing CI/CD pipelines (e.g., GitLab CI, GitHub Actions, Jenkins).

  • A systems-thinker with a passion for troubleshooting complex problems and improving processes.

Preferred:

  • Experience working within a client-facing Professional Services or technical consulting organization.

  • Background in managing the reliability of APIs, middleware, and complex data integrations.

  • Experience with containerization and orchestration (Docker, Kubernetes).

  • Bachelor's degree in Computer Science or a related technical field.

Get in on all the awesome at Instructure!

We offer competitive, meaningful benefits in every country where we operate. While they vary by location, here's a general idea of what you can expect:

  • Competitive compensation, plus all full-time employees participate in our ownership program - because everyone should have a stake in our success.

  • Flexible work culture. Our remote, hybrid and in-office collaboration spaces vary by role, team and location.

  • Generous time off, including local holidays and our annual “Dim the Lights” period in late December, when teams are encouraged to step back and recharge based on departmental needs.

  • Comprehensive wellness programs and mental health support

  • Learning and development resources, including professional development tools and tuition reimbursement, to support your growth

  • The technology and tools you need to do your best work

  • Motivosity employee recognition program

  • A culture rooted in inclusivity, support, and meaningful connection

We believe in hiring great people and treating them right. The more diverse we are, the better our ideas and outcomes.

Instructure is an Equal Opportunity Employer. We comply with applicable employment and anti-discrimination laws in every country where we operate.

All employees must pass a background check as part of the hiring process. To help protect our teams and systems, we’ve implemented identity verification measures. Candidates may be asked to verify their legal name, current physical location, and provide a valid contact number and residential address, in accordance with local data privacy laws.

Any attempt to misrepresent personal or professional information will result in disqualification.

Skills Required

  • Experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure role
  • Expertise with AWS
  • Hands-on with Infrastructure as Code (Terraform, CloudFormation)
  • Experience with observability platforms (Datadog, New Relic, Prometheus, Grafana)
  • Proficient in at least one scripting or programming language (preferably Ruby)
  • Experience building and managing CI/CD pipelines (GitLab CI, GitHub Actions, Jenkins)
  • Systems-thinking and strong troubleshooting/problem-solving skills
  • Experience working within client-facing Professional Services or technical consulting
  • Background in managing reliability of APIs, middleware, and complex data integrations
  • Experience with containerization and orchestration (Docker, Kubernetes)
  • Bachelor's degree in Computer Science or related technical field
  • Ability to pass a background check and complete identity verification

Instructure Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Instructure and has not been reviewed or approved by Instructure.

  • Fair & Transparent Compensation Pay is considered market-competitive for many engineering, product, and quota-carrying roles, especially when factoring base, variable, and equity. In these tracks, total rewards are often characterized as fair-to-strong for the role and location.
  • Healthcare Strength Health coverage is described as comprehensive, including medical, dental, vision, mental-health support, and HSA/FSA options. Employer contributions are often portrayed as strong, and core medical benefits receive consistently positive marks.
  • Leave & Time Off Breadth Time off offerings include flexible or “unlimited” PTO, paid holidays, and paid sick time, paired with widespread remote/hybrid flexibility. This breadth of time off and work flexibility is often viewed as a meaningful perk that enhances overall value.

Instructure Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Salt Lake City, UT
1,233 Employees
Year Founded: 2008

What We Do

Instructure is helping people grow from the first day of school to the last day of work. More than 30 million people use its Canvas and Bridge platforms for learning management and employee development.

Similar Jobs

Pfizer Logo Pfizer

Scientist

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office
Budapest, HUN
121990 Employees

Pfizer Logo Pfizer

Vice President, Strategy, Value & Innovation

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
43 Locations
121990 Employees

Pfizer Logo Pfizer

Vice President, Build - Data, Engineering & AI

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
43 Locations
121990 Employees

Pfizer Logo Pfizer

Director, Medical Insights

Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
In-Office or Remote
35 Locations
121990 Employees
177K-294K Annually

Similar Companies Hiring

Scrunch  Thumbnail
Artificial Intelligence • Information Technology • Marketing Tech • Software • SEO
Salt Lake City, Utah
Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account