Senior Site Reliability Engineer

Posted Yesterday
Be an Early Applicant
11 Locations
Remote
Senior level
Information Technology • Consulting
The Role
Senior Site Reliability Engineer responsible for developing and operating reliable, scalable, secure distributed systems across Azure and traditional data centers. Duties include automation, monitoring, capacity planning, system design, CI/CD, incident analysis, service reliability improvements, and software platform development. The role requires strong Kubernetes, Terraform, cloud, Linux, software development, and SRE expertise, with experience in distributed systems and application monitoring preferred.
Summary Generated by Built In

We are looking for a Senior Site Reliability Engineer who is interested in an opportunity to work for an innovative hospitality company with cutting-edge technologies, with new development activities and challenges ahead. Our international team members share a common desire to develop brilliant products on reliable and resilient systems, along with their own skills. We run our services in Azure and traditional data centers. Take a chance to make a valuable contribution and enhance your professional skills.

About the job:

Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that cloud services—both our internally critical and our externally-visible systems—have reliability, uptime appropriate to customer needs, and a fast rate of improvement. Additionally, SREs will keep an ever-watchful eye on our systems' capacity and performance.

On the SRE team, you’ll have the opportunity to manage the complex challenges of scale that are unique to the project while using your expertise in coding, algorithms, complexity analysis, and large-scale system design. You will provide scalable, reliable, durable, and secure services using a customer-first approach while innovating technically. You will understand our customer needs and how we can meet them.

Responsibilities:

  • Develop and improve the whole lifecycle of services
  • Establish and improve monitoring capabilities to reduce outage frequency and duration
  • Create sustainable systems through automation and uplifts
  • Develop and scale systems sustainably through mechanisms such as automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Lead designs of major software components, systems, and features to improve the availability, scalability, latency, and efficiency of our services
  • Analyze and support services before they go live via system design consulting, developing software platforms and frameworks, capacity planning
  • Conduct post-incident analysis and reviews with an attitude of continuous improvement

Requirements:

  • Ideally, strong experience in Azure Services and capabilities, but other cloud services (AWS, Google Cloud Platform etc.) will be considered
  • Confidence and strong experience with KubernetesRecent and fluent Terraform and (Chef platform experience nice to have)
  • Extensive expertise in software development/testing, development operations, and site reliability engineering
  • Experience of Unix/Linux administration - an appreciation of systems internals (e.g., filesystems, system calls) is a bonus
  • Experience with Continuous Integration and Deployment (CI/CD) and release orchestration and Configuration Management of VMs
  • Cloud-agnostic approach, with flexibility to work across various cloud platforms
  • Experience programming in one or more of the following languages: C#,, C++, Java, Python, JavaScript, Go, Perl, or Ruby

Nice to have:

  • Bachelor's degree in Computer Science, similar technical field of study, or equivalent practical experience
  • Experience in distributed systems, storage systems, or databases
  • Experience designing, analyzing, and troubleshooting large-scale distributed systems
  • Systematic problem-solving approach, combined with excellent communication skills and a sense of ownership and drive
  • Experience in configuring application monitoring with Azure Monitor and Application Insight
  • Experience with Service Mesh
  • Previous experience as a DevOps engineer is preferred

We offer*:

  • Flexible working format - remote, office-based or flexible
  • A competitive salary and good compensation package
  • Personalized career growth
  • Professional development tools (mentorship program, tech talks and trainings, centers of excellence, and more)
  • Active tech communities with regular knowledge sharing
  • Education reimbursement
  • Memorable anniversary presents
  • Corporate events and team buildings
  • Other location-specific benefits

*not applicable for freelancers

Skills Required

  • Strong experience with Azure services and capabilities, or other cloud platforms such as AWS or Google Cloud Platform
  • Strong experience with Kubernetes
  • Recent and fluent Terraform experience
  • Extensive expertise in software development, software testing, DevOps, and site reliability engineering
  • Experience with Unix/Linux administration
  • Experience with continuous integration and deployment, release orchestration, and virtual machine configuration management
  • Cloud-agnostic approach and flexibility to work across various cloud platforms
  • Programming experience in one or more of C#, C++, Java, Python, JavaScript, Go, Perl, or Ruby
  • Bachelor's degree in Computer Science or a similar technical field, or equivalent practical experience
  • Chef platform experience
  • Experience with distributed systems, storage systems, or databases
  • Experience designing, analyzing, and troubleshooting large-scale distributed systems
  • Experience configuring application monitoring with Azure Monitor and Application Insights
  • Experience with service mesh
  • Previous experience as a DevOps engineer
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Valletta
2,135 Employees
Year Founded: 2002

What We Do

N-iX is a global software solutions and engineering services company that helps world’s leading organizations turn challenges into lasting business value, operational efficiency, and revenue growth using advanced technology. Whether you need to build a custom solution, modernize your digital product or acquire extra tech expertise - we have the experience and capabilities to ensure your success. With over 2,000 professionals in 25 countries across Europe and the Americas, N-iX offers expert solutions in cloud, data analytics, embedded software, IoT, AI, machine learning, and other tech domains. Being in business for over two decades, we have worked with dozens of industry-leading enterprises and Fortune 500 companies creating value across a wide variety of sectors, including finance, manufacturing, supply chain, retail, e-commerce, healthcare, and more. Our unique combination of business domain expertise and technical know-how enables us to effectively collaborate with ISVs, tech companies, and enterprises of all sizes. Thanks to the strong tech ecosystem and partnerships with AWS, GCP, Microsoft, SAP, OpenText, Snowflake, and others, we bring extra speed, scale and efficiency to more than 160 organizations across the globe. N-iX is recognized by numerous industry awards, such as CRN Solution Provider 500, Global Outsourcing 100 by IAOP, ISG Provider Lens™, Modern Application Development services providers by Forrester, etc

Similar Jobs

Alpaca Logo Alpaca

Senior Site Reliability Engineer

Fintech • Information Technology
Remote
13 Locations
132 Employees
Remote
Chile
359 Employees
84K-144K Annually
Remote
6 Locations
77 Employees

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account