Epic Site Reliability Engineer II

Posted 4 Days Ago
2 Locations
In-Office
85K-110K Annually
Mid level
Healthtech • Database
The Role
Design, implement, and maintain observability and performance solutions for production systems. Monitor, analyze, and optimize performance, conduct capacity planning, automate operational tasks, collaborate with development and security teams, run performance testing, document systems, and drive continuous reliability and performance improvements.
Summary Generated by Built In

Pay Range: $85,000-110,000, plus yearly bonus


This position is hybrid and will require 3 days on site at one of the following Quest sites: Secaucus, NJ or Schaumburg, IL.

 

Benefits Information: We are proud to offer best-in-class benefits and programs to support employees and their families in living healthy, happy lives. Our pay and benefit plans have been designed to promote employee health in all respects physical, financial, and developmental. Depending on whether it is a part-time or full-time position, some of the benefits offered may include: 

  • Day 1 Medical, supplemental health, dental & vision for FT employees who work 30+ hours 
  • Best-in-class well-being programs 
  • Annual, no-cost health assessment program 
  • Blueprint for Wellness 
  • healthyMINDS mental health program 
  • Vacation and Health/Flex Time 
  • 6 Holidays plus 1 MyDay off 
  • FinFit financial coaching and services 
  • 401(k) pre-tax and/or Roth IRA with company match up to 5% after 12 months of service 
  • Employee stock purchase plan 
  • Life and disability insurance, plus buy-up option 
  • Flexible Spending Accounts Annual incentive plans 
  • Matching gifts program
  •  Education assistance through MyQuest for Education Career advancement opportunities and so much more!
     
Responsibilities

Responsibilities:

System Monitoring and Analysis:

  • Implement and maintain robust observability solutions to monitor system performance, identifying bottlenecks, and ensuring optimal operation.
  • Utilize tools to gather, analyze, and visualize key performance metrics.

Performance Optimization:

  • Proactively identify and address performance bottlenecks through in-depth analysis and optimization strategies.
  • Work closely with development teams to implement performance improvements and enhance overall system efficiency.

Capacity Planning:

  • Conduct capacity planning exercises based on observed patterns and future growth projections.
  • Collaborate with infrastructure and development teams to ensure adequate resources are available to meet system demands.

Automation and Scripting:

  • Develop and maintain automation scripts for routine tasks, enabling efficient monitoring and response procedures.
  • Implement automated processes for scaling and provisioning resources based on observed workload patterns.

Documentation:

  • Document system architecture, configurations, and observability best practices to facilitate knowledge transfer and onboarding for team members.
  • Keep documentation up-to-date to reflect changes in the system and its monitoring setup.

Collaboration with Development Teams:

  • Work closely with software engineers to integrate observability tools into the development lifecycle.
  • Provide guidance on building observable systems and assist in instrumenting applications for effective monitoring.

Continuous Improvement:

  • Stay informed about industry best practices and emerging technologies related to observability and performance engineering.
  • Drive continuous improvement initiatives to enhance the reliability and performance of systems.
  • Security and Compliance:
  • Collaborate with security teams to implement monitoring and observability measures that align with security requirements and compliance standards.
  • Participate in security incident response activities and contribute to ongoing security assessments.
  • Training and Knowledge Sharing:
  • Conduct training sessions for team members and other stakeholders on observability tools, best practices, and performance engineering concepts.
  • Foster a culture of knowledge sharing within the organization.

And other duties as assigned.

Qualifications

Required Work Experience: 

  • 4+ years of experience with multiple APM tools and extensive experience with Dynatrace
  • 3+ years SRE experience
  • Experience in software development, infrastructure, or operations roles
  • Certifications in relevant technologies (e.g. AWS, DevOps, Kubernetes, Dynatrace, Azure, etc.)
  • Working experience building CI/CD pipelines and version control systems
  • Working experience with scripting languages (e.g. Python, Bash, Go, etc.)
  • Excellent problem-solving and communication skills.
  • Ability to work collaboratively in a fast-paced, agile environment.

Preferred Work Experience: 

  • Working experience with Neoload, Jmeter or equivalent performance testing tool.
  • Experience executing software load and performance testing in an enterprise environment.
  • Experience testing applications hosted in the cloud.
  • Experience with infrastructure as code tools such as Terraform or CloudFormation.
  • Deep understanding of Linux systems administration and networking principles.
  • Experience with containerization and orchestration technologies such as Docker and Kubernetes.
  • Experience or familiarity with IIS, HTML, Java, Jboss.
  • Experience in Chaos Engineering
  • Programming experience using.NET, C, C++, Java, or other popular programming languages.  Perl/Python/JavaScript scripting experience may be considered equivalent.
  • Terraform and Ansible experience.
  • Exposure to Splunk tools.
  • Exposure to microservices.
  • Dynatrace Certifications
  • AWS/Azure/GCP Certifications
  • Chaos Engineering Certifications
  • Agile Certifications

Physical and Mental Requirements: 

  • Ability to sit/stand for long periods of time.
  • Ability to handle high stress situations.
  • Ability to lift up to 50 lbs.

Knowledge: 

  • Site Reliability Engineering Principles
  • DevSecOps Principles
  • Agile (SAFe)
  • Healthcare industry
  • ITLT
  • ServiceNow
  • Jira/Confluence

Skills: 

  • Dynatrace/Prometheus/Grafana
  • Neoload/Jmeter
  • Splunk
  • AWS/Azure/GCP
  • SAFe Agile
  • Strong communication skills (written/verbal)
  • Time management
  • Analytic problem solver
  • Self-starter
  • Result oriented and proven ability in organizing priorities

Education

  • Bachelor’s Degree Bachelor's degree in Computer Science, Engineering, or a related field (Required)

Licenses and Certifications

  • Agile Certification (Project Management) (Preferred)
About the Team Quest Diagnostics honors our service members and encourages veterans to apply.
While we appreciate and value our staffing partners, we do not accept unsolicited resumes from agencies. Quest will not be responsible for paying agency fees for any individual as to whom an agency has sent an unsolicited resume.
Equal Opportunity Employer: Race/Color/Sex/Sexual Orientation/Gender Identity/Religion/National Origin/Disability/Vets or any other legally protected status.

Skills Required

  • 4+ years experience with multiple APM tools and extensive experience with Dynatrace
  • 3+ years Site Reliability Engineering (SRE) experience
  • Experience in software development, infrastructure, or operations roles
  • Certifications in relevant technologies (e.g., AWS, DevOps, Kubernetes, Dynatrace, Azure)
  • Working experience building CI/CD pipelines and using version control systems
  • Working experience with scripting languages (e.g., Python, Bash, Go)
  • Excellent problem-solving and communication skills
  • Ability to work collaboratively in a fast-paced, agile environment
  • Bachelor's degree in Computer Science, Engineering, or a related field
  • Experience with Neoload, JMeter, or equivalent performance testing tools
  • Experience executing software load and performance testing in an enterprise environment
  • Experience testing applications hosted in the cloud
  • Experience with infrastructure-as-code tools such as Terraform or CloudFormation
  • Deep understanding of Linux systems administration and networking principles
  • Experience with containerization and orchestration such as Docker and Kubernetes
  • Experience or familiarity with IIS, HTML, Java, JBoss
  • Experience in Chaos Engineering
  • Programming experience using .NET, C, C++, Java, or other languages
  • Perl/Python/JavaScript scripting experience
  • Ansible experience
  • Exposure to Splunk and microservices architectures
  • Dynatrace, AWS/Azure/GCP, or Chaos Engineering certifications
  • Agile (SAFe) experience and Agile certifications
  • Familiarity with ServiceNow, Jira, and Confluence
  • Ability to lift up to 50 lbs and handle high-stress situations
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Secaucus, NJ
25,839 Employees
Year Founded: 1967

What We Do

Quest Diagnostics (NYSE: DGX) empowers people to take action to improve health outcomes. Derived from the world's largest database of clinical lab results, our diagnostic insights reveal new avenues to identify and treat disease, inspire healthy behaviors and improve health care management. Quest annually serves one in three adult Americans and half the physicians and hospitals in the United States, and our 47,000 employees understand that, in the right hands and with the right context, our diagnostic insights can inspire actions that transform lives. The company offers physicians the broadest test menu (3,000+ tests), is a pioneer in developing innovative new tests, is the leader in cancer diagnostics, provides anatomic pathology (AP) services, & interpretive consultation through its medical & scientific staff of about 900 M.D.s & Ph.D.s. The company reported 2020 revenues of $9.44 billion. Quest Diagnostics offers the most extensive clinical testing network in the U.S., with laboratories in most major metropolitan areas, & in Mexico, the UK & India. The company also operates four esoteric laboratories, 40 outpatient AP laboratories, & 160 smaller, rapid-response laboratories. Patients may have specimens collected in any of the company’s approximately 2,250 patient service centers. On a typical workday, testing is performed for about 550,000 patients. Quest Diagnostics empowers healthcare organizations & clinicians with state-of-the-art connectivity solutions. The company is the leading provider of pre-employment drugs-of-abuse screening for employers & risk assessment services for the life insurance industry. It is the world’s 2nd largest provider of clinical trials testing for new pharmaceuticals. More information is available at www.questdiagnostics.com. Language Assistance / Non-Discrimination Notice Asistencia de Idiomas / Aviso de no Discriminación 語言協助 / 不歧視通知 www.QuestDiagnostics.com/home/nondiscrimination

Similar Jobs

In-Office
Secaucus, NJ, USA
25839 Employees
90K-120K Annually
Hybrid
Jersey City, NJ, USA
289097 Employees
Hybrid
Jersey City, NJ, USA
289097 Employees

PwC Logo PwC

Salesforce Consulting Senior Associate

Artificial Intelligence • Professional Services • Business Intelligence • Consulting • Cybersecurity • Generative AI
Hybrid
64 Locations
370000 Employees
77K-202K Annually

Similar Companies Hiring

Sailor Health Thumbnail
Healthtech • Social Impact • Telehealth
New York City, NY
20 Employees
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account