Production Engineer

Posted Yesterday
Be an Early Applicant
Newark, NJ, USA
In-Office
109K-120K Annually
Mid level
Healthtech
The Role
Manage and improve WebMD’s Windows, Linux, and Kubernetes infrastructure as part of the SRE team. Responsibilities include automation with PowerShell and Python, observability using Prometheus, Grafana, and ELK, distributed-system troubleshooting, CI/CD support, configuration management, load balancing, and maintaining reliable high-traffic web services. The role participates in on-call rotations and drives infrastructure and operational improvements.
Summary Generated by Built In

WebMD is the most recognized and trusted brand of health information and the leading provider of health information services, serving consumers, physicians, healthcare professionals, employers and health plans through our public and private online portals and WebMD the Magazine. The WebMD Health Network includes WebMD, Medscape, MedicineNet, eMedicine, RxList, theheart.org and Medscape Education. Our consumer portals and mobile health applications provide engaging, relevant and credible health and wellness information, personalized health assessment tools and access to online communities.
WebMD is an Equal Opportunity/Affirmative Action employer and does not discriminate on the basis of race, ancestry, color, religion, sex, gender, age, marital status, sexual orientation, gender identity, national origin, medical condition, disability, veterans status, or any other basis protected by law.

About the job 
As a Production Engineer on WebMD’s SRE team, your primary focus will be maintaining, automating, and scaling our core Windows Server Infrastructure. While this is a Windows-first systems and production engineering role, you will have the opportunity to gain exposure and work on Linux and Kubernetes environments. 
You will:
 

  • Manage and optimize our core  Windows Server environments and IIS web applications  to ensure the high availability, security and performance.
     
  • Automate tasks and routines through modern scripting (primarily  PowerShell  and Python), simplifying developer and operations workflows.
     
  • Take ownership of tasks and projects within our Observability stack  (including  Prometheus  and ELK) to maintain full-stack visibility, latency tracking, and scalability.
     
  • Identify opportunities for infrastructure improvement, collaborate with other Operations teams, and act as a driver of change.

 A successful candidate will be able to:
 

  • Demonstrate fluency in working with  Windows Server operating systems , paired with a strong command-line aptitude and a  willingness to learn or work within Linux environments .
     
  • Show hands-on proficiency in deploying, managing, or troubleshooting workloads within  Kubernetes .
     
  • Provide samples of automation scripts built using  PowerShell , Python, or Go.
     
  • Leverage AI-assisted engineering tools  (such as GitHub Copilot, Claude Code, or LLMs) to accelerate scripting, troubleshooting, and daily operational workflows.
     
  • Drive continuous improvement within our  Observability and monitoring ecosystems (Prometheus, Grafana, ELK Stack) .
     
  • Demonstrate operational competency with critical stack technologies:
     
  • Configuration management (e.g., Puppet, Ansible)
     
  • Load balancers (specifically  F5 Big-IP )
     
  • Data and messaging layers (e.g., Redis, Kafka, RabbitMQ)
     
  • Convey a fundamental understanding of CI/CD pipelines (GitLab, Jenkins, or Azure DevOps).
     
  • Participate in the team’s on-call rotation to address production issues outside of working hours.

 About you (Qualifications):
 

  • Bachelor's degree in Computer Science, Technology, Engineering, Math, or equivalent practical experience.
     
  • 4+ years of overall experience in IT/Infrastructure, with 2+ years of relevant experience  supporting web-based applications for large-scale, high-traffic websites.
     
  • Proven experience troubleshooting and performance tuning distributed systems, IIS, and web applications.
     
  • Fundamental understanding of network concepts (TCP/IP, DNS, HTTP) and traffic management via  F5 Big-IP .
     
  • An open, adaptive mindset  toward adopting new methodologies, whether that means jumping into a Linux CLI or integrating AI tools into your daily engineering workflow.
     
  • Excellent communication skills, with the ability to convey detailed technical concepts clearly to both peers and non-technical stakeholders.

Bonus Eligible: This position is also eligible for a discretionary company bonus, based upon business results.

Compensation: $109,000 - $120,000

Benefits: Employees in this position are eligible to participate in the company sponsored benefit programs, including the following within the first 12 months of employment:

  • Health Insurance (medical, dental, and vision coverage)
  • Paid Time Off (including vacation, sick leave, and flexible holiday days)
  • 401(k) Retirement Plan with employer matching
  • Life and Disability Insurance
  • Employee Assistance Program (EAP)
  • Commuter and/or Transit Benefits (if applicable)

Eligibility for specific benefits may vary based on job classification, schedule (e.g., full-time vs. part-time), work location and length of employment.


Skills Required

  • Bachelor’s degree in Computer Science, Technology, Engineering, Math, or equivalent practical experience
  • 4+ years of overall experience in IT or infrastructure
  • 2+ years of relevant experience supporting web-based applications for large-scale, high-traffic websites
  • Experience with Windows Server and command-line environments
  • Experience working with or willingness to work within Linux environments
  • Hands-on experience deploying, managing, or troubleshooting Kubernetes workloads
  • Automation scripting experience with PowerShell, Python, or Go
  • Experience troubleshooting and performance tuning distributed systems, IIS, and web applications
  • Understanding of TCP/IP, DNS, and HTTP networking concepts
  • Experience with F5 Big-IP traffic management
  • Experience with observability and monitoring tools such as Prometheus, Grafana, or ELK Stack
  • Understanding of configuration management tools such as Puppet or Ansible
  • Fundamental understanding of CI/CD pipelines using GitLab, Jenkins, or Azure DevOps
  • Ability to participate in an on-call rotation
  • Excellent communication skills with technical and non-technical stakeholders
  • Experience with Redis, Kafka, or RabbitMQ
  • Experience using AI-assisted engineering tools such as GitHub Copilot, Claude Code, or LLMs
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: New York, NY
2,062 Employees
Year Founded: 1995

What We Do

Our Company: WebMD is a leading provider of health information services to consumers, physicians, healthcare professionals, employers and health plans. Our Business: WebMD is the leading provider of health information and services to consumers and healthcare professionals. The online healthcare information, decision-support applications and communications services that we provide: help consumers take an active role in managing their health by providing objective healthcare information and lifestyle information. make it easier for physicians and healthcare professionals to access clinical reference sources, stay abreast of the latest clinical information, learn about new treatment options, earn continuing medical education credits and communicate with peers. enable employers and health plans to provide their employees and plan members with access to personalized heath and benefit information and decision support technology that helps them make informed benefit, provider and treatment choices.

Similar Jobs

Generant Company, Inc. Logo Generant Company, Inc.

Production Engineer

Energy • Industrial • Automation • Manufacturing
In-Office
Butler, NJ, USA
150 Employees
65K-85K

Barclays Logo Barclays

Support Engineer

Fintech • Financial Services
In-Office
Jefferson Park, NJ, USA
83500 Employees
80K-120K Annually

BAE Systems, Inc. Logo BAE Systems, Inc.

Electrical Engineer

Aerospace • Hardware • Information Technology • Security • Software • Cybersecurity • Defense
Hybrid
Wayne, NJ, USA
40000 Employees
129K-219K Annually
In-Office
Jersey City, NJ, USA
1061 Employees
150K-210K Annually

Similar Companies Hiring

Sailor Health Thumbnail
Healthtech • Social Impact • Telehealth
New York City, NY
20 Employees
Granted Thumbnail
Artificial Intelligence • Healthtech • Insurance • Mobile • Financial Services
New York, New York
23 Employees
OneImaging Thumbnail
Healthtech
Miami, FL
62 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account