Site Reliability Engineer I

Posted 5 Days Ago
Hiring Remotely in San Mateo, CA, USA
In-Office or Remote
Junior
Cloud • Information Technology
The Role
Act as first responder for customer-impacting incidents, monitor and respond to Zabbix alerts, ensure pod and server farm health, run filesystem checks, support Vault deployments and migrations, troubleshoot pod/Ansible/network issues, participate in on-call rotation, document and automate daily tasks, and coordinate escalations with DC techs and management.
Summary Generated by Built In

About Backblaze
Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our partners, we’re helping customers break free from the restrictive, overpriced legacy solutions that hold them back, and blaze forward with the full power of the open cloud in their hands.

Founded in 2007, we scaled the business with less than $3 million in outside funding until 2021, when we did a traditional IPO on the Nasdaq stock exchange. Today, Backblaze generates over $100m in revenue and is the leading specialized storage cloud - managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals.
But while there is a lot to celebrate in our past, there is almost as much opportunity ahead of us. We are seeking a Site Reliability Engineer I to join our team!

What You’ll Do: 

  • Act as first point of contact for all customer affecting issues
  • Be a Key Driver for managing the resolution of technical problems
  • Ensure that incident management processes are following and that incident post-mortems are completed to capture process deviations and areas for improvement
  • Deliver consistent communication to Management
  • Respond to zabbix alerts/regular monitoring of zabbix, either by taking direct action on alerts or escalating. Acknowledge every alert if direct action taken, or with escalation point of contact.
  • Make sure escalations are handed off successfully.
  • Ensure health of pods across all sites (define pod alerts on zabbix).
  • Work through daily filesystem checks for pods.
  • Troubleshoot technical issues for DC Techs -> advanced pod questions, deployment questions, migration troubleshooting, and ansible playbook issues.
  • Identification and escalating any potential issues regarding the network.
  • Vault pre-deployment configuration and testing.
  • Start Vault Migrations, monitor migration pods, handle applicable migration pod health checks.
  • Document/Work on automating Daily Items.
  • Document/Provide Network IP's for upcoming deployments.
  • Monitor Releases/Updates to the Server Farm, escalate issues as they arise.
  • Engaging in on-call rotation shifts.
  • Assist fellow TechOps team members in handling tasks.
  • Making recommendations for improvements in organizational productivity.
  • Be able to work outside of normal business hours(weekend shift, holidays & evenings) as needed

The Right Fit:

  • Must be located in Bangalore.
  • 2 - 4 years of relevant experience.
  • Knowledge of Sysadmin and Linux skills.
  • Desire to learn and develop all necessary technical skills.
  • Strong analytical thinking.
  • Strong skills in working with different teams and communication.
  • Knowledge of network cabling, network classification, and network topology.

At this point, we hope you're feeling excited about the job description you're reading. Even if you don't meet every requirement, we still encourage you to apply. Learning, developing, and growing are key parts of our culture. We're eager to meet people who believe in our mission and can contribute to our team in various ways. We want people to feel comfortable expressing their true selves and to come, stay, and do their best work here.

At Backblaze, we value being fair and good to our customers, partners, and employees. That’s why diversity, equity, and inclusion are at the core of our values. We are committed to fostering a workforce where all employees feel a sense of belonging regardless of race, ethnicity, nationality, gender, sexual orientation, age, religion, socio-economic status, ability, veteran status, and education. We believe that our dedication to cultivating a diverse workspace not only allows us to better serve our customers in over 175 countries, but further reinforces our commitment to doing the right thing. We are proud to be an Equal Opportunity Employer.

To understand more about the data we collect and process as part of your application, please view our Backblaze Employee Privacy Notice.

Skills Required

  • Located in Bangalore
  • 2-4 years of relevant experience
  • Sysadmin and Linux skills
  • Experience monitoring and responding to Zabbix alerts
  • Experience with Kubernetes pods and pod health management
  • Experience with Ansible playbooks and automation
  • HashiCorp Vault configuration and migration experience
  • Knowledge of network cabling, network classification, and network topology
  • Strong analytical thinking and cross-team communication skills
  • Willingness to participate in on-call rotation and work outside normal hours
  • Desire to learn and develop technical skills
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Mateo, CA
363 Employees
Year Founded: 2007

What We Do

Backblaze provides cloud storage and online backup that’s astonishingly easy to use and affordable. We are entrusted with over an exabyte of data from customers in 175 countries. Our approach is guided by honesty, transparency, and a commitment to doing the right thing for our customers and co-workers. Our customers are happy, and so are our co-workers: In a recent survey, 97% of our team rated Backblaze as “a great place to work.” While there is a lot to celebrate in our past, there is almost as much opportunity ahead of us.

Similar Jobs

AuthZed Logo AuthZed

Senior Site Reliability Engineer

Artificial Intelligence • Information Technology • Software • Database
Remote
2 Locations
30 Employees
150K-195K Annually

MongoDB Logo MongoDB

Site Reliability Engineer

Big Data • Cloud • Software • Database
Easy Apply
Remote or Hybrid
10 Locations
5550 Employees
127K-249K Annually
Remote
United States of America
1612 Employees
185K-227K Annually

Similar Companies Hiring

Standard Template Labs Thumbnail
Artificial Intelligence • Information Technology • Software
New York, NY
25 Employees
NODA AI Thumbnail
Artificial Intelligence • Information Technology • Software • Cybersecurity
Sydney, AU
54 Employees
Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account