Senior Software Engineer, Infrastructure

Reposted One Month Ago
Palo Alto, CA, USA
Hybrid
30K-30K Annually
Senior level
Artificial Intelligence • eCommerce
The Role
As a Senior Site Reliability Engineer, you'll ensure the reliability and scalability of production systems, define SLOs and SLIs, lead incident responses, and improve cloud infrastructure performance.
Summary Generated by Built In
About Us

We're living through a fundamental shift in how people discover, evaluate, and purchase products. The next generation doesn't respond to traditional marketing -- they build relationships with brands through authentic social interactions, seek recommendations from communities they trust, and expect personalized experiences that feel human, not corporate.

At Nectar Social, we're building the AI-native social operating system that enables this new era of commerce. We believe every social interaction should deepen the relationship between brands and their communities while creating genuine value for both sides.

Founded by ex-Meta product and engineering leaders, we've raised over $30M in total capital from investors including GV and True Ventures. We work with brands like Oura Health, Caraway, e.l.f. Cosmetics, Kosas, OLIPOP, and many more. We're building the future of social commerce -- where community, conversation, and commerce converge.

The Role

We're looking for a Senior Software Engineer, Infrastructure to own the reliability, scalability, and operational excellence of the production systems that power Nectar's platform. We run high-volume data ingestion pipelines and real-time AI agents on top of a fast-growing customer base, and we need a seasoned SRE to help us scale these systems safely and keep them running flawlessly.

As one of our first dedicated SREs, you'll have outsized impact and ownership. You'll define how we measure, operate, and harden our infrastructure -- establishing the reliability foundations that the rest of the engineering team builds on as we scale.

What You'll Be Doing
  • Own the reliability and scalability of our production systems as they handle rapidly growing volumes of social data and real-time AI workloads

  • Define and drive SLOs, SLIs, and error budgets, and build the observability, alerting, and on-call practices to support them

  • Lead incident response and blameless postmortems, then turn what we learn into systemic improvements that prevent recurrence

  • Improve performance, cost efficiency, and capacity planning across our cloud infrastructure as the platform scales

  • Harden our infrastructure-as-code, deployment, and CI/CD pipelines for resilience and repeatability

  • Partner with engineering teams to embed reliability into system design and raise the operational bar across the org

What We're Looking For
  • 5+ years of experience operating production systems as an SRE, infrastructure, or platform engineer

  • Experience scaling databases, data infrastructure, or complex production platforms under significant load

  • Hands-on expertise with cloud infrastructure (AWS or similar) and infrastructure-as-code tooling

  • Solid programming skills for building automation, tooling, and operational services

  • Comfortable operating in fast-moving startup environments with high ownership and autonomy

  • A reliability-first mindset balanced with pragmatism about velocity and cost

Bonus Points
  • Experience standing up or maturing an SRE practice at an early-stage or rapidly scaling company

  • Familiarity with our tech stack: AWS, Pulumi, Postgres, ClickHouse, Turbopuffer, or Temporal

  • Background in capacity planning, performance engineering, or cost optimization at scale

What We Offer
  • Competitive compensation and early equity

  • Health, vision, and dental benefits + 401(k) match

  • Hybrid work schedule (4 days in office) with 3 flex remote days per quarter (available after 3 months)

  • Comprehensive stipends: $1,000/month housing stipend for living near the office, $50/month mobile stipend, $50/month internet stipend for remote employees, and commuter benefits (train stipend or Palo Alto parking permit reimbursement)

  • Free lunch in the heart of University Ave. in Palo Alto

Skills Required

  • 5+ years of experience operating production systems as an SRE, infrastructure, or platform engineer
  • Experience scaling databases, data infrastructure, or complex production platforms under significant load
  • Hands-on expertise with cloud infrastructure (AWS or similar)
  • Solid programming skills for building automation, tooling, and operational services
  • Familiarity with tech stack: AWS, Pulumi, Postgres, ClickHouse, Turbopuffer, or Temporal
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Seattle, WA
20 Employees

What We Do

AI Community Commerce platform for social-first Brands.

Similar Jobs

ServiceNow Logo ServiceNow

Senior Software Engineer

Artificial Intelligence • Cloud • HR Tech • Information Technology • Productivity • Software • Automation
Hybrid
Mountain View, CA, USA
29000 Employees
143K-243K Annually
Hybrid
2 Locations
289097 Employees

VSCO Logo VSCO

Senior Software Engineer

Digital Media • Mobile • Productivity • Social Media • Software
In-Office
San Francisco, CA, USA
110 Employees
165K-185K Annually

Crunchyroll Logo Crunchyroll

Senior Software Engineer

Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Hybrid
Los Angeles, CA, USA
1300 Employees
183K-229K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account