Staff Software Engineer, Agentic Tools

Posted Yesterday
Hiring Remotely in United States
Remote
210K-245K Annually
Expert/Leader
Software
Cribl is the AI Platform for Telemetry.
The Role
Build and operate production agentic workflows, runtimes, orchestration systems, tool integrations, and AI platform infrastructure. Develop guardrails, observability, evaluation, secure access controls, event-driven architectures, and continuous delivery practices. Partner across Engineering to drive adoption, improve software delivery, and take systems from design through production operations. The role includes evaluating AI tooling, supporting agentic coding workflows, defining success metrics, and potentially participating in on-call duties.
Summary Generated by Built In

Join the company that’s building the telemetry infrastructure for the AI era. At Cribl, we partner with IT and Security teams at many of the world’s biggest enterprises, including half of the Fortune 100, to bridge the gap between AI ambition and infrastructure reality. As the AI Platform for Telemetry, we give customers the choice, control, and flexibility to manage and analyze telemetry for both humans and agents, so they can build what’s next.

We’re one of the fastest‑growing private companies and a leading player in a massive, fast‑moving market. With a global workforce, we’re remote‑first and grounded in a simple idea: software is a people business. Cribl is the place where curious, collaborative people can do their best work, grow fast, and bring their full selves to the herd.

Why You’ll Love This Role

You will help build Cribl Engineering’s AI platform and productivity rails: the shared systems, runtimes, and workflows that let engineers use autonomous AI across the software delivery lifecycle. Instead of one-off prompts or single-feature AI experiments, you will focus on intent engineering, orchestration and harnesses, and production agent infrastructure that other teams can trust and extend.

You will work with a small, high-impact team alongside tech leads, principals, and partner engineering groups. You will ship systems that plan, implement, review, test, and operate work with strong guardrails, then drive adoption until those tools become part of how Cribl builds software. You will also bring a builder mentality: use the product yourself to solve real team and Engineering problems, then turn that lived experience into better systems. Cribl strives to be a great place to work for everyone.


As An Active Member Of Our Team, You Will…

  • Design, build, and operate production agentic workflows and the platform harnesses they run on (orchestration, tool integrations, shared context, extension points for other teams)
  • Practice intent engineering: turn goals into clear specifications, rules, constraints, and acceptance criteria that AI systems can execute reliably
  • Stay current on AI tooling and practices, and evangelize what works across Engineering so partners adopt proven patterns
  • Engage hard in design discussions and healthy debate while the direction is open, then pivot with equal energy into implementation once the team decides
  • Build observability and guardrails: tracing, regression detection, human-in-the-loop controls, safe rollout, and operability for agentic systems others depend on
  • Own critical pieces of agent runtime: job isolation, scheduling, execution environments, secrets and access, and production operations in the cloud
  • Design and operate event-driven architectures where it fits (queues, streams, webhooks, async job fan-out) so agentic systems stay scalable and loosely coupled
  • Ship safely and often: keep agentic systems on automated CI/CD paths with progressive delivery and clear rollback 
  • Bring a builder mentality: dogfood what we ship. Use the platform to unblock yourself and the team, surface sharp edges, and turn real usage into product and platform improvements
  • Compress software delivery loops by improving how AI helps engineers write, test, review, debug, and validate changes in real repositories and pipelines
  • Partner across Engineering to understand workflows, ship tools that fit how people work, and drive adoption through playbooks, examples, demos, and enablement
  • Evaluate build vs. buy; stay current on models, agent frameworks, MCP-style tool protocols, and the broader AI tooling ecosystem
  • Define and track success metrics for adoption, quality, reliability, and satisfaction; use feedback and data to decide where to invest
  • Partner on security and data access so tools are useful while respecting permissions and company policy
  • Take designed projects from zero to production: own the path from agreed design through implementation, rollout, and day-two operability
  • This position may include stand-by, on-call, or off-hours duties for systems you own

If You’ve Got It - We Want It

  • Staff-level (or equivalent) professional software engineering experience building and operating production distributed systems
  • Strong TypeScript (and modern JavaScript) plus Node.js experience shipping production services; polyglot comfort is welcome, TypeScript is the primary stack for this team
  • Strong software engineering fundamentals and the ability to ship quickly: design, testing, debugging, APIs/services, and code quality
  • You have lived (or thrived in) a continuous deployment culture: automated pipelines, progressive delivery or safe rollout, monitoring and rollback, and a bias toward frequent, reversible production changes rather than infrequent manual releases
  • Hands-on fluency with modern LLM and agentic coding workflows in production or serious internal platforms (not curiosity-only)
  • Experience building backend services, integrations, automation, and internal tools; comfort across product surfaces when needed
  • Professional experience with agent orchestration, tool-calling systems, evaluation or guardrail techniques, and connecting agents to reliable backend systems
  • Familiarity with agent frameworks, orchestration layers, and integrating external tools and data sources into LLM-based systems (MCP or equivalent experience is a plus)
  • Experience with event-driven systems: queues, streams, pub/sub, webhooks, or similar async patterns in production
  • Observability fluency: metrics, logs, traces, and using them to operate and improve production systems
  • Clear communication, documentation, and teaching ability; comfort driving adoption, not only writing code
  • Good judgment around security, permissions, data access, and safe tool rollout
  • Ability to problem-solve from first principles, make sound trade-offs, and drive work independently through ambiguity
  • Nice to have
    • Hands-on Kubernetes in production
    • Terraform or similar infrastructure-as-code for cloud provisioning
    • Temporal or other workflow-platform experience
    • Deeper AWS / cloud-native ops fluency (IAM, networking, running production workloads end to end)

#LI-JB1
#LI-Remote

The salary for this role is dependent on geographic location and will be based on the individual candidate's job-related knowledge, skills, and experience.
In addition to base salary, for sales and some sales-adjacent roles, employees are eligible to earn incentive compensation (commission). For all other roles, employees are eligible to participate in the Cribl Corporate Bonus Program.
In addition to a competitive salary, Cribl also offers a generous benefits package which includes health, dental, vision, short-term disability, and life insurance, paid holidays and paid time off, a fertility treatment benefit, 401(k), and equity.

Base Salary Range
$210,000—$245,000 USD

Bring Your Whole Self

Diversity drives innovation, enables better decisions to support our customers, and inspires change for the better. We’re building a culture where differences are valued and welcomed, and we work together to bring out the best in each other. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, or any other applicable legally protected characteristics in the location in which the candidate is applying.

Interested in joining the Cribl herd? Learn more about the smartest, funniest, most passionate goats you’ll ever meet at cribl.io/about-us. 

Skills Required

  • Staff-level or equivalent professional software engineering experience building and operating production distributed systems
  • Strong TypeScript and modern JavaScript experience
  • Strong Node.js experience shipping production services
  • Strong software engineering fundamentals, including design, testing, debugging, APIs, services, and code quality
  • Experience working in a continuous deployment culture with automated pipelines, progressive delivery or safe rollout, monitoring, and rollback
  • Hands-on fluency with modern LLM and agentic coding workflows in production or serious internal platforms
  • Experience building backend services, integrations, automation, and internal tools
  • Professional experience with agent orchestration, tool-calling systems, evaluation or guardrail techniques, and reliable backend integrations
  • Familiarity with agent frameworks, orchestration layers, and integrating external tools and data sources into LLM-based systems
  • Experience with event-driven systems such as queues, streams, pub/sub, or webhooks in production
  • Observability fluency with metrics, logs, and traces
  • Clear communication, documentation, and teaching ability
  • Experience driving adoption of engineering tools and practices
  • Sound judgment around security, permissions, data access, and safe tool rollout
  • Ability to solve problems from first principles, make trade-offs, and work independently through ambiguity
  • Hands-on Kubernetes experience in production
  • Terraform or similar infrastructure-as-code experience for cloud provisioning
  • Temporal or other workflow-platform experience
  • Deeper AWS or cloud-native operations fluency, including IAM, networking, and production workload management

Cribl Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Cribl and has not been reviewed or approved by Cribl.

  • Affordable Benefits — Medical and dental premiums are fully covered for individuals in the U.S., with low costs for dependents, and the plans are described as low‑cost overall. This positions healthcare expenses favorably for many employees.
  • Leave & Time Off Breadth — Unlimited PTO, paid holidays, and periodic company “refresh” or winter‑break days provide ample time away. Flexible schedules further support taking time when needed.
  • Wellbeing & Lifestyle Benefits — A monthly stipend for home office, phone, and internet, plus strong remote‑work setup support, underpin the remote‑first model. Additional perks like recharge days and equipment support bolster day‑to‑day wellbeing.

Cribl Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: San Francisco, CA
1,000 Employees
Year Founded: 2018

What We Do

Cribl, the AI Platform for Telemetry, empowers enterprises to manage and analyze telemetry for both humans and agents. Trusted by organizations worldwide, including half of the Fortune 100, Cribl bridges the gap between AI ambition and infrastructure reality. No lock-in. No data loss. No compromises. Cribl’s vendor-agnostic platform ensures data remains portable and interoperable. By cost-effectively handling increasing data volume and variety without delay, Cribl gives enterprises the choice, control, and flexibility to build what’s next.

Why Work With Us

We are building the company that will become the industry leader in IT and Security data. But, doing that doesn’t mean we’re always serious. We approach our work fearlessly, learn quickly, improve constantly, and celebrate our wins at every turn. And more importantly, we laugh a lot.

Gallery

Gallery

Similar Jobs

Ellevation Education Logo Ellevation Education

Product Designer

Edtech • Social Impact • Software
In-Office or Remote
2 Locations
225 Employees
105K-145K Annually
Remote
USA
3000 Employees

Takeda Logo Takeda

Specialty Business Manager, Derm - Stamford, CT

Healthtech • Software • Analytics • Biotech • Pharmaceutical • Manufacturing
Remote or Hybrid
Connecticut, USA
50000 Employees
64-87 Hourly
Remote or Hybrid
2 Locations
289097 Employees

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
60 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account