Senior Software Engineer

Posted Yesterday
Hiring Remotely in United States
Remote
Senior level
Artificial Intelligence • Cloud • Software • Consulting
The Role
Build and evolve Stratusphere, an agentic cloud optimization platform spanning web applications, backend services, infrastructure, data pipelines, and agent tooling. Review and improve AI-generated code, architecture, testing, reliability, observability, and safety. Collaborate with engineering, product, customer success, and customers to deliver measurable solutions, improve agent workflows, and influence the roadmap. Own work end to end while maintaining documentation, change control, communication, and operational standards.
Summary Generated by Built In

About StratusGrid

StratusGrid is building Stratusphere™, a multi-agent platform that turns cloud complexity into measurable outcomes. Stratusphere coordinates specialized infrastructure agents to observe AWS and Azure environments, simulate policy-safe plans, and execute approved changes with auditability and rollback-ready safety, so savings compound over time, security improves, delivery accelerates, and teams spend less time on toil and fire drills.

We’re a team of builders and operators who care about trustworthy automation and real business impact. Our work sits at the intersection of cloud engineering, product, and customer outcomes, helping customers make infrastructure changes they can measure, explain, and stand behind.

About the Position:

StratusGrid delivers a high-trust, high-competence customer experience around cloud optimization, and we’re productizing that expertise into our agentic platform, Stratusphere. This role is for a versatile software engineer who can help build the core software platform, the backend infrastructure running agents, and the agents and tooling that discovers, contextualizes, plans, and executes work for our customers.

You will work closely with our customer success and human in the loop teams, as well as directly with customers, to rapidly test, implement, and mature solutions that make our system increasingly autonomous. Problems we are solving range from the discovery and contextualization of complex cost inefficiencies in cloud infrastructure, to designing and implementing web interfaces that enable customers, to making agents and tooling that can plan and execute changes inside of complex cloud environments. To do this, you will be using agent-first development practices in a product that was designed to be developed agent first, bringing critical judgement, maintainability, and design expertise to the rapid AI development cycle.

Responsibilities:

  • Agent-First Software Development: Actively utilize and oversee AI agents to rapidly test, implement, and mature software solutions across the web platform, backend services, infrastructure, data pipelines, and agent tooling that make up the Stratusphere platform.

  • Engineering Judgment & Scrutiny: Serve as the ultimate source of judgment for AI-generated plans and code contributed to the Stratusphere platform. Scrutinize all agent-driven implementations to ensure they are rigorously tested, scalable, maintainable, and meet StratusGrid's engineering standards.

  • Add to the Foundation: Stay current with the latest methods and best practices for improving velocity, quality, and maintainability with agent-first development approaches. Share knowledge and actively contribute to maturing processes, standards, and tooling to which makes the whole team better.

  • Collaborative System Design: Partner closely with engineering, product, and customer-facing teams to design and evolve the Stratusphere platform, ensuring technical decisions align with strategic goals and customer needs.

  • Agent Improvement Feedback Loop: Improve agentic architecture, tooling, and prompts to solve for patterns in agent errors and blind spots (e.g., missing context, risky sequencing, unclear rollback, incomplete stakeholder info). Propose rubric changes, training examples, and workflow improvements; partner with Product/Human in the Loop teams to measurably improve quality and safety.

  • Cross-Functional Collaboration: Work in high-visibility channels with Engineering, Product, and Customer teams; share context broadly; ask questions early; and support a safe culture where we learn fast without introducing risk through isolation.

  • Execution with Reliability & Urgency: Own work end-to-end with a strong sense of responsibility. Capture commitments, hit deadlines, communicate status proactively, escalate early, and close the loop visibly—no surprises.

  • Product Partnership & Roadmap Input: Partner with Product to test solutions to customer problems and recurring friction, iterate from learnings, and implement features which are based on data and proven results.

  • Operational Excellence & Standards: Follow StratusGrid customer experience standards, change control processes, documentation expectations, and work-system hygiene to ensure consistency, traceability, and scalability.

Required Experience:

  • Multi-Language Proficiency: Deep expertise in at least one language (preferably Typescript) combined with proficiency in overseeing agentic delivery of solutions in other languages. Ability to quickly adapt to new languages/frameworks as required.

  • Skillful Tester: Skilled in architecting services to allow for layered testing, and then consistently using testing to enforce quality and maintainability while rapidly developing solutions with AI Agents.

  • Frontend Experience: Experience contributing to frontend applications, especially NextJS or other React frameworks.

  • Familiar with Coding Agents: Familiar with, and excited about using, coding agents such as Claude Code, Codex, and Gemini CLI to rapidly plan, write, and test high quality and maintainable code.

  • Multifaceted Database Expertise: Experience querying large datasets with tools like AWS Athena or DuckDB, creating applications backed by relational databases like

  • Postgres, and experience working with NoSQL solutions like AWS DynamoDB.

  • CICD Experience: Experience building, maintaining, and using tools like github actions to test code and catch quality issues early.

  • AWS and IaC Expertise (AWS): Proven ability to build and change infrastructure in production AWS environments using infrastructure code, including container environments and cloud native services like lambda, S3, and cloudfront.

  • Observability: Experience integrating observability tooling into software and using it to quickly troubleshoot complex problems and improve system performance.

  • Excellent Communication: Exceptional written and verbal communication, able to translate complex technical topics into decision-ready narratives for technical teammates and business stakeholders, with clarity, completeness, and empathy. Ability to listen to customer concerns and turn them into actionable solutions.

  • Problem Solving Mindset: Proven ability to work through ambiguity, identify problems to solve to reach goals, and test and implement solutions to novel problems.

  • Strong Ownership & Urgency: Track record of meeting commitments, proactively communicating status/risks, escalating early, and maintaining high standards of reliability and follow-through.

  • AI as a Daily Tool: Demonstrated habit of using AI tools to deliver high quality work at high velocity while maintaining rigorous quality, security awareness, and sound judgment.

  • Travel: Willingness and ability to travel periodically (as needed) for team planning sessions, or onsite work.

  • Remote-Work-Ready: Equipped to work effectively in a distributed team environment, including a reliable high-speed internet connection, a professional and distraction-limited workspace, and the ability to consistently communicate, collaborate, and execute independently.

  • Authorized to work: Applicants must be legally authorized to work in the United States at the time of hire. This position does not offer visa sponsorship.

Nice-to-have / Differentiators:

  • Located in the Chattanooga, Nashville, Knoxville, or Atlanta Metro area and able to come into the Chattanooga HQ office for coworking on a regular basis.

  • Experience working with iceberg or other datalake formats and tools like DBT to build high quality and cost effective data pipelines creating unique datasets from mulitple sources. and policy-as-code exposure (OPA/Sentinel/Azure Policy)

  • Strong grasp of DevOps principals and operational tooling, practices, and processes used by Cloud Infrastructure teams.

  • Strong working knowledge of cloud IAM/RBAC and least-privilege operations: ○ AWS IAM roles, cross-account access, SCPs, permission boundaries ○ Azure Entra ID, RBAC, PIM, Azure Policy

  • Ability to work effectively within constrained access and compliance requirements.

Equal Opportunity Statement

Siftwell celebrates diversity and is proud to be an Equal Opportunity Employer. We consider all qualified applicants without regard to race, color, religion, sex, national origin, disability status, veteran status, or any other characteristic protected by law.

Qualified applicants with arrest or conviction records will be considered in accordance with applicable laws, including the California Fair Chance Act.

ADA Accommodation Statement

If you require a reasonable accommodation during a selection process, please let us know through our general contact form or indicate this in your application.

#StratusGrid

Skills Required

  • Deep expertise in at least one programming language, preferably TypeScript, with proficiency overseeing agentic delivery in other languages
  • Ability to quickly adapt to new programming languages and frameworks
  • Experience architecting services for layered testing and consistently using testing to enforce quality and maintainability
  • Frontend application experience, especially with Next.js or other React frameworks
  • Familiarity with coding agents such as Claude Code, Codex, and Gemini CLI
  • Experience querying large datasets with AWS Athena or DuckDB
  • Experience building applications backed by relational databases such as PostgreSQL
  • Experience with NoSQL solutions such as AWS DynamoDB
  • Experience building, maintaining, and using CI/CD tools such as GitHub Actions
  • Ability to build and change production infrastructure in AWS using infrastructure as code, including container environments and services such as Lambda, S3, and CloudFront
  • Experience integrating observability tooling into software and using it to troubleshoot problems and improve performance
  • Excellent written and verbal communication skills, including translating technical topics into decision-ready narratives
  • Ability to work through ambiguity, identify problems, and implement solutions to novel problems
  • Strong ownership, urgency, reliability, proactive communication, and follow-through
  • Demonstrated daily use of AI tools while maintaining quality, security awareness, and sound judgment
  • Willingness and ability to travel periodically for team planning or onsite work
  • Ability to work effectively in a distributed remote team with reliable high-speed internet and a professional workspace
  • Legally authorized to work in the United States at the time of hire
  • Experience with Iceberg or other data lake formats and tools such as dbt
  • Exposure to policy-as-code tools such as OPA, Sentinel, or Azure Policy
  • Strong understanding of DevOps principles and cloud infrastructure operational tooling and practices
  • Working knowledge of cloud IAM/RBAC and least-privilege operations across AWS and Azure
  • Ability to work effectively within constrained access and compliance requirements
  • Located in the Chattanooga, Nashville, Knoxville, or Atlanta metropolitan area and able to regularly cowork at the Chattanooga headquarters
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
30 Employees
Year Founded: 2013

What We Do

StratusGrid is a cloud engineering and software company building Stratusphere™, a multi-agent platform for managing complex AWS and Azure infrastructure. The platform observes environments, simulates policy-safe plans, and executes approved changes with auditability and rollback readiness. Its solutions aim to reduce cloud costs and operational toil, strengthen security, accelerate delivery, and help engineering teams achieve measurable, explainable infrastructure outcomes over time.

Similar Jobs

Samsara Logo Samsara

Senior Software Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
6 Locations
4000 Employees
155K-260K Annually

Coinbase Logo Coinbase

Senior Software Engineer

Artificial Intelligence • Blockchain • Fintech • Financial Services • Cryptocurrency • NFT • Web3
Easy Apply
Remote
USA
4700 Employees
186K-219K Annually

Optum Logo Optum

Senior Software Engineer

Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Remote or Hybrid
Minnetonka, MN, USA
160000 Employees
92K-164K Annually

NBCUniversal Logo NBCUniversal

Senior Software Engineer

AdTech • Cloud • Digital Media • Information Technology • News + Entertainment • App development
Remote or Hybrid
New York, NY, USA
120K-170K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account