Help us build reliable production software that keeps our hardware moving through manufacturing.
MatrixSpace develops AI-enabled radar and sensing systems that help people understand what's happening in the world around them. By combining advanced radar, edge computing, and AI, we deliver situational awareness in environments where traditional sensing solutions struggle.
We're looking for a Senior Production Test Software Engineer to own and improve the software we use to configure, provision, and test our products in production, with a focus on improving first-pass yield, reducing test cycle time, increasing test reliability, and lowering production test cost. You'll work across software, automation, infrastructure, and connected hardware, partnering with manufacturing and engineering teams to solve problems quickly and make our production systems more reliable and efficient.
If you're hands-on, methodical, and energized by solving problems across software and hardware, we'd love to talk.
THIS IS NOT A FULLY REMOTE ROLE
What You'll Do
- Own and improve the software that supports production testing and product provisioning, and serve as a key escalation point when software issues affect production.
- Implement automated production data collection for tracking quantitative production test results and trends including first-pass yield, retest rate, test cycle time, data distributions, station efficiency, and recurring failure modes, and utilize this data to drive continuous improvement of process performance metrics.
- Develop the strategic roadmap for the provisioning and end-of-line test process, including managing capacity, scalability and test coverage.
- Troubleshoot problems across applications, automation, Linux systems, networks, containers, and connected hardware to find the root cause and get production moving again.
- Build and improve Python-based provisioning and diagnostic tools, backend services and APIs, and operator-facing workflows.
- Maintain the automation and containerized services behind the system, including deployments, configuration, monitoring, and releases.
- Make the system easier to use and recover when things go wrong through clearer errors, better diagnostics, automated recovery, and practical troubleshooting guidance.
- Bring new products, firmware, configuration steps, and tests into production in partnership with product and test engineering.
- Use production data, recurring failures, and operator feedback to prioritize improvements, prevent repeat issues, and keep documentation current.
What We're Looking For
This position requires working directly or indirectly with the US Government in restricted environments. Candidates must be legally authorized to work in the United States without employer sponsorship and may be required to obtain and maintain a U.S. government security clearance in the future.
Required
- 8+ years of professional experience in software engineering, site reliability, DevOps, systems integration, test automation, or a closely related field.
- 3+ years building or supporting Python-based automation that connects with devices, command-line tools, APIs, or other systems.
- 2+ years running containerized services or production systems using Docker and Kubernetes or similar technologies.
- Experience building or supporting backend services and APIs. Go experience is helpful, but you should be comfortable learning it if it's new to you.
- Strong Linux and networking troubleshooting skills, including services, logs, connectivity, remote systems, cloud deployments and networked devices.
- Experience with infrastructure automation tools such as Ansible or AWX, plus Git-based version control and reliable release practices.
- Experience with manufacturing test, factory automation, product provisioning, embedded systems, robotics, radar, networking equipment, or other connected hardware.
- Experience troubleshooting systems that combine software and hardware, working directly with engineers and production users to diagnose problems, restore operation, and prevent repeat failures.
- Ability to work onsite in Hoffman Estates, IL and travel approximately 10%, including initial travel to Burlington, MA for training and knowledge transfer.
Someone Who Will Thrive in This Role
- Self-starter willing to lead and drive resolution of issues requiring collaboration across engineering, radar software and AI teams.
- Takes ownership of a problem and stays with it until it's understood, fixed, and documented.
- Enjoys engaging across software, infrastructure, networks, and hardware to understand what's really going wrong.
- Knows when production needs a fast fix and when a problem calls for a longer-term solution.
- Makes sound technical decisions, shares clear recommendations, and brings in the right people when a problem crosses team boundaries.
- Spots patterns in failures and production data and turns them into better tools and more reliable systems.
Bonus Points
- Hands-on experience with Go, Next.js, React, JavaScript or TypeScript, Kubernetes, Ansible, or AWX.
- Experience building monitoring, health checks, automated recovery, or diagnostics for production systems.
- Familiarity with firmware deployment, device configuration, credential management, sensors, or automated functional testing.
- Experience taking over an existing production system and improving it without disrupting the people who depend on it.
At MatrixSpace, this software has a direct impact on how efficiently we build and test our products. You'll solve real production problems, improve the systems our teams use every day, and help us scale manufacturing as we grow. If that sounds exciting, we'd love to hear from you.
Skills Required
- 8+ years of professional experience in software engineering, site reliability, DevOps, systems integration, test automation, or a closely related field
- 3+ years building or supporting Python-based automation connected to devices, command-line tools, APIs, or other systems
- 2+ years running containerized services or production systems using Docker and Kubernetes or similar technologies
- Experience building or supporting backend services and APIs
- Strong Linux and networking troubleshooting skills, including services, logs, connectivity, remote systems, cloud deployments, and networked devices
- Experience with infrastructure automation tools such as Ansible or AWX, Git-based version control, and reliable release practices
- Experience with manufacturing test, factory automation, product provisioning, embedded systems, robotics, radar, networking equipment, or other connected hardware
- Experience troubleshooting systems combining software and hardware while working with engineers and production users
- Ability to work onsite in Hoffman Estates, Illinois
- Ability to travel approximately 10%, including initial travel to Burlington, Massachusetts for training and knowledge transfer
- Go experience
- Experience with Next.js, React, JavaScript, TypeScript, Kubernetes, Ansible, or AWX
- Experience building monitoring, health checks, automated recovery, or production diagnostics
- Familiarity with firmware deployment, device configuration, credential management, sensors, or automated functional testing
- Experience improving an existing production system without disrupting users
What We Do
We're a technology leader in Autonomous Flight Systems. Our products fly themselves, incorporate cutting edge AI, and sense what's around them. This is a place for people that have the intellectual talent and tenacity to invent new things and get them to work. It's demanding but a lot of fun.
Why Work With Us
Our culture is for innovators who want to build real systems and see them work. Our team of engineers and scientists come from many different disciplines, and our technology is truly cutting edge. This is an elite team for the curious, talented, and hard working professionals that want to make history with their products.
Gallery








