Computer Vision Engineer

Posted 15 Hours Ago
Be an Early Applicant
Watertown, MA, USA
In-Office
100K-300K Annually
Entry level
Artificial Intelligence • Software
The Role
Build and deploy computer-vision perception systems for robot manipulation. Responsibilities include detection, segmentation, tracking, pose estimation, 3D scene understanding, camera calibration, dataset creation, evaluation, model selection, and diagnosing failures on physical robots. The role requires close collaboration across research, software, mechanical, and electrical engineering teams, plus maintaining reliable, measurable, and diagnosable perception systems.
Summary Generated by Built In

Tutor Intelligence builds AI robotics systems and deploys them into the facilities that need them. We believe research and deployment should improve each other: real-world operation reveals the problems worth solving, and better technology makes robots more useful. Join a team where your designs, experiments, and code have a direct impact on physical systems.

Our Culture

We value technical excellence, collaboration, and respect. Engineers take ownership, make tradeoffs explicit, and help colleagues do better work. We move quickly through building, testing, and learning from what happens on the robot.

About the Role

Build perception that makes robot manipulation experiments reliable. You will own a significant vision system from problem definition and data collection through evaluation, integration, and ongoing improvement on physical robots. The work combines camera geometry and learned models, with decisions grounded in measured performance.

You will work closely with research, software, mechanical, and electrical engineers to understand failures across sensing, algorithms, and the robot environment. This role focuses on perception and its integration into manipulation systems.

Responsibilities

    • Define the technical approach and success criteria for ambiguous perception problems in robot manipulation.
    • Build and integrate capabilities such as detection, segmentation, tracking, object pose estimation, and 3D scene understanding.
    • Own camera calibration, coordinate transforms, synchronization assumptions, and checks that expose degraded sensor quality.
    • Create reproducible datasets and evaluations covering changes in objects, lighting, viewpoint, and occlusion; prevent evaluation leakage.
    • Select and adapt geometric methods and learned models based on accuracy, latency, robustness, and operational constraints.
    • Deploy and maintain perception in the robot stack; diagnose failures using recorded data and physical experiments.
    • Improve shared evaluation tools and documentation, contribute to design reviews, and mentor colleagues in your domain.

Requirements

    • Evidence of independently owning a substantial computer-vision system from an ambiguous problem through tested integration.
    • Strong Python engineering and experience writing maintainable code, tests, and reproducible experiments.
    • Practical command of camera models, calibration, coordinate frames, and 3D geometry.
    • Experience with geometric vision or learned visual models, with the judgment to explain when each approach is appropriate.
    • Ability to build representative evaluations, isolate failure modes, and distinguish model improvements from data or measurement artifacts.
    • Ability to communicate technical tradeoffs and resolve interfaces with engineers in other disciplines.

Nice to have

    • Perception deployed on physical robots, especially manipulation systems.
    • Multi-camera or depth systems, sensor fusion, SLAM, or object pose estimation.
    • OpenCV, Open3D or PCL; a modern deep-learning framework; C++ or GPU optimization.
    • What Success Looks Like

      • Establish a reproducible baseline, meaningful evaluation, and documented sensing assumptions for an agreed perception problem.
      • Deliver a measured improvement that holds up on the physical robot and meets runtime constraints.
      • Keep the system diagnosable and usable by colleagues through reliable calibration checks, failure analysis, and documentation.

About our Roles & Titles

    At Tutor, we believe great engineers and researchers are defined by what they build and the impact they have — not where they sit in an org chart or what title they have. Therefore, everyone in our R&D org holds the title Member of Technical Staff (MoTS). Our job postings use standard titles so you can find us, but if you join Tutor, you'll be a MoTS — with a level that is determined through the interview process.

    That also means we hire people, not slots. Work at Tutor evolves every quarter, and we set the expectation of flexibility from day one — it's common for people to start on one thing and shift to another based on where the team needs them most. A high technical bar across the board is what makes that flexibility possible: it's what allows people to contribute meaningfully whatever problem they take on.

Skills Required

  • Independently own a substantial computer-vision system from an ambiguous problem through tested integration.
  • Strong Python engineering skills and experience writing maintainable code, tests, and reproducible experiments.
  • Practical command of camera models, calibration, coordinate frames, and 3D geometry.
  • Experience with geometric vision or learned visual models, including judgment about when each approach is appropriate.
  • Ability to build representative evaluations, isolate failure modes, and distinguish model improvements from data or measurement artifacts.
  • Ability to communicate technical tradeoffs and resolve interfaces with engineers in other disciplines.
  • Experience deploying perception systems on physical robots, especially manipulation systems.
  • Experience with multi-camera or depth systems, sensor fusion, SLAM, or object pose estimation.
  • Experience with OpenCV, Open3D, or PCL and a modern deep-learning framework.
  • Experience with C++ or GPU optimization.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Boston, MA
22 Employees

What We Do

Tutor Intelligence is a full-service robotics and automation provider built to serve contract packagers. We partner with the world's largest 3PLs to automate what has stumped the industry for decades: short-run packaging and post-packaging where SKUs, patterns, orders, and volumes are constantly changing

Similar Jobs

Symbotic Logo Symbotic

Staff Software Engineer

Artificial Intelligence • Machine Learning • Robotics • Automation
Hybrid
Wilmington, MA, USA
3500 Employees
149K-223K Annually

Berkshire Grey Logo Berkshire Grey

Senior Machine Learning Engineer

Artificial Intelligence • Robotics • Business Intelligence
In-Office
Bedford, MA, USA
289 Employees
In-Office or Remote
2 Locations
222 Employees

Tycho.AI Logo Tycho.AI

Machine Learning Engineer

Aerospace • Artificial Intelligence • Robotics • Defense
Hybrid
Cambridge, MA, USA
26 Employees

Similar Companies Hiring

Revel Thumbnail
Aerospace • Hardware • Robotics • Software
Marina Del Rey, California
70 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees
Vega Thumbnail
Artificial Intelligence • Automotive • Insurance • Transportation
US
43 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account