Senior Systems Engineer, Diagnostics and System Behavior Analysis
Atoms
Job at a glance
Who we are Atoms is building the machines that power the next era of progress. Over the last decade, software has transformed the digital world. But the physical world, where food is made, minerals are mined, goods are moved, and industries are run, remains far less intelligent, far less efficient, and far more constrained. We’re changing that. Atoms builds Physical AI— real-world robots for the industries that move civilization forward, starting with food, mining, and transport.
Our systems are designed to understand, predict, and control the real world with precision, turning complex physical operations into something more reliable, more scalable, and more productive. This work requires more than robotics. It requires deep integration across hardware, software, AI, operations, manufacturing, and real estate. We don’t just build machines in a lab. We deploy them into real environments, operate them, learn from them, and improve them until they work at scale.
We are roboticists, engineers, operators, and builders. We believe the next great technology companies will not only transform information, but the physical systems that shape everyday life. If you want to work on hard problems with real-world impact, join us. About the role We are seeking a Senior Systems Engineer to own diagnostics and failure analysis across our development fleet. In this role, you will take a failure from a log to an assigned root cause, owning the triage classification, the severity model, and the diagnostic tooling that determines how quickly we know what failed.
You will build the diagnostic coverage and automated classification that turns failure analysis from an engineering activity into a process that runs such that only new failures reach a person. The role spans multiple vehicle platforms with different hardware, compute, and software, and requires regular time in the garage and around vehicles. What you’ll do Triage. Classify failures as software, firmware, hardware, data path, or system-level regressions, and route them with evidence.
Root cause analysis. Investigate novel failures using logs, telemetry, bus traces, and video, and establish the root cause. Diagnostic tooling. Build the parsing, correlation, and visualization that make large log datasets tractable for engineers outside this role. Automated classification. Build the frameworks that detect known failure signatures automatically, so recurrence is caught by software.
Data path integrity. Own monitoring for the logging path itself: dropped frames, timestamp drift, bandwidth and storage limits, and the alerting that catches a vehicle producing unusable data the same day. Severity and trend analysis. Categorize severity consistently and identify the failure clusters that should set engineering priority. Documentation. Set the standard for what a complete root cause analysis looks like here, and author the troubleshooting guides.
What we’re looking for Bachelor's or master's degree in Computer Science, Computer Engineering, Electrical Engineering, Robotics, or a related field required. 6+ years in systems integration, diagnostics, or hardware and software interface engineering on vehicles. L4 program or production ADAS experience. A track record of diagnosing failures that spanned hardware and software, with specific examples where you established the root cause.
Ability to correlate hardware symptoms and video evidence with telemetry and bus data into a concrete timeline. Sensor-level troubleshooting depth on cameras, lidar, radar, GNSS and IMU, including timing failure modes. Strong working knowledge of vehicle networks and diagnostics, including CAN, CAN FD, Automotive Ethernet, and DTCs, and how failures present at each layer. Python and command line log analysis at the level of building production tools Automated triage or anomaly detection built over fleet data.
Experience with monitoring and observability platforms and with log management at fleet scale. Working experience wi