Sr. Software Engineer, Perception Data Infrastructure
Job description
About the role
This position focuses on building the systems that manage and process massive amounts of perception data for autonomous driving technology. You will design and maintain the infrastructure required to train, validate, and deploy the software that allows vehicles to understand their surroundings. In this capacity, you will own the end-to-end lifecycle of critical data systems that directly influence the perception models powering self-driving operations. You will be responsible for ensuring that data flows seamlessly from collection vehicles into the training environments used by machine learning teams. The role requires a deep commitment to building robust, scalable solutions that can handle the evolving needs of autonomous development. You will partner closely with software and machine learning engineers to translate high-level requirements into performant data infrastructure. Ultimately, your work will ensure that the perception systems are trained on high-quality, well-managed data throughout the entire development cycle.
Key facts
What you'll do
- Architect and scale data pipelines that handle large-scale perception datasets across distributed environments.
- Improve the efficiency of data storage, retrieval, and processing workflows to reduce latency and operational costs.
- Collaborate with machine learning engineers to optimize data availability for model training and accelerate iteration cycles.
- Build tools that monitor data quality and infrastructure performance to ensure reliability and consistency in production.
- Design and implement data ingestion frameworks that support the high-volume inputs from autonomous vehicle sensors.
- Develop automation for data versioning and labeling to maintain traceability and reproducibility in model training.
- Partner with infrastructure teams to integrate data processing workflows with cloud-based compute resources.
- Create interfaces that enable data scientists to easily access and query large perception datasets for analysis.
- Implement logging and alerting mechanisms to proactively identify bottlenecks and failures in the data pipeline.
- Contribute to the definition of standards for data formats, storage layouts, and access patterns across the team.
- Optimize data throughput from edge collection systems to centralized training clusters.
- Support the deployment of perception models by ensuring data infrastructure meets the requirements for validation and testing.
- Evaluate new data processing technologies and frameworks to improve scalability and resilience over time.
- Act as a technical lead in designing solutions that balance performance, maintainability, and operational simplicity.
Requirements
- Hold a Bachelor degree or higher in Computer Science, Engineering, or a related technical field.
- Bring professional experience building large-scale distributed systems or data infrastructure for production environments.
- Demonstrate proficiency in C++ or Python for developing high-performance data processing components.
- Show a strong understanding of data storage systems and cloud infrastructure used in scalable applications.
- Have a proven track record of writing clean, maintainable code that supports long-term infrastructure goals.
- Exhibit strong problem-solving skills when dealing with complex data flows and system interactions.
- Possess the ability to work effectively in a fast-paced engineering environment with evolving priorities.
- Display solid communication skills to collaborate with cross-functional teams on technical decisions.
- Understand the fundamentals of sensor data types and the challenges associated with processing perception inputs.
- Have experience with version control systems and collaborative development practices.
- Be comfortable debugging issues in distributed data pipelines under tight operational constraints.
- Demonstrate knowledge of infrastructure principles such as scalability, fault tolerance, and observability.
- Commit to following engineering best practices for code reviews, testing, and documentation.
Nice to have
- Experience working with autonomous vehicle perception stacks and the associated data formats.
- Background in high-performance computing or big data frameworks such as Hadoop or Spark.
- Familiarity with GPU-accelerated data processing for training and inference workflows.
- Prior work in environments that require strict data governance and compliance standards.
- Exposure to containerization and orchestration tools used in modern data platforms.
Practical notes
- Nuro is currently focused on licensing its autonomous driving software for robotaxis and personally owned vehicles.
- The company recently secured 106 million dollars in funding to support this strategic shift.
- This is a full-time position based in Mountain View, California.
- Candidates must be eligible to work in the United States without sponsorship for this role.
- No relocation assistance or visa sponsorship is provided for this position.
- Applicants are encouraged to submit their applications promptly to be considered for the role.