Associate Data Scientist
Job description
About the role
Hayden AI is seeking an Associate Data Scientist to join their team at their San Francisco HQ Office. This on-site position offers an exciting opportunity for a candidate interested in leveraging data to improve transit systems and government agency operations. As part of the team, you will support a variety of stakeholders by delivering high-quality data insights, developing dashboards, and ensuring data accuracy. You will work at the intersection of data engineering, analytics, and data science, collaborating with teams across the organization to solve real-world challenges related to transportation safety, efficiency, and sustainability. The role involves translating complex business questions into technical data requirements, creating impactful visualizations, and maintaining data quality standards to support Hayden AI's mission of transforming transit through computer vision technology.
Key facts
What you'll do
- Develop and enhance standardized metrics from foundational datasets using dbt models and AWS Glue jobs, ensuring data consistency and reliability across various projects.
- Create engaging and insightful data stories, visualizations, and dashboards based on stakeholder feedback and user experience insights, enabling better decision-making.
- Respond promptly to ad hoc data requests from cross-functional teams, including impact analyses, anomaly investigations, and root cause analyses, to support operational and strategic initiatives.
- Act as the first responder for data discrepancy issues and data freshness concerns, troubleshooting problems and coordinating fixes to maintain high data quality standards.
- Collaborate with business teams to translate their questions into precise data requirements, acting as a bridge between customer-facing teams and the data engineering and science teams.
- Monitor data quality and completeness throughout the data processing pipeline across multiple fleets, identifying issues early and driving corrective actions to ensure data integrity.
- Support data scientists by preparing datasets, performing exploratory data analysis, and providing review and feedback on analytical models and insights.
- Clearly communicate findings, insights, and recommendations to both technical and non-technical stakeholders, ensuring understanding and facilitating data-driven decision-making.
- Assist in maintaining and improving existing data pipelines, ensuring scalability and robustness for future data needs.
- Work closely with engineering teams to optimize data workflows and implement best practices for data management and transformation.
- Contribute to documentation of data processes, models, and dashboards to support team knowledge sharing and onboarding of new team members.
- Stay updated on industry best practices in data analytics, visualization, and data pipeline management to continuously improve data solutions.
Requirements
- Master's degree in Data Science, Statistics, Computer Science, Economics, Transportation Engineering, or a related field.
- At least 6 months of experience in data science, data projects, or internships, demonstrating practical skills in data analysis and visualization.
- Strong SQL skills and experience working with data warehouse systems such as Amazon Redshift, Google BigQuery, or Snowflake.
- Proficiency in Python for data manipulation and analysis, including libraries such as Pandas and NumPy.
- Hands-on experience with data transformation and pipeline tools like dbt and AWS Glue, or similar platforms.
- Experience building and maintaining dashboards in BI tools such as Tableau or Looker, with an understanding of best practices for data visualization.
- Solid understanding of descriptive statistics and the ability to interpret analytical results accurately.
- Excellent communication skills, capable of presenting complex data insights clearly to both technical and non-technical audiences.
- Ability to work collaboratively in a team environment, managing multiple priorities and deadlines effectively.
- Strong problem-solving skills and attention to detail, especially in troubleshooting data discrepancies and pipeline issues.
Nice to have
- Prior experience working with geospatial analytics and spatial datasets, such as GIS data or spatial coordinates.
- Familiarity with large-scale time-series and mobility datasets, including GTFS, GPS traces, and transit logs.
- Knowledge of operational monitoring tools like Grafana or similar platforms.
- Exposure to cloud platforms, especially AWS, and understanding of cloud-based data workflows.
- Experience working in a startup environment, demonstrating adaptability and a desire to make a significant impact.
- Knowledge of computer vision applications and how they integrate with data analytics workflows.
- Experience working with cross-disciplinary teams, including engineers, product managers, and customer success teams.
Skills & tools
- Python
- AWS
- Computer Vision
- AI
- Mobile
- Data Engineering
- Data Science
- SQL
- Snowflake
- BigQuery
- dbt
- Tableau
- Looker
- UX
- Finance
- Support
- Customer Success
- Engineering
Practical notes
This position is based on-site at Hayden AI's San Francisco headquarters. Candidates should be prepared to work in the office environment and collaborate closely with team members across departments. The role involves engaging with multiple stakeholders, including transit agencies, customer success, finance, and engineering teams, to deliver actionable insights. Prior experience with geospatial data or large-scale mobility datasets is advantageous but not mandatory. The candidate should have a strong interest in transportation, computer vision, and data-driven solutions to support Hayden AI's mission. The position offers an opportunity to contribute to impactful projects that improve street safety, transit efficiency, and sustainability through innovative data analytics and visualization.