Data Engineer
Job description
About the role
SteerBridge is a modern technology company delivering innovative, mission‑focused solutions to the U.S. Government and private sector. Leveraging deep expertise in federal acquisition, digital transformation, and emerging technologies, we deliver agile, commercial‑grade capabilities that accelerate operational effectiveness and drive measurable mission success. At the core of SteerBridge is our people - especially the veterans whose leadership, problem‑solving mindset, and commitment to excellence elevate every project we support. We don't simply hire exceptional talent; we cultivate it, creating meaningful career pathways for veterans, military spouses, and professionals who share our passion for advancing technology and strengthening the missions we serve. You will own the design and execution of data infrastructure that directly enables high‑impact defense and aerospace analytics. You will be responsible for ensuring data reliability, performance, and security across cloud platforms used by mission partners. This role grants you ownership of end‑to‑end pipeline development from ingestion to insight delivery. You will mentor team members and operational users on best practices for data usage and integrity. Your work will directly support decision‑making processes that affect operational readiness and logistics. You will collaborate closely with data scientists, analysts, and government stakeholders to align technical solutions with mission needs.
Key facts
What you'll do
- Lead the design, development, and maintenance of AWS‑based ETL/ELT pipelines using AWS Glue, Amazon Redshift, and Amazon S3 with orchestration via Apache Airflow, dbt, or Prefect.
- Utilize Azure Databricks (Spark) and Azure Data Factory to manage and schedule data pipelines and workflows in hybrid cloud environments.
- Build and optimize data models in cloud data warehouses such as Snowflake, BigQuery, or Redshift, and maintain structured data stores in S3 buckets and Blob storage.
- Integrate data from diverse sources including REST and SOAP APIs, event streams such as Kafka, relational and non‑relational databases, and multiple SaaS platforms.
- Create, index, query, and update SQL tables and servers; write and maintain Python and/or JavaScript code to parse, transform, and validate complex data structures.
- Monitor pipeline health, enforce rigorous data quality controls including schema validation, null checks, and duplicate detection, and troubleshoot data issues with observability best practices.
- Develop and implement data acquisition, quality assurance, and management protocols, and document all data collection, cleaning, and analysis activities for both internal and external users.
- Use schedulers and APIs to obtain near real‑time data, and automate workflows and processes using Python or other scripting languages to reduce manual effort and increase reliability.
- Partner with data scientists and analysts to deliver clean, well‑documented datasets and data products, and provide active support to the data science team to ensure timely delivery of analytical outcomes.
- Contribute to data governance standards by implementing lineage tracking, data cataloging, and access control mechanisms across platforms.
- Mentor and collaborate with Marines at the squadron level to improve data entry and indexing practices, translating technical concepts into actionable guidance for operational users.
- Assist with the maintenance and development of internal analytics data architecture, ensuring alignment with evolving mission requirements and data strategies.
- Design, write, and disseminate innovative and visually appealing reports for diverse audiences, using clear narratives and effective visualizations to communicate insights.
- Participate actively in code reviews, architecture discussions, and cross‑functional planning sessions to promote engineering excellence and shared learning.
- Evaluate and recommend new tools and technologies to improve the data platform, scalability, and performance in support of mission objectives.
Requirements
- 3-5 years of professional experience in data engineering or a closely related role demonstrating delivery of production‑grade data solutions.
- Bachelor's Degree in Computer Science or a related field; alternatively, three (3) years of additional relevant experience may substitute for the degree, with a minimum of six years total experience without a degree.
- U.S. Citizenship is mandatory for this position due to the nature of the mission and data sensitivity.
- An active security clearance or the ability to obtain one is required before beginning work.
- Demonstrated strong proficiency in Python, including frameworks such as PySpark and pandas, and advanced SQL skills covering CTEs, window functions, and complex joins.
- Hands‑on experience with at least one major cloud data warehouse such as Snowflake, BigQuery, or Redshift, and with object storage in S3 or equivalent blob stores.
- Practical experience building and maintaining ETL/ELT pipelines in cloud environments using tools such as AWS Glue, Azure Data Factory, Apache Airflow, dbt, or Prefect.
- Familiarity with REST and SOAP APIs, event streaming platforms like Kafka, and interaction with both relational and non‑relational databases.
- Ability to write and maintain code in scripting languages such as Python and JavaScript for data parsing, transformation, and automation tasks.
- Experience with data quality frameworks, monitoring, and alerting practices to ensure pipeline reliability and data integrity.
- Understanding of data governance concepts including lineage, cataloging, and access control, with exposure to implementing these in enterprise settings.
- Willingness to mentor operational users, including Marines at the squadron level, and to improve data entry and indexing practices through clear, practical guidance.
- Capability to work on‑site within existing systems of record that host multiple databases and data stores in a secure government environment.
- Commitment to continuous improvement, collaboration, and adherence to strict timelines in a mission‑focused setting.
Nice to have
- Experience with the National Center for Manufacturing Sciences (NCMS) or similar collaborative defense programs.
- Familiarity with F-35 logistics, spares, or aviation supply chain datasets.
- Background in coaching or mentoring personnel in technical environments, particularly within military or veteran contexts.
Practical notes
This role is based in Miramar, California and requires full‑time on‑site engagement. U.S. citizenship and the ability to obtain a security clearance are required. Candidates must be able to work on‑site within the existing systems of record without remote or hybrid alternatives. No travel, visa, or application deadlines are specified in this posting.