Senior Data Engineer
Job description
About the role
The owns the design and execution of the data infrastructure that underpins our active grid response platform. You will architect and build high-throughput data pipelines that process continuous streams of sensor and telemetry readings into structured, reliable datasets. This role requires you to ensure that raw grid signals are converted into timely and accurate information that supports grid reliability and safety decisions. You will establish the ingestion frameworks that allow diverse telemetry formats to be consumed without interruption by monitoring applications. A core responsibility involves creating transformation logic that enriches and validates data to enable precise fault identification and prediction. You will also define and enforce data model standards so that downstream teams can trust the definitions and lineage of every dataset. Collaboration with reliability engineers and analytics groups will guide schema evolution and ensure that data products meet operational needs. You will implement monitoring and verification practices that maintain data quality as system volume and complexity continue to scale.
Key facts
What you'll do
- Execute the technical scope for the Senior Data Engineer role as defined by Gridware's active grid response mission.
- Develop ingestion pipelines that validate and normalize diverse sensor telemetry formats before they are stored in the warehouse.
- Construct scalable data processing workflows capable of handling high-velocity streams essential for real-time grid monitoring.
- Transform raw measurements into curated datasets that power accurate fault detection and system reliability analytics.
- Record data models and definitions to ensure clarity and consistency for teams consuming grid information.
- Partner with analytics and reliability teams to align data warehouse structures with reporting, alerting, and operational needs.
- Verify integration points between ingestion, processing, and visualization layers to confirm that dashboards reflect grid status correctly.
- Guide long-term schema and architecture decisions to support new grid devices and measurement types without service disruption.
- Implement data quality checks and testing procedures that safeguard integrity as data volumes and pipeline complexity increase.
- Coordinate change management processes so that updates to data pipelines meet reliability, safety, and compliance standards.
- Optimize query performance and data access patterns to enable rapid insights for grid operations teams.
- Maintain documentation that explains data lineage, business logic, and configuration for critical datasets used in active grid response.
- Support the deployment of data infrastructure components in cloud environments where grid information is stored and processed.
- Continuously assess tools, formats, and protocols used for telemetry to ensure compatibility with evolving grid monitoring requirements.
Requirements
- Bring three to five years of professional experience designing and managing complex data systems in production environments.
- Demonstrate advanced proficiency in SQL across varied datasets, including time-oriented and high-cardinality telemetry information.
- Show comfort with Python or comparable general-purpose languages for building data processing logic and integration scripts.
- Exhibit fluency in distributed processing frameworks necessary to manage the scale and velocity of sensor data streams.
- Possess expertise with cloud storage services where grid telemetry, logs, and processed datasets are preserved long-term.
- Understand time series data patterns and best practices for structuring data to support trend analysis and anomaly detection.
- Have knowledge of grid monitoring concepts or active grid response methodologies, which is preferred for context alignment.
- Commit to following defined engineering processes that uphold safety, reliability, and data integrity for grid-related applications.
Nice to have
- Hands-on experience with visualization tools that display grid status in operations centers, enhancing situational awareness for responders.
Skills & Tools
- The role requires hands-on proficiency in Python, SQL, Cloud Storage, Apache Kafka, Time Series Databases, and Grafana.
- You should be comfortable managing data ingestion from protocols and formats common in industrial and utility environments.
- Experience with infrastructure-as-code practices for data platforms is valued to ensure reproducible deployments.
- Familiarity with monitoring and alerting frameworks that integrate with operational technology enhances reliability workflows.
- Understanding of data governance, access controls, and auditability supports compliance and operational trust.
Practical notes
- This is a full-time position based in San Francisco, California.
- Compensation details are provided as part of the total reward package.
- The role requires engagement in collaborative practices with cross-functional teams responsible for grid reliability.
- Applicants should be prepared to discuss how their background aligns with the technical and safety-critical nature of active grid response.