Infrastructure Software Engineer
Job description
Infrastructure Software Engineer at Coram Ai.
About the role
You will own the design and evolution of the edge and cloud infrastructure that powers our AI-first video security platform. You will write and maintain production-grade software that spans on-prem hardware, AWS, and Kubernetes environments. You will build the automation and observability that lets our distributed IoT devices operate reliably at scale. You will directly influence how AI models are deployed from the cloud to the edge in real-world security scenarios. You will collaborate closely with product and data teams to turn camera and sensor streams into actionable intelligence. You will help maintain rigorous security and compliance standards while maximizing the productivity of our developers. You will shape the future of physical security by making infrastructure that is both resilient and intuitive.
Key facts
What you'll do
Write and maintain production-grade software code for our custom edge infrastructure stack that underpins our IoT product portfolio.
Provision and maintain resources running on AWS using infrastructure as code approaches that keep our environment consistent and auditable.
Build provisioning and management workflows across hundreds of thousands of connected IoT devices deployed in the field, ensuring secure and reliable onboarding.
Build and refine CI and CD pipelines for multiple layers of our stack, from infrastructure code to application deployments and rollbacks.
Build observability and telemetry across our cloud applications and edge devices, enabling rapid detection and diagnosis of issues.
Help maintain compliance with security and operational standards such as SOC2 and HIPAA through infrastructure design and automation.
Maximize developer productivity by streamlining development workflows, tooling, and internal platforms used by engineers.
Implement robust over-the-air update mechanisms and device management capabilities for edge hardware running custom embedded stacks.
Design and operate hybrid and multi-cloud patterns that integrate AWS with on-prem systems and IoT gateways.
Construct high-volume data-processing pipelines that handle video and sensor data efficiently, reliably, and securely.
Operate and self-host critical observability tools such as Grafana and Prometheus to provide deep insights into system health.
Partner with hardware and firmware teams to ensure edge software integrates cleanly with physical devices and sensors.
Support on-site deployments and remote troubleshooting for customers who require tailored or complex installations.
Continuously iterate on infrastructure reliability, performance, and scalability as our global fleet of devices grows rapidly.
Requirements
3+ years of experience writing production infrastructure running on AWS using infrastructure as code tools, such as Pulumi or Terraform.
Experience with Docker and Kubernetes, particularly Amazon EKS, including cluster operations and networking.
3+ years of experience with Python, Go, or any other modern programming language for building scalable services.
A strong track record of building and maintaining CI/CD pipelines and automation across diverse parts of the stack.
Experience self-hosting and maintaining observability tools such as Grafana and Prometheus in production environments.
Hands-on background with edge and IoT infrastructure, including Yocto, device provisioning, and over-the-air updates.
Proven ability to remotely manage on-prem infrastructure, hybrid environments, and multi-cloud setups.
Experience helping maintain compliance frameworks such as SOC2 and HIPAA in infrastructure and deployment processes.
History of building high-volume data-processing pipelines that handle streaming and storage of large data sets.
Thrive in fast-paced, challenging environments where priorities shift and problems are solved with clarity.
Exceptional written and verbal communication skills in English, with the ability to influence stakeholders at all levels.
Demonstrated ability to work effectively in an onsite environment as part of a colocated team.
Practical notes
This role is based in London and requires the ability to work onsite. The position is full-time with an expectation of long-term growth and impact. There are no explicit visa sponsorship details, relocation support, or specific working hour ranges mentioned in the source description. Compensation details are not provided in the source material.