Cloud Platform Engineer
Job description
Cloud Platform Engineer at Airspace Intelligence.com.
About the role
This position leads technology delivery for aviation, defense, energy, and critical infrastructure missions. The role is responsible for deploying and operating secure cloud infrastructure that directly supports government customer needs. You will work in close collaboration with product engineers to ship new features and integrate diverse data sources into coherent solutions. The focus is on building systems that enhance operational decision superiority in demanding environments. Platform work centers on automation, scaling, and removing friction so product teams can move quickly. Systems thinking is the core skill required to succeed in this position. Teams in this role operate on-call rotations and own incident response as a standard part of the job.
Key facts
What you'll do
Deploy, scale, and maintain services on secure cloud infrastructure to support operational decision superiority.
Write production Python code to add features and integrate new data sources that amplify warfighter value.
Build and manage large-scale, complex systems that power mission-critical workflows across multiple domains.
Collaborate across multiple teams to align infrastructure work with product goals and delivery timelines.
Travel as required to support customers, conduct on-site assessments, and engage in government engagements.
Adapt quickly to changing business priorities to keep delivery on track while maintaining system integrity.
Manage infrastructure, CI/CD pipelines, observability, and reliability using infrastructure-as-code practices.
Automate repetitive tasks and improve tooling to reduce manual effort and increase platform efficiency.
Support the full lifecycle of cloud services from initial deployment through ongoing operations and optimization.
Contribute to design reviews and technical documentation that ensure solutions meet security and compliance standards.
Implement monitoring and alerting strategies using tools like Grafana to maintain high system reliability.
Leverage modern LLM tools to enhance development workflows, code quality, and internal productivity.
Ensure all deployed solutions follow best practices for security, scalability, and maintainability.
Assist in debugging complex issues in distributed systems and help drive incidents toward resolution.
Requirements
Ability to maintain and grow infrastructure stack using AWS, Kubernetes, Docker, Terraform, Helm, PostgreSQL, Grafana.
Deep understanding of CI/CD pipelines to automate builds, tests, and deployments in a reliable manner.
Experience managing large-scale and complex distributed systems across multiple environments.
Experience writing production-grade Python code for services, integrations, and automation tasks.
Advanced knowledge of Kubernetes, Docker, and an object-oriented programming language for building robust services.
Proficiency with modern LLM tools to enhance development workflows and improve code quality.
Aptitude to lead initiatives and work independently with minimal supervision while delivering results.
Collaboration across multiple teams to integrate infrastructure and product workflows effectively.
Willingness to travel domestically as required by customer needs and project demands.
Flexibility to adjust to changing business priorities and shifting requirements without losing focus.
Ability to timely obtain all required U.S. authorizations for government-contracted work, including security clearances.
Commitment to following secure software development practices and infrastructure operations standards.
Capacity to communicate clearly with both technical and non-technical stakeholders during planning and execution.
Nice to have
Experience with export-controlled technologies and restricted U.S. employment scenarios.
Background in defense, aviation, or critical infrastructure sectors and related compliance frameworks.
Familiarity with structured post-incident analysis and incident response best practices.
Understanding of SRE principles and reliability engineering in cloud environments.
Knowledge of security and compliance requirements for government cloud workloads.
Experience contributing to open source projects or public technical documentation.
Practical notes
ASI works with export-controlled technology and restricted U.S. Employment offers depend on obtaining required U.S. immigration status and location restrictions. Typical interview steps
Platform interviews usually include an infrastructure scenario, a scripting or coding exercise, and operational questions. Candidates may be asked to design a deployment pipeline or debug an outage. Incident experience and an automation mindset are tested. Interviewers often ask about a past outage and how you handled it. Structured post-incident thinking, not heroics, is what they look for.