IT Operations Engineer
Job description
About the role
You will design, build, and maintain integrations and automations that connect our IT and business systems, eliminating manual, repetitive work across the organization. You will administer, configure, and optimize core IT platforms and business SaaS applications, ensuring consistent performance and reliability for a global workforce. You will script and deploy solutions using Python, Bash, or similar languages to automate repeatable operational processes at scale. You will integrate systems through APIs and webhooks, carefully handling authentication, rate limits, retries, and error cases to ensure resilience. You will handle advanced technical troubleshooting across IT systems, endpoints, identity, and integrations, and support escalations from the IT Support team. You will implement proactive measures for high availability, monitor system performance, and remediate issues before they impact users. You will contribute to disaster recovery planning and execution, ensuring swift restoration of critical services when needed. You will proactively identify improvements to systems and processes, then scope and deliver them without waiting to be told.
Key facts
What you'll do
- Design, build, and maintain integrations and automations that connect our IT and business systems, eliminating manual, repetitive work.
- Administer, configure, and optimize core IT platforms and business SaaS applications, ensuring consistent performance and reliability.
- Script and deploy solutions organization-wide using Python, Bash, or similar languages to automate repeatable operational processes.
- Integrate systems through APIs and webhooks, handling authentication, rate limits, retries, and error cases with attention to reliability.
- Perform advanced technical troubleshooting across IT systems, endpoints, identity, and integrations, supporting escalations from the IT Support team.
- Implement proactive monitoring and high availability measures, watching system performance and remediating issues before users are affected.
- Contribute to disaster recovery planning and execution, ensuring swift restoration of critical services during incidents.
- Identify and propose improvements to systems and processes, then scope and deliver them without waiting for direction.
- Create and maintain clear documentation for systems, processes, and integrations to support knowledge sharing and continuity.
- Partner closely with Security, Compliance, and Data teams to ensure all solutions remain governed and aligned with organizational standards.
- Communicate technical updates and progress clearly to both engineers and non-technical stakeholders, facilitating collaboration across teams.
- Own end-to-end responsibilities for projects, seeing tasks through from investigation and design through deployment and follow-up.
Requirements
- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related field, or equivalent practical experience.
- 3+ years of relevant experience in IT operations, systems administration, or a similar technical role.
- Strong proficiency with Python and Bash scripting to automate tasks and integrate systems.
- Solid understanding of IT service management practices and tools, including monitoring, logging, and incident response.
- Experience with SaaS platform administration and API integrations in a business-critical environment.
- Demonstrated ability to investigate unfamiliar systems, reverse-engineer workflows, and implement practical solutions quickly.
- Excellent problem-solving skills and comfort working with complex, distributed systems where clear thinking is essential.
- Strong written and verbal communication skills to collaborate effectively with both technical and non-technical stakeholders.
- Ability to work autonomously, manage priorities, and maintain ownership of tasks in a dynamic, fast-paced environment.
- Commitment to following processes and documentation standards to ensure consistency and auditability.
- Willingness to partner closely with Security, Compliance, and Data teams to uphold governance and regulatory standards.
Practical notes
ID.me is a full-time, in-office culture. Unless a specific job description explicitly states otherwise, all roles are on-site five days per week at one of our offices in McLean, VA; Mountain View, CA; New York City, NY; or Tampa, FL. Certain roles - such as field-based sales or other remote-by-design positions - may have different work arrangements as noted in their individual postings. At ID.me, we embrace the thoughtful use of AI tools in our daily work and there are even occasions where we leverage AI in our hiring process. However, during the interview process, we want to understand your individual skills and experiences. Therefore, we have guidelines on how AI can be appropriately used during your application and interviews which can be found here.