DevOps / Platform Infrastructure Engineer in United States Embassy at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a DevOps / Platform Infrastructure Engineer based in United States.
This role is an opportunity to shape and operate the cloud infrastructure powering critical healthcare applications and services.
You will work hands-on across Azure infrastructure, Kubernetes platforms, CI/CD automation, security, and reliability engineering.
The position focuses on building scalable systems, improving developer experiences, and ensuring production environments remain secure and resilient.
You will collaborate closely with engineering and security teams while owning key platform capabilities and operational improvements.
This role is ideal for an engineer who enjoys solving complex infrastructure challenges, automating workflows, and improving systems through thoughtful engineering practices.
You will play a key role in maintaining reliable technology platforms where security, compliance, and operational excellence are essential.
The DevOps / Platform Infrastructure Engineer will design, maintain, and improve cloud infrastructure while ensuring reliable, secure, and efficient operations. This role combines infrastructure engineering, automation, platform reliability, and technical collaboration to support high-performing application environments.
- Build, maintain, and improve Microsoft Azure infrastructure using Infrastructure-as-Code practices with Bicep.
- Operate and optimize Azure Kubernetes Service (AKS) environments, including upgrades, scaling, reliability improvements, and cost optimization.
- Own and enhance CI/CD pipelines using GitHub Actions to improve deployment speed, reliability, and developer productivity.
- Automate operational processes through scripting with PowerShell and Python.
- Implement monitoring, logging, alerting, and observability solutions to proactively identify and resolve issues.
- Reduce cloud costs while maintaining security, availability, and operational stability.
- Support database environments including PostgreSQL, SQL Server, and MySQL through maintenance, backups, performance optimization, and disaster recovery planning.
- Design and maintain network architectures including VNets, private endpoints, subnet segmentation, routing, peering, and hybrid connectivity.
- Support identity management, access controls, secrets management, and security compliance initiatives.
- Improve developer experience through better tooling, automation, documentation, and platform standards.
- Participate in incident response, troubleshooting, root-cause analysis, and operational improvement initiatives.
- Create and maintain technical documentation including runbooks, architecture diagrams, onboarding guides, incident reports, and operational procedures.
- Participate in production support and on-call rotations for critical incidents while focusing on long-term reliability improvements.
The ideal candidate has strong experience operating production infrastructure and building secure, scalable cloud platforms. You should be comfortable working independently, solving complex technical problems, and collaborating with engineering teams in regulated environments.
- 5+ years of experience managing and supporting production infrastructure.
- Deep hands-on experience with Microsoft Azure cloud services.
- Strong Infrastructure-as-Code experience, particularly with Bicep.
- Extensive networking knowledge, including cloud network design and troubleshooting.
- Production Kubernetes experience, preferably with Azure Kubernetes Service (AKS).
- Strong experience building and managing CI/CD pipelines with GitHub Actions.
- Advanced scripting skills with PowerShell and Python.
- Experience supporting both Linux and Windows Server environments.
- Experience managing PostgreSQL and SQL Server database environments.
- Experience with observability platforms such as Azure Monitor or similar solutions.
- Knowledge of RBAC, IAM, secrets management, and security control implementation.
- Experience working in regulated environments involving standards such as HIPAA, HITRUST, or SOC 2.
- Strong technical writing and documentation skills with a commitment to treating documentation as code.
- Experience participating in on-call support, incident response, and production troubleshooting.
- Strong problem-solving skills, curiosity, and the ability to communicate clearly during technical challenges.
- Competitive salary range of $128,600 to $181,600.
- Fully remote work environment with flexible scheduling.
- Comprehensive health, dental, and vision insurance options for employees and families.
- Wellness support including a yearly stipend for health-related activities.
- Paid parental leave program.
- Opportunity to work on technology platforms supporting meaningful healthcare outcomes.
- Collaborative engineering culture focused on learning, knowledge sharing, and continuous improvement.
- Access to challenging infrastructure projects involving cloud, automation, security, and reliability engineering.
- Supportive on-call approach focused on reducing recurring issues through automation and engineering improvements.
- Travel opportunities of up to 10% when required.