Principal DevOps Architect in United States Embassy at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Principal DevOps Architect based in United States.
This role offers the opportunity to define and shape the future of cloud infrastructure for a global technology platform supporting critical healthcare and research solutions. As a senior technical leader, you will architect scalable, secure, and highly reliable cloud environments while driving modern DevOps practices across the organization. You will combine deep hands-on engineering expertise with strategic architectural vision, influencing teams through technical excellence and innovation. The position focuses on infrastructure automation, AI platform operations, reliability engineering, and compliance-driven cloud solutions. You will play a key role in advancing infrastructure as code, improving developer experiences, and ensuring systems remain secure, efficient, and audit-ready. This is an impactful opportunity for an experienced architect who thrives in complex environments and enjoys solving challenging technical problems.
The Principal DevOps Architect will serve as the senior technical authority responsible for designing, implementing, and evolving cloud platforms, DevOps practices, and AI infrastructure. This hands-on individual contributor role will influence engineering standards, improve operational excellence, and enable teams to deliver reliable and secure solutions.
- Own the cloud platform architecture, infrastructure roadmap, deployment strategies, observability practices, and AI platform direction while partnering with engineering leadership.
- Establish engineering standards, reference architectures, and best practices for infrastructure as code, CI/CD pipelines, cloud operations, and AI tooling adoption.
- Design, maintain, and optimize cloud infrastructure using Terraform as the source of truth, including reusable modules, automated deployments, and policy-driven controls.
- Build and operate scalable AWS environments across development, testing, staging, and production while ensuring performance, availability, security, and compliance.
- Develop and improve CI/CD pipelines, release processes, containerized workloads, and deployment automation using technologies such as Docker, Kubernetes, GitHub Actions, and related tools.
- Lead reliability initiatives by defining SLIs, SLOs, error budgets, monitoring strategies, and incident response processes.
- Operate and govern AI/ML platforms, including model-serving infrastructure, AI observability, security controls, cost management, and responsible AI practices.
- Implement security and compliance measures supporting healthcare data protection, audit readiness, secrets management, vulnerability management, and supply-chain security.
- Collaborate with product, engineering, QA, support, and operations teams to improve automation, documentation, and cross-functional delivery.
- Mentor engineers through technical leadership, architecture reviews, and hands-on contributions without direct management responsibility.
The ideal candidate is a highly experienced cloud and DevOps professional with a strong background in architecture, automation, security, and large-scale platform engineering. They should be comfortable leading technical initiatives, influencing teams, and delivering solutions in regulated environments.
- Bachelor’s degree in software engineering, computer science, or equivalent technical experience.
- 10+ years of experience in DevOps, SRE, platform engineering, cloud architecture, or related disciplines, including senior individual contributor or architect-level responsibilities.
- Proven experience designing and operating large-scale cloud environments, preferably on AWS.
- Strong hands-on expertise with Terraform, infrastructure as code practices, reusable modules, automated provisioning, and CI/CD-based infrastructure management.
- Experience building and managing CI/CD pipelines using tools such as GitHub Actions, Jenkins, GitLab, or AWS-native solutions.
- Strong knowledge of Docker, Kubernetes, container orchestration, Linux administration, scripting, and automation.
- Experience with cloud observability, monitoring, troubleshooting, OpenTelemetry, and reliability engineering practices.
- Familiarity with security and compliance frameworks such as HIPAA, HITECH, HITRUST, SOC 2, PCI DSS, CIS Controls, or FedRAMP.
- Experience managing regulated data environments involving PHI, PII, or other sensitive information.
- Strong communication skills with the ability to explain complex technical concepts to both technical and business stakeholders.
- Demonstrated ability to introduce and scale new engineering practices, platforms, or operational improvements across teams.
- Experience with PHP, MySQL, SQL, or similar technologies is a plus.
Preferred qualifications include:
- Experience operating AI/ML or generative AI platforms in production environments.
- Knowledge of LLMOps practices, including prompt management, evaluation frameworks, AI monitoring, cost attribution, and AI incident response.
- Experience supporting LLM applications, AI agents, or MCP-based services.
- Familiarity with AI security frameworks, OWASP LLM Top 10, NIST AI Risk Management Framework, and key management practices.
- Experience with policy-as-code, secrets management, SBOM generation, image scanning, and software supply-chain security.
- Experience with FinOps practices and optimizing large-scale cloud spending.
- Background in healthcare technology, clinical research, or other highly regulated industries.
- Competitive annual salary range of approximately $158,000–$190,000, depending on experience and qualifications.
- Remote work flexibility within the United States.
- Health insurance coverage.
- Long-term disability and life insurance benefits.
- Unlimited paid time off.
- Paid holidays.
- Paid parental leave.
- 401(k) retirement plan with employer matching contributions.
- Monthly connectivity stipend reimbursement.
- Employee recognition programs and anniversary rewards.
- Opportunity to work on innovative cloud, AI, and healthcare technology solutions.
- Collaborative environment with opportunities for professional growth and technical leadership.