Principal Platform Engineer in Abbeyville, Colorado at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Principal Platform Engineer based in United States.
We are looking for a senior-level platform engineering leader responsible for building the foundation that enables engineering teams to deliver reliable, scalable, and secure products.
This role combines hands-on technical execution with architectural ownership across cloud infrastructure, developer platforms, automation, and reliability engineering.
You will design and operate critical systems that accelerate software delivery while improving developer experience and operational excellence.
Working closely with engineering, security, AI, and business stakeholders, you will shape the future of a modern technology platform in a regulated healthcare environment.
The position offers the opportunity to influence engineering practices, introduce innovative tooling, and build solutions adopted across the organization.
You will play a key role in creating a self-service, high-performing platform that empowers teams to move faster with confidence.
As a Principal Platform Engineer, you will own the architecture, development, and operation of the core engineering platform that supports product delivery. You will drive platform reliability, automation, developer productivity, and cloud excellence while partnering with cross-functional teams to establish scalable engineering practices.
- Design, build, and maintain core cloud infrastructure, including compute, networking, identity management, and container orchestration within AWS environments.
- Own and evolve CI/CD platforms, including build pipelines, artifact management, deployment workflows, environment promotion, and release automation.
- Develop and improve observability capabilities through logging, metrics, tracing, alerting, and reliability frameworks such as SLIs and SLOs.
- Lead infrastructure-as-code practices and manage Kubernetes environments, including clusters, operators, and production workloads.
- Build internal developer platforms, tooling, and self-service capabilities that allow engineering teams to deploy and operate services efficiently.
- Partner with security and compliance teams to implement appropriate access controls, secrets management, network policies, and governance practices.
- Establish reliability engineering standards, improve incident response processes, and support major production issue resolution.
- Collaborate with engineering leadership and finance teams on infrastructure cost management, capacity planning, and optimization initiatives.
- Create documentation, runbooks, and clear platform interfaces to improve adoption and operational consistency.
- Participate in technical reviews, code reviews, and engineering improvement initiatives while promoting automation, simplicity, and scalable design practices.
- Support teams working with AI-powered applications and emerging workloads by helping build reliable infrastructure foundations.
The ideal candidate is an experienced platform engineering professional with strong architectural expertise, hands-on development skills, and a track record of building scalable internal platforms.
- Bachelor’s degree or higher in Computer Science, Engineering, or a related technical discipline, or equivalent practical experience.
- 8+ years of experience in platform engineering, infrastructure engineering, site reliability engineering, or related fields with demonstrated Staff, Principal, or Architect-level ownership.
- Proven experience designing and operating enterprise-scale cloud infrastructure and production systems.
- Deep AWS expertise, including IAM, networking, compute services, and Kubernetes-based environments.
- Strong experience building CI/CD systems for high-performing engineering organizations, including deployment automation and release management.
- Hands-on programming experience in Python, Go, or TypeScript with the ability to contribute directly to production code.
- Experience building internal developer platforms, tools, or automation solutions adopted by engineering teams.
- Practical knowledge of infrastructure-as-code tools such as Terraform, CloudFormation, or Pulumi.
- Experience with modern observability solutions including Prometheus, Grafana, Datadog, New Relic, or OpenTelemetry.
- Strong understanding of reliability engineering, distributed systems, security practices, and operational excellence.
- Excellent communication skills with the ability to collaborate across engineering, security, finance, and business teams.
- Strong documentation habits and the ability to create clear technical guidance and operational standards.
Preferred qualifications include:
- Experience supporting AI or LLM-based workloads, including GPU infrastructure, model serving, or AI platform operations.
- Background working in healthcare, pharmacy, or other regulated data environments.
- Kubernetes operator or custom controller development experience.
- FinOps experience and ability to optimize infrastructure investments.
- Familiarity with distributed systems technologies such as Kafka, gRPC, Redis, or similar platforms.
- Experience with security scanning, quality automation, and modern Agile engineering practices.
- Competitive salary range of $190,000 - $225,000 annually.
- Fully remote work environment within the United States.
- Opportunity to work on impactful healthcare technology solutions.
- Ability to influence platform architecture and engineering practices at scale.
- Collaboration with experienced engineering, AI, security, and technology teams.
- Inclusive workplace culture focused on innovation, teamwork, and continuous improvement.
- Opportunities to contribute to modern cloud, automation, and developer productivity initiatives.