Staff Engineer, Platform & Infrastructure in New York at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Staff Engineer, Platform & Infrastructure based in United States.
This is a high-impact staff-level engineering role responsible for owning and scaling the infrastructure behind a developer-focused platform.
You will take end-to-end ownership of cloud infrastructure, including compute, networking, data systems, and infrastructure automation.
A major focus will be building and productizing Bring Your Own Cloud (BYOC) and private cloud deployments for enterprise customers.
You will establish strong foundations for reliability, security, compliance, observability, and efficient software delivery.
The role combines deep technical leadership with product thinking, strategic decision-making, and close collaboration across a small, autonomous engineering team.
You will help define infrastructure standards and shape the technical roadmap as the platform scales.
This is a remote-first opportunity open to experienced engineers across North America, LATAM, and Europe.
- Own and scale the cloud infrastructure supporting the platform, including compute, networking, storage, and core data services.
- Lead the development and productization of BYOC and private cloud deployments, including provisioning, upgrades, observability, and operational scalability.
- Build and evolve infrastructure-as-code and GitOps foundations that enable a small engineering team to deploy safely and efficiently.
- Establish reliability as a core product capability by defining meaningful service-level objectives, improving observability, and developing trusted incident response processes.
- Lead the infrastructure and operational strategy for Postgres, Redis, Elasticsearch, ClickHouse, and other critical data systems as usage grows.
- Own infrastructure-related security and compliance initiatives, including SOC 2, GDPR, HIPAA, penetration testing, and enterprise security reviews.
- Design scalable Terraform modules, multi-account cloud architectures, and safe infrastructure state-management practices.
- Operate and improve production Kubernetes environments while making pragmatic decisions about when alternative technologies are more appropriate.
- Support software deployments in customer-controlled environments, including single-tenant, on-premises, air-gapped, and other constrained environments.
- Collaborate directly with customers and engineering teams to understand infrastructure challenges and turn customer needs into practical technical solutions.
- Contribute to product specifications, technical roadmaps, and broader company strategy through strong technical judgment and ownership of key decisions.
- Provide technical leadership during incidents, unblock teammates, and establish engineering practices that improve reliability and delivery speed.
- 10+ years of professional experience in platform engineering, infrastructure, DevOps, SRE, or closely related disciplines, with demonstrated experience operating high-traffic production systems.
- Deep hands-on experience with Kubernetes in production, combined with strong judgment around its appropriate use and alternatives.
- Proven experience operating software in environments outside your direct control, such as BYOC, private cloud, single-tenant, on-premises, or air-gapped environments.
- Advanced AWS expertise; familiarity with GCP and Azure is a plus.
- Strong Terraform experience, including reusable modules, secure state management, and multi-account cloud architectures.
- Significant database expertise, particularly with PostgreSQL, including performance optimization, migrations under load, replication, and production operations.
- Practical experience with security and compliance programs such as SOC 2 or ISO audits, penetration tests, and enterprise security reviews.
- Product-oriented mindset with strong developer empathy and the ability to turn customer problems into technical solutions.
- Familiarity with TypeScript and Node.js is beneficial for working effectively within the existing technology environment.
- Excellent communication and collaboration skills, with the ability to work effectively in a remote-first, distributed environment.
- Proactive, pragmatic, and low-ego approach, with a willingness to take ownership during incidents and help unblock colleagues.
- Strong autonomy, ownership mentality, and comfort navigating ambiguity in a fast-moving startup environment.
- Ability to influence technical direction and communicate complex infrastructure decisions clearly to both technical and non-technical stakeholders.
- Availability to work remotely from North America, LATAM, or Europe.
- Fully remote, remote-first working environment.
- Opportunity to work from North America, LATAM, or Europe.
- High level of autonomy and ownership over critical platform and infrastructure decisions.
- Staff-level opportunity to shape technical strategy, infrastructure standards, and product direction.
- Work on challenging developer infrastructure and cloud technologies at a growing technology company.
- Exposure to Kubernetes, AWS, Terraform, GitOps, BYOC, private cloud, distributed systems, and large-scale data infrastructure.
- Opportunity to build scalable infrastructure products rather than maintaining bespoke solutions.
- Close collaboration with an experienced, highly autonomous engineering team.
- Significant influence on reliability, security, compliance, and the future scalability of the platform.
- Startup environment offering broad technical scope, rapid decision-making, and the opportunity to make a direct impact.