Senior AI Infrastructure Engineer in United States Embassy at Jobgether
Explore Related Opportunities
Job Description
This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior AI Infrastructure Engineer based in United States.
This role offers the opportunity to build and scale the next generation of AI-powered infrastructure supporting advanced customer experiences.
You will work at the intersection of artificial intelligence, backend engineering, API architecture, and production reliability.
As a key technical contributor, you will design systems that enable intelligent agents, automate workflows, and improve how AI capabilities are delivered at scale.
The position requires strong ownership, deep engineering expertise, and the ability to solve complex infrastructure challenges in a fast-moving environment.
You will collaborate with engineering, product, and operations teams to create reliable, scalable, and innovative AI solutions.
This is an impactful opportunity for an experienced engineer who wants to shape emerging AI platforms and influence the future of intelligent systems.
As a Senior AI Infrastructure Engineer, you will design, develop, and operate production-critical AI infrastructure that enables reliable agent-based experiences. You will own complex technical initiatives, contribute to platform architecture, and help establish engineering best practices for scalable AI systems.
- Design and implement multi-tenant AI infrastructure, including tenant isolation, rate limiting, and caller attribution systems.
- Build and maintain durable agent runtimes supporting long-running workflows, task lifecycle management, and failure recovery.
- Develop and enhance agent orchestration systems, including context management, memory, tool usage, and multi-agent coordination.
- Create evaluation-driven development frameworks with automated testing, regression detection, quality gates, and continuous improvement processes.
- Extend AI communication interfaces and enable agent-to-agent interactions across internal and customer-facing systems.
- Develop progressive discovery strategies for tools, agents, skills, and contextual information using semantic filtering and intelligent retrieval approaches.
- Build and maintain API-first infrastructure that supports both traditional applications and AI-powered consumers.
- Contribute to LLM abstraction layers that support flexibility across multiple model providers.
- Partner with engineering, product, and operations teams to align technical solutions with business needs.
- Participate in production support, incident response, postmortems, and reliability improvements.
- Create technical documentation, architecture diagrams, design specifications, and implementation plans.
- Support internal AI enablement by sharing expertise in LLM tooling, agentic workflows, and emerging AI development practices.
The ideal candidate is an experienced software engineer with strong expertise in AI infrastructure, backend systems, and production-scale platform development. You should be comfortable designing complex systems, collaborating across teams, and driving projects from concept through deployment.
- 6+ years of software engineering experience, including 1+ years focused on AI/ML infrastructure, LLM platforms, or agentic systems.
- Strong proficiency in Go (Golang) for backend engineering, with experience using Python and SQL.
- Experience designing, building, and operating production APIs and developer-facing platforms at scale.
- Strong understanding of multi-tenant architectures, including rate limiting, tenant isolation, and secure shared infrastructure.
- Hands-on experience with LLM APIs such as OpenAI or Anthropic, including prompt engineering and function/tool calling.
- Practical experience building agent workflows involving context management, state handling, caching strategies, and multi-step execution.
- Experience with MCP or similar tool-layer frameworks for AI agents and platform integrations.
- Familiarity with agent-to-agent communication protocols and durable workflow systems.
- Experience creating automated evaluation frameworks for AI systems, including testing suites, regression detection, and CI/CD quality controls.
- Ability to collaborate effectively with engineering, product, and operational stakeholders.
- Strong technical communication skills with experience creating design documents, system diagrams, and postmortems.
Preferred qualifications:
- Experience in advertising technology, media measurement, streaming, OTT, CTV, or digital video environments.
- Experience working with open-weight AI models such as Llama or Mistral.
- Knowledge of model fine-tuning approaches such as LoRA or PEFT.
- Familiarity with Rust or additional backend programming languages.
- Competitive base salary ranging from $155,000 to $180,000, depending on experience.
- Equity participation opportunities.
- Fully remote work opportunity within the United States.
- Flexible and discretionary paid time off.
- Company-wide spring, summer, and winter breaks.
- Comprehensive medical, dental, and vision insurance.
- 401(k) retirement plan with employer matching.
- Health Savings Account (HSA) and Flexible Spending Account (FSA) options.
- Paid maternity and parental leave for all family additions.
- Cell phone and internet reimbursement.
- Commuter benefits.
- Opportunity to work on cutting-edge AI technologies with a highly collaborative engineering team.