Staff / Principal Platform Engineer

As a Staff / Principal Platform Engineer, you'll take end-to-end ownership of building, securing, and scaling AI products. You'll be the driving force behind cloud infrastructure, partnering with engineers across the organization to deploy and evolve services across major cloud providers using Terraform, ArgoCD, and other tooling. In this high-impact role, you'll identify what needs to be done and move it forward, directly shaping how operations and innovation happen.

Responsibilities

  • Work closely with engineers to design, deploy, and maintain reliable, high-performance, and secure cloud infrastructure for TTS and LLM Router
  • Drive engineering velocity by identifying and building AI-powered tooling and workflows that improve how teams develop and deploy software
  • Facilitate a "you build it, you run it" culture by providing tools and processes for monitoring reliability, availability, and performance of services
  • Manage pipelines to ensure smooth and efficient code integration and deployment
  • Conduct root cause analysis to identify critical issues and develop automated solutions to prevent recurrence

Requirements

  • 8-10 years of experience in software engineering
  • 3+ years of experience with infrastructure-as-code
  • Proficiency in managing Kubernetes clusters and applications, including creating Kustomize manifests/Helm charts for new applications
  • Experience in creating and maintaining CI/CD pipelines for both applications and infrastructure deployments (using tools like Terraform/Terragrunt, ArgoCD, GitHub Actions, Ansible, etc.)
  • Deep knowledge of at least one major cloud provider (Google Cloud Platform, Microsoft Azure, Oracle Cloud)
  • Proficient in at least one backend programming/scripting language such as Golang, Python, and Bash
  • Candidates must be based in the SF Bay Area or willing to relocate

Benefits

  • Equity

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available