Staff DevOps Engineer

MEDvidi is an AI-powered mental healthcare platform setting a new standard for safe, effective, and scalable psychiatric care in the United States.

We combine licensed providers with proprietary AI tools to deliver consistent, outcomes-driven treatment for conditions like ADHD, anxiety, depression, and more. MEDvidi's technology automates charting, follow-ups, and treatment planning, freeing providers to focus on patient care while improving efficiency and clinical quality.

As our Staff DevOps Engineer, you'll decide where our infrastructure goes next and take the rest of engineering there. This is a hands-on role in the truest sense: you set the technical direction, ship it with your own hands, and own the outcomes for reliability, cost, security, and developer experience across the company. You'll partner directly with our product engineering teams, embedded in what they're building, removing infrastructure friction, and shaping infrastructure around real product needs.

Why this role

  • Real ownership, end to end. You own outcomes, not tasks, including your metrics. Reliability, cost, performance, deployment health, and developer velocity belong to you: you set the targets and bring the numbers to the table proactively.
  • Technical leadership without the management overhead. Lead by expertise and initiative where the most senior engineers are also the ones in the code. You set direction and raise the bar while staying hands-on.
  • AI-native by default. Agentic AI is a first-class part of how we work. You'll delegate to autonomous agents, integrate their output into production changes, and help shape how the whole org builds with AI.

Responsibilities

  • Set and drive the technical vision and quarterly roadmap for our infrastructure proactively, with clear trade-offs and measurable goals.
  • Run and evolve our AWS + Kubernetes (EKS) infrastructure: cluster management, autoscaling (Karpenter), policy enforcement (Kyverno), and zero-downtime operations.
  • Own Infrastructure as Code end to end (Terraform, AWS CDK in TypeScript) and our GitLab CI/CD (reusable/shared templates, OIDC, self-managed GitLab).
  • Build and own observability that teams actually use (Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch; log pipelines, APM) a shared view of system health, not a dashboard nobody opens.
  • Keep blue-green deployments and health-gated automated rollback fast and boring; own zero-downtime PostgreSQL schema migrations (expand/contract) and CI migration gating.
  • Own security engineering in a HIPAA environment: secrets hygiene (rotation, short-lived credentials, leak scanning), PHI-aware handling of logs and data, and Vault managed as code.
  • Partner directly with product teams to remove infrastructure friction and improve developer experience by design.
  • Use agentic AI as a core part of your workflow, integrating autonomous-agent output into production.

Requirements

  • 6+ years in DevOps/infrastructure engineering, with strong systems fundamentals and solid Linux administration and troubleshooting (performance analysis, resource management, process debugging).
  • Hands-on AWS (EC2, EKS, RDS, ElastiCache, Lambda, SQS, EventBridge, API Gateway, ALB, S3) and production Kubernetes/EKS (cluster management, node scaling, policy enforcement; Karpenter, Kyverno, or similar).
  • Strong Infrastructure as Code (Terraform and AWS CDK in TypeScript) and CI/CD ownership (GitLab CI/CD: reusable/shared templates, OIDC id_tokens, self-managed GitLab).
  • Monitoring and observability in practice (Prometheus, Grafana, OpenTelemetry, OpenSearch, CloudWatch; log-shipping and error tracking/APM).
  • Practical security engineering (secrets rotation, short-lived credentials, leak scanning, PHI-aware logging) and HashiCorp Vault as code (KV, JWT/OIDC auth for CI, po

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available