Padmi
Cyngn logo
Cyngn

autonomous mobile robots (AMRs) · autonomous tuggers

Senior DevOps Lead - Cloud & Autonomous System

San Francisco Bay Area · OnsitePosted 17 months ago
Infrastructure And DatabasesUnspecified
Apply at Cyngn

Opens the source posting on jobs.lever.co

Source description

About the role

View original

About Cyngn

Based in Mountain View, CA, Cyngn is a publicly-traded autonomous technology company. We deploy self-driving industrial vehicles - specifically autonomous tuggers - to factories, warehouses, and other facilities throughout North America. To build this emergent technology, we are looking for innovative, motivated, and experienced leaders to join us and move this field forward. If you like to build, tinker, and create with a team of trusted and passionate colleagues, then Cyngn is the place for you. Key reasons to join Cyngn:

We are small and big. With under 100 employees, Cyngn operates with the energy of a startup. On the other hand, we’re publicly traded. This means our employees not only work in close-knit teams with mentorship from company leaders—they also get access to the liquidity of our publicly-traded equity.

We build today and deploy tomorrow. Our autonomous vehicles aren’t just test concepts—they’re deployed to real clients right now. That means your work will have a tangible, visible impact.

We aren’t robots. We just develop them. We’re a welcoming, diverse team of sharp thinkers and kind humans. Collaboration and trust drive our creative environment. At Cyngn, everyone’s perspective matters—and that’s what powers our innovation.

About this Role

As a Senior DevOps Leadat Cyngn, you will play a vital role in architecting and managing infrastructure across cloud and autonomous vehicle systems. This position combines traditional cloud DevOps leadership with specialized expertise in robotics and autonomous systems infrastructure. You will bridge the gap between cloud operations and edge computing while leading a team of DevOps engineers to build and maintain scalable, reliable infrastructure for our autonomous vehicle platform.

What you will do in this role Lead and architect cloud and vehicle infrastructure initiatives across AWS and ROS/Linux environments

Design and implement scalable solutions for both cloud services and autonomous vehicle systems

Establish and maintain DevOps best practices, CI/CD pipelines, and infrastructure as code

Design, develop, and operate production Kubernetes platforms (EKS and edge/lightweight distributions), including cluster lifecycle and upgrades, multi-tenancy, RBAC, networking (CNI, ingress, service mesh), autoscaling, and workload reliability

Develop reusable Kubernetes deployment tooling and packaging (Helm charts, Kustomize overlays, GitOps with Argo CD/Flux) and, where needed, custom controllers/operators and CRDs to automate platform operations

Develop and own the Terraform codebase as a product: reusable modules, remote state and workspace strategy, versioning and release process, provider upgrades, drift detection, policy-as-code guardrails, and automated plan/apply pipelines

Drive observability, monitoring, and incident response strategies

Optimize performance and cost efficiency of cloud and edge computing resources

Mentor team members and foster a developer-friendly environment

Manage on-call rotations and incident response processes

Architect solutions for processing and storing large-scale vehicle telemetry data

Lead security initiatives and compliance efforts across infrastructure

Design and implement solutions for both cloud services and autonomous vehicle systems

Optimize system performance for real-time processing of high-bandwidth sensor data

Develop and maintain documentation for system architecture and integration procedures

Who you are

10+ years of relevant DevOps/Infrastructure experience

Proven track record as a technical lead in platform or infrastructure teams

Advanced expertise in AWS services, infrastructure as code (Terraform), and Kubernetes

5+ years of hands-on Kubernetes development and operations at production scale, including cluster architecture and upgrades, RBAC and admission control, networking and ingress, resource requests/limits and autoscaling (HPA/VPA/Cluster Autoscaler/Karpenter), and troubleshooting workloads end to end

5+ years writing production Terraform, with proven authorship of reusable modules, remote state/workspace design, provider and version upgrade strategy, and CI-driven plan/apply workflows with code review and automated testing

Strong experience with service mesh (Istio) and Helm/Kustomize

Deep understanding of ROS/ROS2 and Linux kernel configurations

Experience with GPU configurations and ML infrastructure

Expertise in ARM and NVIDIA CUDA platform configurations

Strong programming skills in Python and shell scripting

Experience with infrastructure automation (Ansible)

Expertise in CI/CD tools (Jenkins, GitHub Actions)

Strong system architecture and design skills

Excellence in technical documentation

Outstanding problem-solving abilities

Strong leadership and mentoring capabilities

Nice to haves Experience with autonomous vehicle systems

Track record of optimizing GPU-based ML infrastructure

Experience with large-scale IoT deployments

Contributions to open-source projects

Experience with real-time systems and low-latency requirements

Expertise in security implementations including SSO, IdP, and AWS Cognito

Experience with JFrog artifactory and container registry management

Proficiency in AWS IoT Greengrass

Experience with container resource management on edge devices

Understanding of CPU affinity and priority scheduling

Track record of implementing cost optimization strategies

Experience with scaling systems both horizontally and vertically

Contributions to Kubernetes or Terraform ecosystem projects (upstream code, operators, providers, or published modules)

Experience running GitOps at scale (Argo CD/Flux) across multiple clusters and environments, including progressive delivery (Argo Rollouts, canary/blue-green)

Experience authoring custom Terraform providers or extending Terraform with external data sources and automation

Certifications such as CKA/CKAD/CKS or HashiCorp Certified: Terraform Associate

Benefits & Perks

  • Health benefits (Medical, Dental, Vision, HSA and FSA (Health & Dependent Daycare), Employee Assistance Program, 1:1 Health Concierge)

  • Life, Short-term, and long-term disability insurance (Cyngn funds 100% of premiums)

  • Company 401(k)

  • Commuter Benefits

  • Flexible vacation policy

  • Sabbatical leave opportunity after five years with the company

  • Paid Parental Leave

  • Daily lunches for in-office employees

  • Monthly meal and tech allowances for remote employees

One address, no account. We’ll tell you when matching roles go live.

More at Cyngn

Related open roles

View all roles