Job Description
About the Role
We're hiring a Senior DevOps Engineer to help build and run the AWS platform. You'll work
inside a multi-account AWS organization that hosts production workloads for enterprise
customers across manufacturing, distribution, and chemicals designing infrastructure that's
reliable by default, secure by design, and economical at scale.
What You'll Do
Build the platform
- Design and deploy highly available, scalable AWS infrastructure across multiple
accounts governed by AWS Control Tower
- Author and maintain Terraform modules and deployment pipelines that make customer
environments reproducible and audit-ready
- Build and operate EKS clusters as the runtime for Vendavo's core product platforms
Run it well
- Monitor service health and cluster performance using Grafana, Prometheus, and Graylog;
convert recurring signals into proactive controls
- Lead production incident response — triage, mitigate, write the RCA, and close the loop
with durable fixes
- Tune resource allocation, autoscaling, and right-sizing to keep performance high and unit
costs predictable
Make it safer and cheaper
- Partner with Security to enforce hardening baselines, IAM least-privilege, network
controls, and compliance evidence
- Drive FinOps practices: Reserved Instance and Savings Plan coverage, anomaly
detection, tag hygiene, and per-customer cost visibility
Be the expert
- Act as the cloud SME for escalations from Customer Success, Support, and Product
- Mentor peers, document the platform, and codify lessons learned so the team gets stronger with every incident
What You Bring
- 8+ years operating production AWS infrastructure at scale
- Hands-on production experience with EKS (or equivalent managed Kubernetes) —
networking, ingress, autoscaling, upgrades
- Deep familiarity with core AWS services: EC2, VPC, IAM, S3, RDS, Route 53, KMS
- Production experience with Amazon RDS — administration, backups, replication,
failover testing
- Strong Terraform skills, including module design, state management, and drift
remediation
- CI/CD experience with Jenkins, GitLab CI, or Azure DevOps Pipelines
- Working knowledge of Prometheus, Grafana, and a centralized logging stack (Graylog,
ELK, or similar)
- Solid scripting skills in Python (Bash and Go are a plus)
- Strong networking fundamentals applied to Kubernetes — DNS, ingress, load balancing,
service-to-service patterns
- A bias toward writing things down, and a calm head under incident pressure
Bonus Points For
- CKA or CKAD certification
- AWS certification (Solutions Architect, DevOps Engineer, or equivalent)
- HashiCorp Certified: Terraform Associate
- Experience with configuration management (Ansible, Puppet, or Chef)
- Exposure to Azure (some legacy workloads remain post-migration)
- FinOps Certified Practitioner or equivalent cost-engineering background
No Referrers Available
There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.