Senior Devops Engineer

STL Digital
Posted on
STL Digital logo

Experience
8 - 12 yrs
Salary (CTC)
₹15L - ₹25L
Job Location
Pune, India
Vacancy
1
Designation
Senior Devops Engineer
Job Type
Not specified

Job Description

Required Skills-

  • Bachelor's Degree with 8+ years of professional experience handling large scale production systems.
  • Hands on experience in Designing and Deploying EKS / GKE/ AKS Clusters
  • Understanding of Kubernetes security best practices, including RBAC, network policies, and PodSecurityPolicies.
  • Identifying and resolving issues related to Kubernetes, networking, storage, and application deployments.
  • Experience in migrating workloads to Kubernetes
  • Experience with AWS or GCP cloud providers with certification.
  • Hands on experience with Ruby, Terraform and configuration management tools like Chef, Ansible or equivalent.
  • Excellent knowledge of large scale web applications/distributed systems.
  • Experience in observability tools like NewRelic, Datadog etc
  • Expertise in problem solving and analyzing global scale distributed systems.
  • Excellent written and verbal communication skills.
  • Critical thinking, continuously challenging how and why we do things to help us improve

Responsibilities

  • Led the migration of legacy AWS workloads  to Amazon EKS (Elastic Kubernetes Service), leveraging Terraform for reproducible infrastructure and Helm for application packaging.
  • Provide weekend support for migration activities.
  • Developed Python and Bash-based automation to streamline containerization, secret management (AWS Secrets Manager), and resource tagging.
  • Implemented deep-trace monitoring using Observability tools to maintain visibility during and after the migration process.
  • Acted as the primary point of contact for resolving K8s-specific incidents, including Pod CrashLoopBackOffs, OOMKills, and VPC CNI networking issues.
  • Managed the full incident lifecyclefrom real-time weekend troubleshooting to Post-Mortem analysis—to build automated guardrails against recurrence.
  • Working closely with development, QA, and operations teams to ensure seamless collaboration and efficient workflows
  • Coordinate incident, problem and change management.
  • Participate in on-call rotation for after-hours emergencies

No Referrers Available

There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.