Job Description
Please share your profile to rehan@itcan.biz or connect on WhatsApp: +601115489275
Role:Production Support Engineer
Location: KualaLumpur-Malaysia (Onsite)
*Role Summary*
We are looking for production Support Engineer to support the reliability and day-to-day operation of large-scale, distributed production systems. This role focuses on troubleshooting, deployment support, and coordination across application, platform, and infrastructure layers in a fast-paced environment.
*Core Responsibilities*
- Support production stability by monitoring, troubleshooting, and resolving service issues across distributed systems.
- Perform initial triage and coordination for production problems, following defined procedures and escalation paths.
- Support Kubernetes-based workloads, including pod health, service availability, resource usage, and basic networking issues.
- Assist with deployment and change activities, validating configurations, monitoring rollouts, and ensuring minimal disruption.
- Support application build and delivery workflows, including basic code compilation, container image builds, and artifact validation.
- Assist in troubleshooting data platforms and pipelines, including batch and streaming workloads.
- Provide on call and troubleshooting support for big data platforms and services, including Flink, Hive, Yarn, Spark, and related data processing workflows.
- Maintain clear operational communication, documentation, and shift handoffs while collaborating with engineering and platform teams.
*Required Skills & Experience*
- Experience in a production operations or platform support role.
- Strong Linux knowledge, including log analysis, processes, networking, filesystem basics, and resource monitoring.
- Hands-on troubleshooting exposure to Kubernetes or containerized environments at a support level.
- Basic understanding of middleware and service communication layers, such as service mesh, traffic routing, or service discovery.
- Introductory to intermediate experience with application build processes, including compiling code, building container images, or working with CI/CD pipelines.
- Basic familiarity with big data technologies and concepts such as Spark, Hive, Flink, HDFS, and analytical platforms such as ClickHouse, along with Redis and relational databases.
- Strong communication skills, collaboration skills, and a clear sense of ownership when handling production issues
No Referrers Available
There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.
