TymblHub

© 2026 TymblHub

Production Support Systems Engineer (Night Shift)

Vcheck
Posted on

Experience
2 - 4 yrs
Job Location
Pune, India
Vacancy
1
Designation
System Support Engineer
Job Type
Not specified

Job Description


About the role
The Production Support Systems Engineer (Night Shift) maintains the stability and availability of production systems during overnight hours (10:00 PM 6:00 AM). Youll serve as the primary technical point of contact for incidents, monitoring alerts, and escalations ensuring issues are triaged, resolved, or escalated with minimal business impact. Working across our AWS-hosted Python/Django/FastAPI backend and React JS frontend, youll manage incidents end-to-end via JIRA Service Management and monitor system health through Sentry and CloudWatch. This role includes night differential pay and requires on-call availability for critical escalations outside shift hours.

What youll be doing
  • Monitor Sentry and CloudWatch for errors, exceptions, and anomalies throughout the night shift
  • Triage, prioritize, and respond to alerts per SLA; execute runbooks for known issues
  • Create, update, and resolve JIRA Service Management tickets with full incident documentation
  • Coordinate escalations to Tier 3 and on-call engineers with detailed summaries
  • Perform shift handoffs via written reports and overlap with the day team
  • Diagnose failures across Django/FastAPI services, PostgreSQL, and the React JS frontend
  • Inspect CloudWatch Logs for ECS task errors, AWS Batch job failures, and Lambda/API issues
  • Restart failed ECS tasks and services; triage Batch job queue stalls and retry failures
  • Identify and escalate PostgreSQL query performance issues and connection pool exhaustion
  • Distinguish frontend (React JS) issues from backend API failures using Sentry error traces
  • Manage and respond to CloudWatch alarms covering CPU, memory, error rates, and 5xx responses
  • Perform environment health checks at shift start and end; validate batch job completion
  • Monitor RDS metrics and identify slow query patterns; escalate with supporting evidence
  • Execute scheduled overnight maintenance including deployments, database migrations, and config changes
  • Validate nightly backup and data replication processes; escalate failures promptly
  • Document incident root cause, resolution steps, and preventive notes in JIRA
  • Contribute to and maintain runbooks and SOPs in the team knowledge base
  • Support planned change windows by monitoring post-deployment behavior and coordinating rollbacks via Git

About you
Key requirements:
Were looking for someone who is passionate about joining a diverse team and is driven to achieve results through ownership, process optimization, and upstanding character. If this describes you, we encourage you to apply, even if you dont meet every requirement listed.
  • 2 4 years of experience in Production Support, NOC, DevOps, or SRE roles
  • Hands-on experience with AWS services including ECS (Fargate/EC2), AWS Batch, CloudWatch, RDS, S3, IAM, VPC, Load Balancers, and API Gateway
  • Proficiency in Python for scripting, log parsing, and automation
  • Working knowledge of Django and FastAPI, including reading error logs and identifying ORM/API issues
  • Experience with PostgreSQL diagnostic queries, slow log analysis, and connection management
  • Familiarity with Sentry for exception triage, breadcrumb analysis, and release tracking
  • Comfortable working with Docker reading container logs, inspecting images, and restarting containers
  • Experience with JIRA Service Management for incident ticketing, SLA management, and escalation workflows
  • Proficiency with Git for reviewing commits related to incidents and coordinating rollbacks
  • Ability to differentiate frontend (React JS) issues from backend API failures using Sentry and browser-level error traces
  • Calm and methodical under pressure during production incidents
  • Strong written communication skills for shift handoffs and incident reports
  • Self-directed with sound judgment on when to escalate versus resolve independently
  • Comfortable working nights including weekends and holidays on rotation
  • AWS certification (SysOps Administrator, Developer, or Solutions Architect Associate) is a plus
  • Experience with JIRA automation rules and SLA escalation configuration is a plus
  • Familiarity with CI/CD pipelines (GitHub Actions, CodePipeline) and deployment artifacts is a plus
  • Prior night shift or 24/7 NOC/on-call production support experience is a plus
  • ITIL Foundation or equivalent operational framework knowledge is a plus
  • Exposure to microservices architecture and RESTful API design patterns is a plus

Why us
You will be joining a cutting-edge company, where you will tackle complex challenges and work with the very best in the industry. In addition, we offer:
  • Competitive compensation package
  • Comprehensive benefits, including GHI coverage for you & your loved ones
  • Flexible vacation policy, encouraging you to take the time you need
  • Comfortable shift(s) to maintain work life balance
  • Annual wellness allowance to support your health and well-being
  • Quarterly team events, fun team activities monthly happy hours to refresh mind and soul.
  • A fun and collaborative work environment where youll be supported by a team of dedicated and collaborative colleagues
  • Additional equipment support, if needed, for your workplace
  • A vital role in shaping our companys future  

No Referrers Available

There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.