Job Description
About the Role We are setting up a dedicated Reliability Operations Team responsible for proactively overseeing critical business and technology processes. As a Reliability Operations Analyst, you will play a key role in ensuring operational continuity, minimizing risks, and driving timely issue detection and resolution. You will be part of the first line of defense, ensuring smooth operations and effective communication across stakeholders.
- Real-Time Monitoring Oversight Continuously monitor systems, applications, and business-critical processes using dashboards, alerts, and monitoring tools
- Detect anomalies early to prevent downstream impact
- Incident Detection Response Identify, log, categorize, and escalate incidents as per defined SLAs and escalation matrices
- Act as the central coordination point during live incidents
- First-Level Troubleshooting Perform initial triage to validate alerts and reduce false positives
- Gather relevant information before escalation to technical or business teams
- Operational Continuity (24/7 Coverage) Ensure seamless handovers between shifts with complete context
- Maintain uninterrupted operations across rotational shift.
- Governance Documentation Maintain shift logs, incident records, and escalation notes
- Update SOPs and contribute to internal knowledge bases
- Cross-Functional Collaboration Work closely with Tech Ops, Business Ops, Risk, Product, and Engineering teams for timely issue resolution
- Reporting Insights Share real-time updates during incidents
- Prepare daily/weekly operational reports
- Highlight recurring issues, trends, and reliability risks
You ll Excel in This Role If You Have
- Ability to remain calm and act decisively under pressure
- Clear written and verbal communication for structured incident reporting
- Strong sense of ownership and attention to detail
- Willingness to work in rotational 9-hour shifts for 24/7 coverage
- Basic analytical aptitude with Excel proficiency
- Technical Exposure (Preferred)
- Familiarity with monitoring and observability tools (Slack or similar)
- Basic understanding of system operations, networks, or process monitoring
Qualifications Bachelor s degree in Computer Science, IT, Operations, Finance, Trading, or a related field (or equivalent practical experience) 1 3 years of experience in Operations, Reliability, NOC, or SOC roles (Fresh graduates with strong aptitude may be considered for junior roles) Foundational understanding of business or technology operations
You ll Know You re Winning When You consistently remain calm and decisive during high-pressure incidents
- You consistently remain calm and decisive during high-pressure incidents
- Incident updates are clear, structured, and timely
- Monitoring alerts are validated efficiently, reducing false alarms
- Shift handovers are seamless with zero loss of context
- You effectively use monitoring tools (Slack or similar) to coordinate responses
- You adapt comfortably to rotational shifts while maintaining performance
- You demonstrate strong aptitude, learning agility, and Excel-based analysis
No Referrers Available
There are currently no referrers available for this job. You can still apply, will let you know once there is any referrer available.
