Site Reliability Engineer
Job Description
Key Skills
4 candidate(s) have already applied for this Job. Apply now
Job Overview
We are looking for an experienced Site Reliability Engineer (SRE) to drive enterprise monitoring, observability, application performance, and reliability initiatives across platforms and applications.
The ideal candidate will have strong hands-on experience with Azure Monitor, Application Insights, Azure Log Analytics, KQL, ServiceNow Event Management, Grafana, APM, and observability tools. This role will focus on improving reliability, reducing alert noise, accelerating incident response, strengthening SLO/SLI-driven operations, and enabling proactive monitoring and automation.
Key Responsibilities
Monitoring, Observability & Reliability
Design and define enterprise standards, patterns, and automation opportunities to improve monitoring and reliability across platforms and applications.
Develop and maintain monitoring and observability solutions using Azure Monitor, Application Insights, Log Analytics, KQL, Grafana, and APM/Synthetics tooling.
Implement proactive monitoring using Azure monitoring services, telemetry, and synthetic transactions.
Improve application performance monitoring and overall system reliability.
Contribute to observability and event management strategy, tooling intake, and governance activities.
Event Management & Incident Response
Engineer enterprise monitoring and event management patterns.
Develop and maintain reference architectures, runbooks, and event management models covering alert → event → incident workflows.
Design actionable alerts and ensure appropriate incident routing.
Work with ServiceNow Event Management to improve event correlation, monitoring, and operational response.
SLO / SLI & Reliability Engineering
Partner with product teams to implement SLO/SLI-driven operations.
Use service-level objectives, service-level indicators, postmortems, and operational data to identify reliability improvements.
Support self-healing initiatives and automation opportunities.
Incorporate postmortem learnings into monitoring patterns, rules, standards, and pipelines.
Collaboration & Enablement
Collaborate with IT Operations, Platform, Cybersecurity, Network, and Product teams.
Translate complex monitoring and reliability concepts into practical and consumable standards.
Continuous Improvement & AIOps
Identify opportunities to automate monitoring, event management, and operational processes.
Support proactive reliability engineering and self-healing capabilities.
Apply AIOps concepts and awareness to improve monitoring, event correlation, and operational efficiency.
Use analytics and systems thinking to improve reliability, MTTR, alert quality, and service performance.
Required Qualifications & Experience
Bachelor's degree in Information Technology, Computer Science, Engineering, or a related field.
Strong experience in Monitoring, Observability, or Site Reliability Engineering (SRE) roles.
Hands-on experience with Azure Monitor, Application Insights, Azure Log Analytics, and KQL.
Hands-on experience with ServiceNow Event Management / ITOM.
Strong understanding of telemetry and Application Performance Management (APM).
Experience with proactive monitoring and synthetic transactions.
Strong understanding of monitoring and application performance management.
Experience collaborating across IT Operations, Platform, Cybersecurity, Network, and Product teams.
Knowledge of SLO/SLI concepts and reliability engineering practices.
Mandatory / Core Skills
Site Reliability Engineering (SRE)
Monitoring & Observability
Azure Monitor
Azure Application Insights
Azure Log Analytics
KQL
Telemetry
Application Performance Monitoring (APM)
ServiceNow Event Management
ServiceNow ITOM
Preferred Skills
Grafana
Prometheus
AppDynamics
ThousandEyes
Nobl9
AIOps
Self-Healing Automation
Event Correlation
Soft Skills
Strong analytical and systems-thinking abilities.
Excellent written and verbal communication.
Strong collaboration and stakeholder management skills.
Ability to translate complex technical concepts into practical standards.
Coaching and knowledge-sharing capabilities.
Strong problem-solving and troubleshooting skills.
Work Details
Work Location: Eton Centris, Quezon Avenue, Quezon City, Manila
Work Setup: Hybrid – 2 Days Work From Home & 3 Days Return to Office
Work Schedule: Night Shift
Employment Type: Full-Time
Start Date: ASAP
Headcount: 1
Eligibility Requirements
Bachelor's degree in IT, Computer Science, Engineering, or a related field.
Candidates should have relevant experience in SRE, Monitoring, or Observability.
Must have hands-on experience with Azure Monitor / Application Insights / KQL and ServiceNow Event Management.
Must be amenable to a hybrid work setup in Quezon City.
Must be willing to work on a night shift schedule.
Candidates should demonstrate stable employment history and career progression.
Candidates who are current or former employees of Wipro are not eligible for this opportunity.
Benefits & Other Information
Hybrid work arrangement with 2 WFH and 3 RTO days.
Opportunity to work on enterprise-level monitoring, observability, and reliability initiatives.
Exposure to Azure cloud monitoring and ServiceNow ITOM/Event Management.
Recruitment Process
Paper Screening – Profile is endorsed to the Operations team to determine whether the candidate meets the basic requirements and qualifications for the position.
L1 Interview – Interview with the Practice / Operations Team.
L2 Interview – Optional, depending on the recruitment process.
Technical Assessment / Final Interview – Conducted by the customer's Operations team.
Pre-Screening Questions
What is your highest educational attainment?
How many years of relevant experience do you have in Monitoring, Observability, or SRE roles, with hands-on experience in Azure Monitor / Application Insights (KQL) and ServiceNow Event Management?
How many years of relevant experience do you have in SRE roles with a strong understanding of monitoring and Application Performance Management (APM)?
What was your last drawn salary?
What is your salary expectation?
Are you amenable to working on a hybrid setup in Quezon City with a night shift schedule?
When are you available to start once hired?
Are you currently or formerly employed by Wipro?
Role
Computer Operators
Timings
Night Shift (Contract To Hire)
Industry
IT-Software / Software Services
Work Mode
Hybrid
Process
Non-Voice
Functional Area
IT Software/Hardware
Note: Myglit doesn't charge any money from candidates. If you have been asked to pay money to get this job then report to us immediately at support@myglit.com.
Interview Tips
- Giving the VNA round?
- What are the most important skills you acquired as a Soft Skills/VNA trainer?
- How would you handle an irate customer?
Similar Jobs
CRM Lead – Microsoft Dynamics 365
Gratitude Inc5 - 8 Year(s)
₱ 35 - ₱ 45K p.m
Quezon Calabarzon, Philippines
Site Reliability Engineer
Gratitude Inc5 - 8 Year(s)
₱ 70 - ₱ 90K p.m
Quezon Calabarzon, Philippines
Service Desk Transformation Consultant
Gratitude Inc6 - 8 Year(s)
₱ 110 - ₱ 140K p.m
Quezon Calabarzon, Philippines
QA Lead with TOSCA
Gratitude Inc5 - 8 Year(s)
₱ 50 - ₱ 95K p.m
Quezon Calabarzon, Philippines
Monitoring, Observability and Event Management Architect
Gratitude Inc3 - 8 Year(s)
Confidential
Quezon Calabarzon, Philippines
Monitoring, Observability and Event Management Architect
Gratitude Inc3 - 15 Year(s)
₱ 100 - ₱ 110K p.m
Quezon Calabarzon, Philippines
Firewall Specialist
Gratitude Inc8 - 14 Year(s)
₱ 50 - ₱ 110K p.m
Quezon Calabarzon, Philippines
CRM lead
Gratitude Inc5 - 6 Year(s)
Confidential
Quezon Calabarzon, Philippines
Operational Technology (OT) Specialist / Team Lead
Gratitude Inc3 - 5 Year(s)
₱ 65 - ₱ 200K p.m
Quezon Calabarzon, Philippines

