Site Reliability Engineer (SRE)
Job Description
Key Skills
9 candidate(s) have already applied for this Job. Apply now
JOB TITLE: Site Reliability Engineer (SRE)
JOB SUMMARY
We are looking for a Site Reliability Engineer (SRE) with strong experience in monitoring, observability, application performance management, and Azure monitoring technologies. The role will focus on improving reliability, reducing alert noise, accelerating incident response, and establishing enterprise monitoring and event-management standards.
The ideal candidate should have hands-on experience with Azure Monitor, Application Insights, Log Analytics, KQL, ServiceNow Event Management, Grafana, and APM/Synthetic monitoring tools.
IMPORTANT ELIGIBILITY REQUIREMENTS
Bachelor's degree in IT, Computer Science, Engineering, or a related field.
3+ years of relevant experience in Monitoring, Observability, or SRE roles.
Hands-on experience with Azure Monitor, Application Insights, KQL, and ServiceNow Event Management.
Strong knowledge of Azure Log Analytics, telemetry, APM, and proactive monitoring.
5+ years of SRE experience is preferred for candidates with deep monitoring and application performance expertise.
Knowledge of SLO/SLI platforms such as Nobl9 is an advantage.
Hands-on knowledge of ITSM processes and ServiceNow.
Understanding of network architecture and security, including WAN/LAN, TCP/IP, and PKI.
Awareness of AIOps concepts and vision.
Strong communication and collaboration skills.
Must be willing to work hybrid in Quezon City on a night shift.
Must not be a current or former Wipro employee.
Stable employment history; candidates should not be frequent job hoppers.
KEY RESPONSIBILITIES
Design and define monitoring, observability, reliability standards, patterns, and automation opportunities across platforms and applications.
Work extensively with Azure Monitor, ServiceNow ITOM Event Management, Grafana, APM, and synthetic monitoring tools.
Partner with product teams to implement SLI/SLO-driven operations.
Reduce alert noise and improve incident response and overall system reliability.
Develop and maintain reference architectures, runbooks, monitoring standards, and event-management models.
Establish actionable alert and incident routing using the alert → event → incident model.
Contribute to Monitoring, Observability, and Event Management strategy and governance.
Coach product and IT operations teams on monitoring and reliability best practices.
Use postmortem learnings to improve monitoring rules, patterns, and automation pipelines.
Support initiatives focused on reducing MTTR and improving change success and operational KPIs.
Identify opportunities for self-healing and proactive monitoring.
MUST-HAVE SKILLS
Azure Monitor
Application Insights
Azure Log Analytics
KQL
ServiceNow Event Management / ITOM
Cloud Observability
Monitoring & Application Performance Management
Grafana
Prometheus
AppDynamics
ThousandEyes
Telemetry and APM
Synthetic Monitoring
SLI/SLO concepts
ITSM processes
Monitoring automation
AIOps awareness
KEY COMPETENCIES
Communication & Teamwork: Ability to explain complex reliability concepts through standards, documentation, office hours, and community-of-practice sessions.
Technical Depth: Strong hands-on knowledge of monitoring, observability, ServiceNow Event Management, Azure Monitor/KQL, and automation.
Analytical & Systems Thinking: Ability to use SLI/SLOs, postmortems, and CMDB context to reduce alert noise, improve self-healing, and enhance MTTR and operational KPIs.
RECRUITMENT PROCESS
Paper Screening – Profile review by the Operations team against basic requirements.
L1 Interview – Practice / Operations Team.
L2 Interview – Optional.
Technical Assessment / Final Interview – Customer's Operations Team.
PRE-SCREENING QUESTIONS
What is your highest educational attainment?
How many years of relevant experience do you have in Monitoring, Observability, or SRE roles?
How many years of hands-on experience do you have with Azure Monitor, Application Insights/KQL, and ServiceNow Event Management?
How many years of SRE experience do you have with monitoring and application performance management?
What is your current/last drawn salary?
What is your expected salary?
Are you amenable to working in a hybrid setup in Quezon City, with 2 WFH and 3 RTO days, on a night shift?
When are you available to start if selected?
Role
Software Developer/ Tester
Timings
Night Shift (Permanent)
Industry
BPO
Work Mode
Work from office
Process
Non-Voice
Functional Area
IT Software/Hardware
Note: Myglit doesn't charge any money from candidates. If you have been asked to pay money to get this job then report to us immediately at support@myglit.com.
Interview Tips
- Giving the VNA round?
- What are the most important skills you acquired as a Soft Skills/VNA trainer?
- How would you handle an irate customer?
Similar Jobs
8 - 9 Year(s)
₱ 170 - ₱ 176K p.m
Quezon Calabarzon, Philippines
8 - 13 Year(s)
₱ 100 - ₱ 108K p.m
Quezon Calabarzon, Philippines
Team Leader – UK Financial Collections
Gratitude Inc2 - 4 Year(s)
Confidential
Quezon Calabarzon, Philippines
MIS Executive
Gratitude Inc2 - 4 Year(s)
Confidential
Quezon Calabarzon, Philippines
SIEM Platform Engineer (SIEM & SOAR)
Gratitude Inc2 - 3 Year(s)
Confidential
Quezon Calabarzon, Philippines
Customer Service Representative
Gratitude Inc0 - 1 Year(s)
₱ 18 - ₱ 250K p.m
Quezon Calabarzon, Philippines
0 - 10 Year(s)
₱ 115 - ₱ 121K p.m
Quezon Calabarzon, Philippines
Operations Lead
Gratitude Inc10 - 11 Year(s)
₱ 100 - ₱ 125K p.m
Quezon Calabarzon, Philippines
Multilingual Language - Khmer (Cambodia)
Gratitude Inc0 - 1 Year(s)
₱ 70 - ₱ 114K p.m
Quezon Calabarzon, Philippines

