(ACTIVE) Site Reliability Engineer | Hybrid Makati City
Job Description
Key Skills
4 candidate(s) have already applied for this Job. Apply now
Site Reliability Engineer (SRE)
Location: Eton Centris, Quezon Avenue, Quezon City, Manila, Philippines
Work Setup: Hybrid (3 Days Onsite, 2 Days WFH)
Shift: Night Shift
Employment Type: Full-Time
Start Date: ASAP
Job Summary
We are looking for an experienced Site Reliability Engineer (SRE) with strong expertise in cloud observability, application performance monitoring, and IT operations. The ideal candidate will have hands-on experience with Azure Monitor, Application Insights, KQL, ServiceNow Event Management, and monitoring automation.
The role focuses on improving platform reliability, establishing enterprise monitoring standards, reducing alert noise, accelerating incident response, and implementing proactive monitoring and self-healing solutions.
Key Responsibilities
Design and implement enterprise monitoring, observability, and event management standards across applications and platforms.
Develop monitoring solutions using Azure Monitor, Application Insights, Log Analytics (KQL), Grafana, and APM tools.
Establish reference architectures, operational runbooks, and event management models to improve incident detection and resolution.
Collaborate with product teams to implement Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
Configure and optimize ServiceNow ITOM Event Management for effective alert correlation, event management, and incident routing.
Identify automation opportunities to reduce alert noise, improve incident response, and enable self-healing.
Implement proactive monitoring using telemetry, synthetic transactions, and application performance monitoring tools.
Work closely with IT Operations, cloud platform, cybersecurity, network, and product teams.
Contribute to monitoring and observability strategies, governance standards, and continuous improvement initiatives.
Analyze postmortem findings to improve monitoring patterns, reduce Mean Time to Resolution (MTTR), and strengthen operational reliability.
Coach technical teams on monitoring standards, observability practices, and operational best practices.
Qualifications & Requirements
Bachelor's degree in Information Technology, Computer Science, Engineering, or a related field.
Minimum 3 years of hands-on experience in monitoring, observability, or SRE roles using Azure Monitor, Application Insights (KQL), and ServiceNow Event Management.
At least 5 years of experience in Site Reliability Engineering, with a strong understanding of monitoring and application performance management.
Strong expertise in Azure Log Analytics, KQL, telemetry, and APM implementations.
Hands-on experience with ServiceNow ITSM and ITOM Event Management.
Knowledge of Grafana, Prometheus, AppDynamics, and ThousandEyes.
Experience implementing SLOs, SLIs, and monitoring governance standards. Familiarity with platforms such as Nobl9 is preferred.
Understanding of network architecture and security, including WAN/LAN, TCP/IP, and PKI.
Experience with proactive monitoring, synthetic transactions, and monitoring automation.
Familiarity with AIOps concepts, ITSM processes, compliance requirements, and CMDB integration.
Strong analytical, problem-solving, communication, and cross-functional collaboration skills.
Ability to translate complex technical requirements into practical monitoring standards and operational improvements.
Willingness to work a hybrid schedule in Quezon City on a night shift.
Required Technical Skills
Azure Monitor, Application Insights, and Azure Log Analytics
Kusto Query Language (KQL)
ServiceNow ITOM Event Management and ITSM
Grafana, Prometheus, AppDynamics, and ThousandEyes
Cloud Observability and Application Performance Monitoring (APM)
Site Reliability Engineering (SRE)
SLI/SLO Management
Monitoring Automation and Self-Healing
Incident Management and Root Cause Analysis
Telemetry and Synthetic Monitoring
AIOps and IT Operations
Recruitment Process
Paper Screening: Initial profile evaluation by the Operations team.
L1 Interview: Technical interview with the Practice/Operations team.
L2 Interview: Additional interview, if required.
Technical Assessment / Final Interview: Conducted by the customer's Operations team.
Pre-Screening Questions
What is your highest educational qualification?
How many years of experience do you have in monitoring, observability, or SRE roles involving Azure Monitor, Application Insights, KQL, and ServiceNow Event Management?
How many years of experience do you have in Site Reliability Engineering and Application Performance Management?
What is your current or last drawn salary?
What is your expected monthly salary?
Are you willing to work a hybrid schedule (3 days onsite and 2 days WFH) in Quezon City on a night shift?
What is your earliest available joining date?
Role
Systems Administration
Timings
Night Shift (Permanent)
Industry
IT-Software / Software Services
Work Mode
Hybrid
Process
Non-Voice
Functional Area
IT Software/Hardware
Note: Myglit doesn't charge any money from candidates. If you have been asked to pay money to get this job then report to us immediately at support@myglit.com.
Interview Tips
- Giving the VNA round?
- What are the most important skills you acquired as a Soft Skills/VNA trainer?
- How would you handle an irate customer?
Similar Jobs
2 - 3 Year(s)
₱ 110 - ₱ 116K p.m
Manila, Philippines
5 - 10 Year(s)
Confidential
Manila, Philippines
Security Engineer – SIEM / SOC (CL8)
Gratitude Inc5 - 8 Year(s)
Confidential
Manila, Philippines
Malware Engineer Specialist / Team Lead – CL 9
Gratitude Inc3 - 10 Year(s)
Confidential
Manila, Philippines
5 - 10 Year(s)
Confidential
Manila, Philippines
Manager, pre-sales & solutions - service desk
Gratitude Inc5 - 10 Year(s)
Confidential
Manila, Philippines
7 - 10 Year(s)
₱ 140 - ₱ 150.7K p.m
Manila, Philippines
2 - 5 Year(s)
Confidential
Manila, Philippines
Microsoft MPOS Developer (Dynamics 365 Retail)
Gratitude Inc5 - 8 Year(s)
₱ 170 - ₱ 180K p.m
Manila, Philippines

