workday
Posted 4 days ago
IT Event Management Engineer
Job description
With 75 years of experience, our focus is on helping the most vulnerable children overcome poverty and experience fullness of life. We help children of all backgrounds, even in the most dangerous places, inspired by our Christian faith. Come join our 31,000+ staff working in nearly 100 countries and share the joy of transforming vulnerable children’s life stories! Key
Responsibilities
The IT Event Management Engineer is responsible for designing, implementing, and continuously improving global event detection and alerting systems that ensure proactive identification and resolution of IT issues before they impact users or services. This position operates at the intersection of engineering, automation, and service management, driving reliability, consistency, and operational excellence across hybrid cloud and on-prem environments. The role is foundational to achieving a proactive, data-driven IT operations model—helping transition the organization from reactive firefighting to predictive service assurance. Core Competencies - Technical Proficiency: Understanding of ITOM systems, monitoring frameworks, APIs, and integrations. - Process Orientation: Ability to document, standardize, and institutionalize best practices through templates and guides. - Collaboration: Works effectively across infrastructure, application, and service management domains. - Analytical Thinking: Strong data interpretation skills to identify event trends and patterns. - Innovation: Drives automation, simplification, and modernization of event workflows. - Customer Focus: Designs alerting systems that directly enhance reliability, transparency, and user satisfaction.
Qualifications
&
Experience
Education
& Certifications Bachelor’s degree in Information Technology, Computer Science, Engineering, or related field ITIL Foundation (minimum); ITIL Intermediate or ITIL 4 Managing Professional is an advantage Relevant certifications in cloud platforms (Azure, AWS) or monitoring tools are desirable
Experience
5+ years’ experience in IT Operations, Monitoring, or Event Management roles Proven experience working with enterprise monitoring and event management tools (e.g., OpsBridge, Azure Monitor, AWS CloudWatch, Site24x7)
Experience
integrating monitoring tools with ITSM platforms for automated incident management Hands-on experience in automation and scripting (e.g., PowerShell, Python, Ansible) Exposure to hybrid environments (cloud + on-prem infrastructure) Technical & Functional Skills Strong understanding of ITOM and monitoring frameworks, including event correlation, alert tuning, and noise reduction
Experience
designing or operating event-driven systems for proactive detection and service assurance Knowledge of APIs and system integrations
Experience
with dashboards, analytics, and reporting tools (e.g., Power BI, Kibana) Familiarity with ITSM processes (Incident, Problem, Change Management) Core
Responsibilities
1. Event Detection, Ingestion & Correlation • Design and maintain event ingestion pipelines from multiple monitoring sources (e.g., Azure Monitor, AWS CloudWatch, network devices, applications, SaaS systems). • Develop correlation logic and rules to identify related alerts and minimize redundant or noisy notifications. • Maintain event taxonomies and classification standards to ensure consistent event tagging, severity, and categorization across systems. 2. Automation, Orchestration & Remediation • Build and maintain automation scripts and workflows to automatically detect and remediate known issues (e.g., restarting services, clearing caches, resizing disks). • Integrate event management systems with ITSM platforms (e.g., ServiceNow, SMAX) to auto-create and route incidents with contextual data. • Participate in AIOps initiatives—leveraging predictive analytics and machine learning models to forecast incidents and anomalies. 3. Standardization, Templates & Documentation • Develop standard operating procedures (SOPs), runbooks, and knowledge articles for consistent event triage and escalation processes. • Create event configuration templates (e.g., for threshold settings, escalation rules, integration blueprints) to ensure monitoring practices are repeatable and scalable. • Maintain a Monitoring and Event Management Playbook outlining governance, workflows, and automation frameworks. • Document integration patterns, naming conventions, and API schema mappings to enable faster onboarding of new systems. • Ensure all documentation is version-controlled, accessible via Confluence or SharePoint, and updated as systems evolve. 4. Operational Effectiveness & Continuous Improvement • Conduct routine health checks on event management systems to ensure optimal performance, data accuracy, and integration stability. • Analyze event and incident data trends to identify gaps, redundancies, or opportunities for improvement. • Partner with Service Desk, Cloud, and Network teams to optimize event thresholds, escalation rules, and notification logic. • Drive continuous improvement initiatives—reducing false positives and increasing actionable alerts through tuning and refinement. 5. Governance & Quality Assurance • Enforce standardized event handling procedures across global and regional IT teams. • Support governance reviews and audits to demonstrate compliance with ITIL Event Management and ITOM standards. • Ensure event management aligns with organizational KPIs such as system uptime, MTTD, MTTR, and service reliability targets. • Support the development of a Monitoring Maturity Framework to assess and elevate event management capabilities globally. 6. Collaboration & Knowledge Transfer • Partner with technical teams to integrate new services and platforms into the event management ecosystem. • Conduct knowledge transfer sessions, workshops, and training for regional IT and operations staff to embed best practices. • Participate in Agile ceremonies and cross-functional collaboration to align event management efforts with larger transformation initiatives. • Act as a subject matter expert (SME) for monitoring and observability in project design and operational readiness reviews. 7. Tool Evaluation & Innovation • Evaluate emerging monitoring and observability tools to identify opportunities for consolidation or modernization. • Support proof-of-concept activities, benchmark tool performance, and provide recommendations for technology roadmaps. • Contribute to architectural discussions regarding hybrid monitoring strategies, especially in multi-cloud contexts. Applicant Types Accepted: Local Applicants Only World Vision is a Christian humanitarian organisation with a mission centred on following Jesus Christ in service to the world´s most vulnerable children. Therefore, in all locations to the fullest extent legally permissible, the successful applicant will affirm our core documents, observe conduct compatible with Christian principles, serve at a high level of professional ethics and strive to act in accordance with cultural sensitivities. Furthermore, regular attendance with team and office devotions, chapel and prayer gatherings are expected in line with policies in the World Vision host location and its departments. Our vision for every child, life in all its fullness. Our prayer for every heart, the will to make it so. As a global Christian relief, development and advocacy organisation, our focus is on helping the most vulnerable children overcome poverty and experience fullness of life. We help children of all backgrounds, even in the most dangerous places, inspired by our Christian faith. Learn more about our work at wvi.org. Our organisational culture reflects a "Partnership" of World Vision offices in 100 countries and 31,000+ staff working towards one vision: life in its fullness for every child. A career with World Vision is a God-given calling, and we believe that every staff member has been brought to World Vision for God’s purposes. Whether working from home, in an office, or with children and community members, we celebrate and embrace each staff member’s diverse background and talents – knowing that together, we can make a difference. Together, #WeAreWorldVision. Learn more about our culture. Our people are our greatest asset. Each staff member brings their unique experience and God-given talents to the organisation – and in return World Vision provides employees a competitive "Total Rewards" package tailored to the context in which they work. Learn more about benefits. Have questions about applying to a job with World Vision? See our Frequently Asked Questions.
Skills and functions
- Aws
- Azure
- Machine Learning
- Python