EDM Platform Operations, Support Engineer
Quick Summary
Empowering Africa’s tomorrow, together…one story at a time. With over 100 years of rich history and strongly positioned as a local bank with regional and international expertise,
With over 100 years of rich history and strongly positioned as a local bank with regional and international expertise, a career with our family offers the opportunity to be part of this exciting growth journey, to reset our future and shape our destiny as a proudly African group.
Job Summary
The EDM Platform Operations, Support Engineer is responsible for the day-to-day operation, monitoring, support and recovery of the Enterprise Data Management (EDM) platform ecosystem across production and non-production environments. The role is accountable for ensuring the successful execution of operational processes including job scheduling, ETL pipelines, Managed File Transfer (MFT), API integrations, Kafka streaming, cloud-based services, patching, vulnerability remediation, backup and recovery validation, disaster recovery testing and platform monitoring. The role drives operational excellence through proactive monitoring, incident management, service restoration, operational readiness and continuous improvement initiatives.- Oversee the execution of scheduled BAU jobs, ensuring they run successfully and on time.
- Manage CA-WA schedules, dependencies, re-runs, calendar changes and batch recovery activities
- Monitor file deliveries, failed transfers, reconciliation controls, file acknowledgements and exception processing
- Monitor job queues and system alerts to proactively identify and resolve failures or delays.
- Investigate failed jobs, delayed processing, missed schedules and abnormal processing patterns and perform recovery activities where required.
- Execute reruns, reprocessing activities and operational recovery procedures using approved runbooks and support processes
- Maintain operational dashboards and logs to track job performance and system health.
- Collaborate with relevant teams to optimize job scheduling and reduce operational risk.
- Monitor and manage production and non-production EDM platform services and operational processes.
- Support and monitor Managed File Transfer (MFT), ATMDL, ETL and CA-WA scheduling services supporting EDM platform operations.
- Monitor and support data ingestion, transformation, file movement and provisioning processes across the EDM ecosystem.
- Support Kafka streaming services, API integrations and AWS-hosted platform components from an operational perspective.
- Investigate and resolve integration failures impacting file processing, data movement, messaging services and downstream provisioning
- Validate successful end-to-end execution of data movement and processing activities.
- Coordinate and support platform patching, maintenance windows and operational implementation activities.
- Perform post-maintenance and post-patching validation to ensure operational stability.
- Support vulnerability remediation activities and track closure of platform-related security findings.
- Participate in production readiness reviews and validate operational controls before go-live activities.
- Ensure monitoring, alerting, escalation procedures and support documentation are maintained and operationally effective.
- Plan and execute system updates, patches, and upgrades with minimal user disruption.
- Test updates in controlled environments before deployment.
- Ensure systems remain secure, stable, and compliant with standards.
- Support operational activities across Development, Test, UAT and Production environments
- Implement and maintain robust backup and recovery processes.
- Regularly test backup systems to ensure data integrity and availability.
- Execute and validate backup, restore and recovery procedures in accordance with operational standards.
- Participate in disaster recovery exercises and recovery validation activities.
- Support validation of RTO and RPO objectives during recovery testing exercises.
- Maintain disaster recovery documentation, recovery runbooks and operational readiness procedures.
- Maintain and regularly update the Disaster Recovery (DR) plan in collaboration with IT leadership.
- Perform scheduled DR drills to validate recovery procedures and system resilience.
- Ensure critical systems and data are recoverable within defined RTO/RPO parameters.
- Document DR test results and follow up on remediation actions.
- Coordinate with infrastructure and application teams to ensure DR readiness across environments.
- Support business continuity planning by identifying system dependencies and recovery priorities.
- Participate in Major Incident (Sev1/Sev2) response activities, technical triage and service restoration efforts.
- Perform root cause analysis and contribute to corrective and preventative action plans.
- Support problem management activities aimed at reducing recurring incidents and operational failures.
- Coordinate with infrastructure, cloud, vendor and application support teams to resolve operational issues.
- Produce operational reports and service metrics relating to availability, job execution, incidents and platform performance.
- Diagnose and resolve hardware, software, and network issues across platforms.
- Perform root cause analysis and implement corrective actions.
- Escalate complex issues to appropriate teams and follow through to resolution
- Maintain accurate documentation of system configurations, procedures, and policies.
- Update documentation regularly to reflect changes and improvements.
- Contribute to knowledge bases and runbooks for operational consistency.
- Ensure systems adhere to security policies and regulatory requirements.
- Monitor for potential security breaches and collaborate with security teams to mitigate risks.
- Stay current with cybersecurity best practices and apply them proactively.
- Work with cross-functional teams to implement new technologies and system enhancements.
- Participate in project planning, execution, and post-implementation reviews.
- Communicate effectively to align technical solutions with business needs.
- Generate reports on system performance, job execution, and support metrics.
- Analyse trends to identify areas for improvement.
- Present insights to management to support strategic decision-making.
- Provide support to EDM stakeholders, application users and downstream consumers relating to platform operations, processing failures and service requests
- Identify and implement process improvements to enhance operational efficiency.
- Stay informed about emerging technologies and industry trends.
- Foster a culture of innovation and proactive problem-solving.
- Identify opportunities to automate manual operational activities and improve service reliability.
- Improve observability, monitoring, alerting and operational controls across the EDM ecosystem.
- Contribute to knowledge management through creation and maintenance of runbooks, known-error records and operational procedures.
- Drive operational excellence through continuous improvement initiatives and service optimisation recommendations
- Continuously monitor system performance, availability, and resource utilization.
- Use monitoring tools to detect anomalies and prevent service disruptions.
- Report findings and recommend improvements to enhance system reliability.
- Monitor and report platform service health, operational risks, SLA performance and recurring issues
Requirements
~1 min read- Education: Bachelor’s degree in Computer Science, Information Technology, or related field preferred.
- Technical Skills: Proficiency in Windows, Linux, macOS, and cloud platforms (AWS, Azure).
- Networking: Solid understanding of TCP/IP, DNS, DHCP, VPNs, and network troubleshooting.
- Scripting: Experience with PowerShell, Python, or Bash for automation.
- Tools: Familiarity with ticketing systems (e.g., JIRA, ServiceNow) and monitoring platforms.
- Soft Skills: Strong communication, customer service orientation, and problem-solving abilities.
- Certifications: CompTIA A+, Network+, MCSA, or equivalent certifications are a plus.
- Security Awareness: Knowledge of cybersecurity principles and compliance standards.
- Adaptability: Ability to learn and adapt to new technologies in a dynamic environment.
- Project Management: Understanding of project management principles and task prioritization.
Education
Bachelor's Degree: Information TechnologyAbsa Bank Limited is an equal opportunity, affirmative action employer. In compliance with the Employment Equity Act 55 of 1998, preference will be given to suitable candidates from designated groups whose appointments will contribute towards achievement of equitable demographic representation of our workforce profile and add to the diversity of the Bank.
Absa Bank Limited reserves the right not to make an appointment to the post as advertised
Location & Eligibility
Listing Details
- Posted
- July 21, 2026
- First seen
- July 21, 2026
- Last seen
- July 21, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 51%
- Scored at
- July 21, 2026
Signal breakdown
Please let absa know you found this job on Jobera.
4 other jobs at absa
View all →Explore open roles at absa.
Similar Operations Support Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.