Business Function
Group Technology enables and empowers the bank with an efficient, nimble, and resilient infrastructure through a strategic focus on productivity, quality & control, technology, people capability, and innovation. In Group Technology we manage the majority of the Bank's operational processes and inspire to delight our business partners through our multiple banking delivery channels.
Responsibilities:
- Production Support & Incident Management: Provide expert-level support for critical trading applications, addressing BAU (Business As Usual) issues for Front Office, Back Office, and Finance users. Lead incident response, troubleshoot complex technical and functional issues, perform root cause analysis, and implement effective long-term solutions to prevent recurrence.
- Reliability & Performance Optimization: Design and implement solutions to improve system reliability, availability, scalability, and performance of trading platforms. Proactively identify and address potential issues through continuous monitoring and analysis.
- Infrastructure Management: Work on infrastructure-related tasks including OS patching, database patching, and setting up UAT environments with production data dumps, ensuring data integrity and system readiness.
- Automation & Tooling: Develop and implement automation tools and scripts using Python to reduce manual operational tasks and improve efficiency. This includes automating health checks, operational tasks, and contributing to CI/CD pipelines.
- Monitoring & Observability: Implement and enhance monitoring, alerting, and logging systems to ensure comprehensive visibility into application health and performance. Define and measure Service Level Objectives (SLOs) and Service Level Indicators (SLIs).
- Collaboration & Communication: Act as a primary point of contact for users, effectively communicating status updates and resolution plans during live issues. Collaborate closely with development, infrastructure, and other technology teams, as well as external vendors, to drive issue resolution and system enhancements.
- Release and Change Management: Support production releases, system changes, and perform post-rollout validation to ensure seamless integration into production environments.
- Documentation & Knowledge Sharing: Maintain comprehensive documentation for support processes, system configurations, and troubleshooting guides. Contribute to team knowledge sharing and best practices.
Required Skills & Experience:
- Experience:
- Minimum of 4-6 years of experience in IT production support within financial institutions, with a strong focus on trading applications.
- Demonstrated experience in a Site Reliability Engineering (SRE) focused role or with significant SRE responsibilities, including production optimization.
- Experience with front-office trading environments is highly desirable.
- Technical Skills:
- Proficiency in Unix/Linux operating systems, including shell scripting.
- Strong SQL skills for database querying and analysis (e.g., MySQL, Oracle, MS SQL).
- Proficiency in Python for scripting and automation.
- Experience with OpenShift or Kubernetes environments.
- Familiarity with CI/CD pipeline tools (e.g., Jenkins, GitLab, Harness, Azure DevOps) is a strong advantage.
- Knowledge of monitoring tools (e.g., Splunk, Grafana, AppDynamics, Dynatrace, ITRS Geneos) is a plus.
- Domain Knowledge:
- Good understanding of financial markets, trading concepts, asset classes, and trade lifecycle.
- Soft Skills:
- Excellent problem-solving and analytical skills with a detail-oriented approach.
- Strong communication skills (verbal and written) for interacting with technical teams and business users.
- Ability to work in a fast-paced, high-pressure financial environment.
- Proactive, self-motivated, and a strong team player.
Location:
DBS Asia Central
Job:
Technology
Schedule:
Regular
Employee Status:
Full time