S

Assoc, Production Eng, WRB Tech

Standard Chartered · Chennai, TN

~₹16L (est.)3–8 yrs experienceShiftPosted 2 days ago
Apply now →

Job description

**#body.unify div.unify-button-container .unify-apply-now:** focus, #body.unify div.unify-button-container .unify-apply-#body.unify div.unify-button-container .unify-apply-now: focus, #body.unify div.unify-button-container .unify-apply- **Requisition Number:** 59281 **Job Location:** Chennai, IND **Global Grade:** Band 7 **Work Type:** Office Working **Employment Type:** Permanent **Posting Start Date:** 30/07/2026 **Posting End Date:** 30/08/2026 : Job Summary The role is accountable for providing tactical and operational support for production services across one or more technology platforms/domains, ensuring optimal service stability, availability, performance, and resilience. Key Responsibilities * Ensure maximum service quality and production stability through rapid and effective response to technical incidents, while driving continual service improvement through trend analysis, problem management, and proactive identification of improvement opportunities. * Manage technical recovery and service restoration for High Severity Incidents (CERT/MIM), service outages, and medium/high severity incidents, providing end-to-end support and implementing timely resolutions within agreed SLAs. * Lead and coordinate incident management activities, including stakeholder communications, escalation management, and technical bridge facilitation during critical incidents. * Perform root cause analysis (RCA) for High Severity and recurring incidents, ensuring corrective and preventive actions are identified, tracked, and implemented to closure. * Own the operational stability, availability, and performance of production systems, directing second- and third-level support teams for problem diagnosis and resolution in accordance with agreed SLAs and OLAs. * Manage production changes, releases, deployments, and rollouts with zero or minimal impact to business services. Ensure comprehensive implementation, validation, rollback, contingency, and communication plans are in place for all production activities. * Review and assess the impact of dependent changes across applications, infrastructure, databases, middleware, cloud platforms, and networks to minimize production risk. * Drive proactive monitoring and event management by identifying opportunities for automation, alert optimization, early issue detection, and operational efficiency improvements. * Support capacity management, resiliency testing, disaster recovery (DR), and business continuity planning (BCP) activities to ensure operational readiness. * Create, maintain, and continuously improve Production Engineering documentation, operational procedures, runbooks, recovery guides, knowledge articles, and contingency plans. * Ensure adherence to operational governance, security, compliance, audit, and risk management requirements across supported services. * Provide inputs to the PE Manager for operational dashboards and service reviews, including incident trends, problem trends, availability metrics, service improvement plans (SIPs), RCA action tracking, and platform health indicators. * Collaborate with application development, infrastructure, security, architecture, and business teams to improve reliability, reduce technical debt, and enhance service resilience. * Participate in and support cross-training, knowledge transfer, mentoring, and capability-building activities within the Production Engineering organization. * Identify and drive opportunities for automation, shift-left initiatives, operational simplification, and reduction of manual effort to improve service reliability and support efficiency. * Act as a technical lead during major incidents, complex problem investigations, production go-lives, and critical business events, providing technical guidance and decision-making support to internal and external stakeholders. Strategy * To be accountable to execute the strategy devised for the business unit. Business * Fully accountable in incident, problem, change and risk management which relates to the production application/system. Processes * Create, Review and update Production Engineering documentation. Update of contingency (DR/BCP) documentation and processes People \& Talent * Participate in cross-training and knowledge transfer activities within support teams Risk Management * Responsible to proactively identify the risks in the application and manage the mitigation actions. Responsible for managing, tracking and timely closure of risks and other compliance related issues in Riskwise (Information Security risks) \& M7 (Operational Risks). Governance * Provide inputs to management for monthly dashboard that provide information on incident and problem trends along with SIP and RCA Action Items. Regulatory \& Business Conduct * Display exemplary conduct and live by the Group's Values and Code of Conduct. * Take personal responsibility for embedding the highest standards of ethics, including regulatory and business conduct, across Standard Chartered Bank. This includes understanding and ensuring compliance with, in letter and spirit, all applicable laws, regulations, guidelines and the Group Code of Conduct. * Effectively and collaboratively identify, escalate, mitigate and resolve risk, conduct and compliance matters. Key stakeholders * PE Manager / PE Lead * Application Development Teams * Business \& Product Stakeholders * CIO / Technology Leadership * Infrastructure, DBA \& Platform Teams * Security, Risk \& Compliance Teams * Change, Release \& Service Management Teams (ITCRM) * MIM / CERT Teams * External Vendors and Service Providers Other Responsibilities * Embed Here for good and Group's brand and values in the SRE team; Perform other responsibilities assigned under Group, Country, Business or Functional policies and procedures; Multiple functions (double hats). Skills and Experience * AWS * Oracle * Linux * Kubernetes * API Qualificat