Senior Software Engineer - Production support
Wells Fargo
- Location
- Bengaluru, India
- Work model
- On-Site
- Level
- Senior
- Posted
- 11h ago
Skills
About this role
About this role: Wells Fargo is seeking a Senior Software Engineer, who has a strong experience in managing and supporting production applications within a 24x7 enterprise environment. The ideal candidate will have expertise in application support, incident and problem management, workload automation, database troubleshooting, Linux administration, reporting platforms, and operational monitoring. This role is focused on maintaining the availability, performance, reliability, and operational health of business-critical systems. The successful candidate will partner closely with development, infrastructure, database, and business teams to proactively identify issues, drive service improvements, and ensure seamless production operations. In this role, you will: Lead moderately complex initiatives and deliverables within technical domain environments Contribute to large scale planning of strategies Design, code, test, debug, and document for projects and programs associated with technology domain, including upgrades and deployments Review moderately complex technical challenges that require an in-depth evaluation of technologies and procedures Resolve moderately complex issues and lead a team to meet existing client needs or potential new clients needs while leveraging solid understanding of the function, policies, procedures, or compliance requirements Collaborate and consult with peers, colleagues, and mid-level managers to resolve technical challenges and achieve goals Lead projects and act as an escalation point, provide guidance and direction to less experienced staff Required Qualifications: 4+ years of Software Engineering experience, or equivalent demonstrated through one or a combination of the following: work experience, training, military experience, education Desired Qualifications: Provide L2/L3 production support for enterprise applications, ensuring adherence to established SLAs and operational standards. Monitor application health, batch processing, interfaces, and data flows to ensure system availability and reliability. Analyze, troubleshoot, and resolve production incidents, service interruptions, application failures, and performance issues. Support and administer enterprise batch scheduling and workload automation processes using AutoSys. Experience supporting large-scale, business-critical applications in a 24x7 production environment. Strong experience in incident response, IT operations, and service management disciplines. Experience supporting ETL, data integration, reporting, and workload automation platforms. Exposure to cloud platforms, DevOps practices, and CI/CD pipelines. Experience working with monitoring, observability, and operational analytics tools. Knowledge of automation and self-healing operational processes. Exposure to AI-driven operations (AIOps), intelligent monitoring, and predictive incident management. Strong troubleshooting and analytical skills. Ability to diagnose complex production issues under pressure. Excellent verbal and written communication skills. Strong stakeholder management and customer engagement capabilities. Ability to prioritize multiple incidents and operational activities. Strong ownership, accountability, and operational mindset. Ability to collaborate effectively across development, infrastructure, database, and business teams. Focus on reliability, service excellence, and continuous improvement. ITIL Foundation Certification. Experience with Splunk, Dynatrace, AppDynamics, Grafana, Prometheus, or similar monitoring platforms. Shell scripting, Python, or automation scripting for operational efficiency. Experience with AIOps, Event Management, and Operational Analytics. Familiarity with Site Reliability Engineering (SRE) concepts and practices. Exposure to cloud operations (Azure, AWS, or GCP). JavaScript and HTML (Support & Debugging Exposure) Log Analysis and Application Diagnostics Production Deployment Support Perform root cause analysis (RCA) for