Senior Lead Software Engineer - DevOps with Equities Electronic Trading
JPMorgan Chase
- Location
- NY, United States
- Work model
- On-Site
- Level
- Senior
- H-1B history
- 1,524 approvals (FY2023)
- Posted
- 4h ago
Skills
About this role
We have an opportunity to impact your career and provide an adventure where you can push the limits of what's possible. As a Lead Software Engineer at JPMorgan Chase within the Commercial & Investment Banking - Markets Tech - Trading / Derivatives Execution Tech team, you are an integral part of an agile team that works to enhance, build, and deliver trusted market-leading technology products in a secure, stable, and scalable way. As a core technical contributor, you are responsible for conducting critical technology solutions across multiple technical areas within various business functions in support of the firm’s business objectives. This position will support the reliability, performance, and operational integrity of electronic and equities trading systems , with a specific focus on FIX protocol connectivity . This role is hands-on and operations-oriented, partnering closely with trading, technology, and development teams to ensure stable order flow, rapid incident response, and disciplined change execution. The position emphasizes Python automation , Linux troubleshooting , and Grafana-based observability , with C++ exposure used primarily to investigate issues and collaborate effectively with Application Job responsibilities Executes creative software solutions, design, development, and technical troubleshooting with ability to think beyond routine or conventional approaches to build solutions or break down technical problems Provide daily production support for electronic trading platforms, including FIX sessions/protocols , connectivity health, and order/trade workflow stability Monitor system health and trading-impacting signals using Grafana dashboards and alerting to improve visibility with latency, errors, throughput, and availability Lead incident triage and restoration activities during service degradation, including structured troubleshooting, stakeholder communications, and post-incident follow-up Perform root cause analysis on recurring issues and implement durable remediation, including runbook improvements, alert tuning, and operational automation Develops secure high-quality production code with reviewing and debugging code by using Python scripts and tools for health checks, operational workflows, reporting, and environment validation (per user-provided role intent) Operate and support Kafka on AWS (e.g., Amazon MSK or equivalent) with topic/partition design, consumer groups, throughput/latency tuning, security/permissions, and production troubleshooting Drives team adoption of enterprise-authorized AI-assisted engineering practices within the work environment to improve code quality, delivery speed, and operational outcomes (e.g., AI-assisted code review / refactoring, test strategy acceleration, incident/root-cause analysis support), while establishing consistent validation standards (secure coding, peer review, automated testing) and promoting reuse of effective patterns across the team Applies knowledge of tools within the Software Development Life Cycle toolchain, including enterprise-authorized AI-assisted development and automaton capabilities, to improve the value realized by automation Troubleshoot Linux based systems using logs, process and resource diagnostics, and network-level checks relevant to connectivity and application behavior (per user-provided role intent) Partner with development teams to investigate complex issues in trading components with read logs, traces, diagnostic output and the ability to interpret and discuss findings in contexts where components are implemented in C++ Required qualifications, capabilities, and skills Formal training or certification on Software engineering concepts and 5+ years applied experience Advanced in one or more programming language(s), framework(s) and tools (e.g., Python , C++, Bash Scripting , Grafana, etc.) Demonstrated experience in DevOps and or SRE in a mission-critical, low-latency environment, with accountability for uptime and incident