Site Reliability Engineer, Global Banking & Markets, Frontline Production Engineering
- Company
- Goldman Sachs
- Location
- New York, NY
- Work type
- Full Time
- Posted
- 2026-08-10
Job description
Who We Look For
As a Frontline SRE, you will be embedded directly within the firm's Global Banking & Markets business. You will play an active role in identifying and resolving inefficiencies in our daily support and operational workflows.
By participating in day-to-day operations, you will diagnose complex system behaviors, coordinate incident responses, manage post-mortems, and oversee safe software deployments. Leveraging your interest in financial markets and the insights gained from managing our production systems, you will directly influence, design, and implement enhancements to our electronic trading product offerings.
We look for creative, collaborative engineers who thrive in fast-paced, dynamic environments. Ideal candidates are innovators and problem-solvers who possess:
A strong passion for automation, system efficiency, and clean code.
Excellent communication skills to collaborate effectively with both technical and non-technical stakeholders.
A proactive mindset focused on continuous improvement and challenging manual processes.
How You Will Fulfill Your Potential
System Ownership: Measure and monitor the availability, latency, performance, and overall health of our production trading services.
Toil Reduction: Design, develop, and maintain software tools and frameworks to automate manual tasks and improve operational efficiency.
Pre-Release Engineering: Support services before they go live through system design consulting, capacity planning, and rigorous release reviews.
Change Management: Safely deploy certified code and new functionalities into the live trading environment.
Collaboration: Act as a key liaison between trading desks, quantitative strategists, and external vendors to resolve day-to-day technical issues.
Continuous Improvement: Drive initiatives to enhance the stability, performance, and disaster-recovery readiness of supported platforms.
Market Microstructure: Understand US Equities and Options market rules to effectively interact with exchanges and resolve execution issues.
Automated Testing: Write and maintain automated checkout tests to validate environment health across trading sessions.
Qualifications
Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related technical field is a plus.
At least 4 years of professional experience in an SRE, Production Engineering, or DevOps role.
Hands-on experience with at least one object-oriented programming language is a plus (e.g., Java, C++) and a scripting language is a must (e.g., Python, Bash).
Proven experience supporting and maintaining high-availability production platforms.
Domain knowledge of US Equities and Options market microstructure and electronic trading.
Solid understanding of the Software Development Life Cycle (SDLC) and version control systems (e.g., Git).
Strong understanding of networking concepts, protocols (e.g., TCP/IP, FIX), and topologies is a plus.
Experience with large-scale automated testing, CI/CD pipelines, and release management.
Excellent written and verbal communication skills.