Any job page, tailored and filled in with one click.A tailored resume, answers and cover letter for any job page, filled in with one click.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

Empowering the next generation with AI education. Custom training for colleges and enterprises.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

SOFTWARE

Site Reliability Engineering (SRE), The Core Engineering, Analyst, Dallas

Goldman Sachs

Senior · 7+ yrsOn-site · Dallas·United StatesListed 1d ago
Apply now

Backed by

VC portfolio

HQ

🇺🇸 New York City, NY, United States

Open roles

40

Experience Senior · 6+ yrs (7–10 years)

About the role

structured by ORI

Site Reliability Engineering (SRE), The Core Engineering, Analyst, Dallaslocation_onDallas, TX, United StatesSite Reliability Engineering (SRE), The Core Engineering, Analyst, DallasApplySite Reliability Engineering (SRE), The Core Engineering, Analyst, Dallaslocation_onDallas, TX, United StatesApplyWHAT WE DO Site…

What you will do

  • Partner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.
  • Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.
  • Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.
  • Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.
  • Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.

What they are looking for

  • Strong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.
  • Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.
  • Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.
  • Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.
  • Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)
  • Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.
  • Knowledge of networking protocols and load balancing strategies in a distributed systems environment.
  • Ability to analyze complex, distributed systems holistically and understand how individual components interact under load.
  • Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.
  • Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.
  • Highly motivated, pro-active and capable of multi-tasking under pressure in a fast-paced environment without compromising quality.
  • Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.

Nice to have

  • Bachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.
  • 7 to 10 years of experience
JavaPythonNode.jsInfrastructure as CodeSDLCalgorithmsdata structuressoftware designTerraformAnsibleCloudFormationDockerKubernetesAWSGCPAzure
Full posting text

Site Reliability Engineering (SRE), The Core Engineering, Analyst, Dallaslocation_onDallas, TX, United StatesSite Reliability Engineering (SRE), The Core Engineering, Analyst, DallasApplySite Reliability Engineering (SRE), The Core Engineering, Analyst, Dallaslocation_onDallas, TX, United StatesApplyWHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help engineer highly reliable, observable, and resilient platforms that support critical business services at scale. You will collaborate with multiple engineering teams to continually improve our production system architecture, facilitate fast delivery of new services, and reduce downtime. This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.Key Responsibilities Partner with engineering leadership to establish service level objectives (SLOs), service level indicators (SLIs), and error budgets.Collaborate with product developers to architect highly available, fault-tolerant, and self-healing systems. Conduct architectural reviews and introduce patterns like circuit breakers, graceful degradation, and rate limiting.Reduce operational toil by building automation, tooling, and self-service capabilities that remove repetitive manual work.Improve production readiness through load testing, performance tuning, capacity forecasting, and reliability reviews.Lead the response to complex, multi-system production incidents. Facilitate blameless post-mortems to identify root causes and drive long-term preventative actions.Promote sustainable operations by helping design healthy on-call models, clear escalation paths, and balanced pager responsibilities. WHAT WE ARE LOOKING FOR Core Technical Skills Strong proficiency in at least one major programming language (e.g., Java, Python, or Node.js) with a focus on writing clean, maintainable code for tooling and automation.Hands-on experience with Infrastructure as Code (IaC) frameworks such as Terraform, Ansible, or CloudFormation.Deep understanding of containerization and orchestration technologies, specifically Docker and Kubernetes (K8s), including service meshes and ingress controllers.Advanced experience with major cloud providers (AWS, GCP, or Azure), specifically building and operating highly resilient cloud-native architectures.Proficiency with Observability stacks, including distributed tracing, logging, and metrics (e.g., Prometheus, Grafana, Splunk, Datadog, OpenTelemetry, ELK, or CloudWatch)Experience with automated testing and SDLC concepts, developing applications in a Linux environment, and sound knowledge of algorithms, data structures and software design.Knowledge of networking protocols and load balancing strategies in a distributed systems environment.Core Competencies & Soft Skills Ability to analyze complex, distributed systems holistically and understand how individual components interact under load.Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.Highly motivated, pro-active and capable of multi-tasking under pressure in a fast-paced environment without compromising quality.Commitment to fostering a blameless culture where failures are treated as opportunities to learn and improve systems.Interest in financial markets and technology.Preferred Qualifications Bachelor’s degree in Computer Science, System Engineering, or a related technical field that involves programming.7 to 10 years of experience ABOUT GOLDMAN SACHS The Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded in 1869, the firm is headquartered in New York and maintains offices in all major financial centers around the world.

Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…
Goldman Sachs

Financial Services

We aspire to be the world’s most exceptional financial institution, united by our shared values of partnership, client service, integrity, and excellence. Operating at the center of capital markets, we act as one firm, mobilizing our people, capital, and ideas to deliver superior results across our clients’ most complex challenges. For 157 years, Goldman Sachs has delivered world-class execution on a global scale across our leading Global Banking & Markets and Asset & Wealth Management businesses. Apprenticeship is central to our culture, with hands-on coaching and access to leaders who bring

Company pageWebsite

More at Goldman Sachs

GBM Private, IB Classic, Investment Banking Vice President - TMT (Digital Infrastructure)

Goldman Sachs

Senior · 4+ yrsFunded

On-site · New York·United States

1d ago
View role
GBM - Private Dept-Bengaluru-Vice President-Software Engineering

Goldman Sachs

Senior · 6+ yrsFunded

On-site · Bengaluru·India

SOFTWARE

1d ago
View role
GBM-Bengaluru-Associate-Software Engineering

Goldman Sachs

3+ yrsFunded

On-site · Bengaluru·India

SOFTWARE

1d ago
View role
GBM - Private Dept-Bengaluru-Associate-Software Engineering

Goldman Sachs

1–3 yrsFunded

On-site · Bengaluru·India

SOFTWARE

1d ago
View role
Apply