Tailored answers, filled into supported job forms.Your tailored resume and answers, filled into supported job forms for you to review.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

OneRoadmap is a career platform built around ORI, its AI career agent. ORI finds overlooked job opportunities, matches them to your profile and shows the skill gaps to close, with roadmaps, challenges, job simulations and certifications to close them. When you are ready, it prepares a tailored resume, application answers and an application strategy, with a cover letter where the application asks for one. The OneRoadmap Chrome extension fills supported application forms for you to review and submit, and you keep track of every application in one place.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • About
  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

SOFTWARE · Engineering

Senior Infrastructure Engineer, Data Compute Platform

Grab

Senior · 2+ yrsRemote · WorldwideListed 4d ago
Apply now

Backed by

Lightspeed India

HQ

🇮🇳 Singapore, Singapore

Open roles

45

Experience Senior · 6+ yrs (2+ years)

About the role

structured by ORI

Senior Infrastructure Engineer, Data Compute Platform Team Engineering Location Petaling Jaya, Malaysia Job Type Full-time The role at a glance Get to Know the TeamThe Data Compute Platform team is an important contributor to Grab's data ecosystem, allowing growth through democratization of data at scale. We build…

What you will do

  • You will design, build, and operate the multi-tenant Kubernetes (EKS) platform.
  • You will build Kubernetes operators and custom resources (kubebuilder / controller-runtime, Go) that automate the provisioning and lifecycle of compute engines and their tenants.
  • You will lead the infrastructure-as-code (Terraform) and CI/CD (GitLab) that provision and change our AWS estate.
  • You will design and implement the platform's identity, access and security model across AWS IAM, Kubernetes RBAC and service identities, working with the storage access and security teams.
  • You will build the observability, alerting, capacity planning, and incident tooling for the platform.

What they are looking for

  • Software Engineering, Computer Science, or related undergraduate degree.
  • You have 3 or more years of experience, with at least 2 years building and operating production infrastructure or platform services at scale.
  • You have programming proficiency in Go and/or Python, with the habit of treating infrastructure as software: tested, reviewed, versioned and automated.
  • You have deep, hands-on Kubernetes experience in production: cluster operations, scheduling, autoscaling, networking, storage, RBAC and multi-tenancy. We prefer experience building custom controllers or operators.
  • You have experience with AWS (EKS, EC2, S3, IAM, VPC) and infrastructure as code with Terraform.
  • You have proficiency in CI/CD tooling (GitLab CI, Jenkins or similar) and GitOps-style delivery.
  • You have solid SRE fundamentals: observability (Datadog, Prometheus, Grafana or equivalent), SLOs, incident management and capacity planning, with experience running reliable services.

Nice to have

  • Working knowledge of at least one of Spark, Ray, Airflow, Trino or Starrocks, and a appetite to learn how distributed data engines behave on Kubernetes.
  • Experience running Apache Spark on Kubernetes at scale (Spark Operator, dynamic allocation, shuffle services) and tuning its interaction with the resource manager.
  • Experience with Kubernetes autoscaling and scheduling tooling such as Karpenter, Cluster Autoscaler, Yunikorn or Volcano.
  • Experience deploying and operating Ray on Kubernetes (KubeRay), including cluster autoscaling and isolation for ML and batch workloads.
  • Experience operating distributed query engines like Trino or Starrocks, including cluster sizing, fault tolerance and workload isolation.
  • Experience with container runtimes and internals (containerd, cgroups, OverlayFS, image build and distribution).
  • Experience with service mesh, ingress and network policy (Istio, Envoy, Cilium or similar) in multi-tenant clusters.
  • Experience with FinOps for compute platforms: cost attribution, and spot / reserved capacity strategy.

Benefits

  • Term Life Insurance and comprehensive Medical Insurance
  • GrabFlex customizable benefits package
  • Parental and Birthday leave
  • Love-all-Serve-all (LASA) volunteering leave
  • Confidential Grabber Assistance Programme
GoPythonKubernetesCI/CDSREGitOpsObservabilityCapacity planningAWSEKSTerraformGitLabS3
Full posting text

Senior Infrastructure Engineer, Data Compute Platform Team Engineering Location Petaling Jaya, Malaysia Job Type Full-time The role at a glance Get to Know the TeamThe Data Compute Platform team is an important contributor to Grab's data ecosystem, allowing growth through democratization of data at scale. We build and operate Grab's data infrastructure and efficient platform that supports internal data processes and company-wide data lake access. Our tech stack uses industry-leading distributed compute engines like Apache Spark, Ray, Trino, and Starrocks, orchestrated by Airflow and Michelangelo and backed by AWS S3. Our evolving Data Lake storage architecture uses modern open-source formats like Apache Iceberg and Delta in addition to traditional Apache Hive Parquet tables.Underneath these engines sits a large, multi-tenant Kubernetes and AWS Infrastructure that we own end to end. This infrastructure includes the clusters, the operators, the autoscaling, the networking, the identity and access model, the observability, and the cost controls. These components work together to keep the platform fast, reliable, and affordable for thousands of pipelines and queries every day. Get to Know the RoleYou will be an important contributor to the infrastructure layer of the Data Compute Platform. This layer consists of the Kubernetes, AWS, and infrastructure-as-code foundations. Spark, Ray, Trino, Starrocks, Airflow, and Michelangelo run on these foundations. You will design how compute is provisioned, scaled, secured, observed and paid for, and you will drive the reliability and cost-efficiency of the platform as it grows. As a senior engineer, you will take ownership of well-scoped infrastructure projects from design through rollout and operation. You will also contribute to the team's engineering and SRE standards. Additionally, you will support other engineers through code review and knowledge sharing. You will also explore new developments in the cloud-native and data infrastructure space and integrate them into our ecosystem to the benefit of the data community at Grab.You will report to our Data Engineering Manager II, and you will based onsite in our Petaling Jaya office.The Critical Tasks You Will PerformYou will design, build, and operate the multi-tenant Kubernetes (EKS) platform. This platform runs Grab's workloads, including Spark, Ray, Trino, Starrocks, Airflow, and Michelangelo. Additionally, your responsibilities will include cluster lifecycle, node provisioning, and autoscaling, and scheduling and resource isolation.You will build Kubernetes operators and custom resources (kubebuilder / controller-runtime, Go) that automate the provisioning and lifecycle of compute engines and their tenants.You will lead the infrastructure-as-code (Terraform) and CI/CD (GitLab) that provision and change our AWS estate. This estate includes EKS, S3, IAM, RDS, VPC, and networking. You will drive it towards safe, reviewable, automated change.You will design and implement the platform's identity, access and security model across AWS IAM, Kubernetes RBAC and service identities, working with the storage access and security teams.You will build the observability, alerting, capacity planning, and incident tooling for the platform. You will contribute to the SRE practice, which includes SLOs, runbooks, on-call, and post-incident reviews. You will reduce toil and MTTR.You will own compute cost efficiency: instance and storage strategy, spot and right-sizing, bin-packing, idle reclamation, and cost attribution back to tenants.You will drive architectural improvements and migrations (for example engine version upgrades, cluster consolidation, new execution backends), managing the design, phased rollout and rollback plan with guidance from senior team members. Read more Skills you need What Essential Skills You Will NeedSoftware Engineering, Computer Science, or related undergraduate degree.You have 3 or more years of experience, with at least 2 years building and operating production infrastructure or platform services at scale.You have programming proficiency in Go and/or Python, with the habit of treating infrastructure as software: tested, reviewed, versioned and automated.You have deep, hands-on Kubernetes experience in production: cluster operations, scheduling, autoscaling, networking, storage, RBAC and multi-tenancy. We prefer experience building custom controllers or operators.You have experience with AWS (EKS, EC2, S3, IAM, VPC) and infrastructure as code with Terraform.You have proficiency in CI/CD tooling (GitLab CI, Jenkins or similar) and GitOps-style delivery.You have solid SRE fundamentals: observability (Datadog, Prometheus, Grafana or equivalent), SLOs, incident management and capacity planning, with experience running reliable services.Skills that are Good to haveWorking knowledge of at least one of Spark, Ray, Airflow, Trino or Starrocks, and a appetite to learn how distributed data engines behave on Kubernetes.Experience running Apache Spark on Kubernetes at scale (Spark Operator, dynamic allocation, shuffle services) and tuning its interaction with the resource manager.Experience with Kubernetes autoscaling and scheduling tooling such as Karpenter, Cluster Autoscaler, Yunikorn or Volcano.Experience deploying and operating Ray on Kubernetes (KubeRay), including cluster autoscaling and isolation for ML and batch workloads.Experience operating distributed query engines like Trino or Starrocks, including cluster sizing, fault tolerance and workload isolation.Experience with container runtimes and internals (containerd, cgroups, OverlayFS, image build and distribution).Experience with service mesh, ingress and network policy (Istio, Envoy, Cilium or similar) in multi-tenant clusters.Experience with FinOps for compute platforms: cost attribution, and spot / reserved capacity strategy.Contributions to open-source cloud-native or data infrastructure projects. Read more What we offer About Grab and Our WorkplaceGrab is Southeast Asia's leading superapp. From getting your favourite meals delivered to helping you manage your finances and getting around town hassle-free, we've got your back with everything. In Grab, purpose gives us joy and habits build excellence, while harnessing the power of Technology and AI to deliver the mission of driving Southeast Asia forward by economically empowering everyone, with heart, hunger, honour, and humility. Read more Life at Grab Life at GrabWe care about your well-being at Grab, here are some of the global benefits we offer:We have your back with Term Life Insurance and comprehensive Medical Insurance.With GrabFlex, create a benefits package that suits your needs and aspirations.Celebrate moments that matter in life with loved ones through Parental and Birthday leave, and give back to your communities through Love-all-Serve-all (LASA) volunteering leaveWe have a confidential Grabber Assistance Programme to guide and uplift you and your loved ones through life's challenges.Balancing personal commitments and life's demands are made easier with our FlexWork arrangements such as differentiated hoursWhat We Stand For At GrabWe are committed to building an inclusive and equitable workplace that provides equal opportunity for Grabbers to grow and perform at their best. We consider all candidates fairly and equally regardless of nationality, ethnicity, race, religion, age, gender, family commitments, physical and mental impairments or disabilities, and other attributes that make them unique. Read more What’s it like to work at Grab? Play video View our location For over a decade, Grab has seamlessly connected Malaysian people to a range of services—food, groceries, rides, and more. Take a tour and get to know this one-stop destination for making life more convenient for the local community. [Click to enlarge image] [Click to enlarge image] [Click to enlarge image] [Click to enlarge image] [Click to enlarge image] Meet the Engineering team From software developers to security specialists, our people thrive on experimenting with fresh ideas and implementing them for users. Solve real-world problems, find purpose in your work and work across borders to help drive Southeast Asia forward. Engineering at Grab is synonymous with holding yourself accountable for impacting the lives of millions of partners, be it drivers, passengers or merchants. Every line of code can have a big positive or negative impact on our consumers, which is why every little change requires thorough thought processes and a deep understanding of systems and scale. “ ” Learning within the Engineering team never stops, especially with the sponsored learning resources. You need to stay curious and be willing to take ownership to join and succeed in the team. “ ” Everybody at Grab understands the necessity of balance. We all believe that what we’re doing here is important, we’re genuinely changing people’s lives for the better, but we also understand that this shouldn’t come at the cost of your personal life. “ ” The Application Process What can you expect when you apply for a role at Grab? Here’s what our hiring process looks like. Step 1: Apply. We have a range of open roles from Marketing to Engineering and everything in between. Find the one that’s best for you and show us why you’re our next teammate. Step 2: Getting to know you. Next, you’ll have a video or phone interview with one of our recruiters to chat about you, the role, and your past work experiences. Step 3: Testing, testing… 123. Some roles might require assessments and/or a technical competency interview. For other roles, there might be an additional interview – speak with your recruiter for application details. Step 4: Is it a culture match? Finally, we’ll find out how we fit together, if your career vision aligns with our mission, how you embody the 4 H's, and your working style and personality.

Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…
Grab

Software Development

Grab is Southeast Asia’s leading superapp, offering a suite of services consisting of deliveries, mobility, financial services, enterprise and others. Grabbers come from all over the world, and we are united by a common mission: to drive Southeast Asia forward by creating economic empowerment for everyone. At Grab, every Grabber is guided by The Grab Way, which explains our mission and the operating principles on how we can achieve it together. We call these principles the 4Hs: Heart We work together as OneGrab to serve communities in Southeast Asia Hunger We work to understand ground truths a

Backed by Lightspeed India

Company pageWebsite

More at Grab

GrabData AnalyticsFunded

On-site · Singapore, Singapore

Senior · 2+ yrs · 1d ago

Analytics Manager II, Financial Services
product analytics+13 more
View role
GrabData AnalyticsFunded

On-site · Jakarta, Indonesia

2+ yrs · 1d ago

Business Analytics Specialist
AnalyticsStatistics+12 more
View role
GrabEngineeringFunded

On-site · HCMC, Vietnam

Senior · 1d ago

Senior Software Engineer, Backend - Agent Experience
GoAlgorithmsData Structures+9 more
View role
GrabEngineeringFunded

On-site · Singapore, Singapore

Senior · 1d ago

Senior Software Engineer
Software Engineering+15 more
View role
Apply