Tailored answers, filled into supported job forms.Your tailored resume and answers, filled into supported job forms for you to review.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

OneRoadmap is a career platform built around ORI, its AI career agent. ORI finds overlooked job opportunities, matches them to your profile and shows the skill gaps to close, with roadmaps, challenges, job simulations and certifications to close them. When you are ready, it prepares a tailored resume, application answers and an application strategy, with a cover letter where the application asks for one. The OneRoadmap Chrome extension fills supported application forms for you to review and submit, and you keep track of every application in one place.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • About
  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

AI / ML · Predictive Science

Data Scientist Lead [Multiple Positions Available]

JPMorgan Chase

Senior · 7+ yrsOn-site · Jacksonville, FL, United StatesListed 2d ago
Apply now

Backed by

VC portfolio

HQ

🇺🇸 New York City, NY, United States

Open roles

1996

How to stand out for Data Scientist Lead [Multiple Positions Available] at JPMorgan Chase

Auto Match agent

Let ORI find you the best jobs.

Set up your Auto Match agent once - your target role, level and where you want to work. It searches every day, scores each opening against your profile and resume, and delivers the ones worth applying to, with a prepared application a click away.

Searches every day Scored against your profile Applications prepared for you
Sign in & set up Auto Match agent

Resume & career call

Get your resume reviewed for this role - 30-minute 1:1 call

Line-by-line resume feedback for this application, how to position your Role Readiness, and a clear plan for what to do next - with a OneRoadmap career coach.

Experience Senior · 6+ yrs (7+ years)

About the role

structured by ORI

Lead advanced analysis of petabyte-scale, multi-source datasets-including streaming, cloud, and legacy systems-to support reporting, predictive modeling, and Al-driven solutions. DESCRIPTION: Duties: Lead advanced analysis of petabyte-scale, multi-source datasets-including streaming, cloud, and legacy systems-to…

What you will do

  • Lead advanced analysis of petabyte-scale, multi-source datasets-including streaming, cloud, and legacy systems-to support reporting, predictive modeling, and Al-driven solutions.
  • Translate business objectives into scalable technical strategies and deliver end-to-end data solutions that maximize product value.
  • Design and implement AI-enabled data transformation pipelines using NLP and generative AI to automate metadata tagging, enhance data quality, and streamline documentation.
  • Oversee multiple data initiatives by managing priorities, milestones, KPIs, and stakeholder communication while mitigating risks and inefficiencies.
  • Partner cross-functionally with Product, Marketing, Operations, Technology, and Data teams to strengthen the data ecosystem and ensure foundational data needs are met.

What they are looking for

  • Bachelor's degree in Computer Engineering, Computer Science, Management Information Systems, Data & Analytics or related field of study plus 7 years of experience in the job offered or as Data Scientist, Business Intelligence Developer, Software Engineer, or related occupation.
  • Three (3) years of experience with developing production-grade data pipelines and ETL workflows using Python including libraries Pandas, NumPy, and SQLAlchemy and PySpark for distributed data processing of datasets
  • Three (3) years of experience with developing, optimizing, and maintaining SQL stored procedures with dynamic parameterization, error handling, and performance tuning for operational reporting systems in Microsoft SQL Server, Oracle database, Teradata and Snowflake
  • Three (3) years of experience with architecting and implementing big data solutions using Hadoop ecosystem components including HDFS, Hive, and MapReduce for processing multi-terabyte datasets
  • Three (3) years of experience with conducting PySpark optimization techniques including partitioning strategies, broadcast joins, and memory management for cluster computing environments
  • Three (3) years of experience with designing and deploying end-to-end data solutions on AWS cloud infrastructure, including S3 for data lake architecture with lifecycle policies and versioning and RDS and Aurora for relational database management with high availability configurations
  • Three (3) years of experience with Using Amazon Redshift for data warehousing, optimizing distribution keys and sort keys for performance and scalability
  • Three (3) years of experience with Conducting serverless SQL querying and analysis of petabyte-scale datasets ssing Amazon Athena
  • Three (3) years of experience with orchestrating and automating data workflows using AWS services including Glue, Lambda, EMR, and Step Functions
  • Three (3) years of experience with Administering and optimizing Snowflake for enterprise data warehousing, including virtual warehousing, time travel, zero-copy cloning, building data pipelines, and implementing streams and tasks for automated and real-time data processing
  • Three (3) years of experience with Designing and optimizing Teradata and Oracle databases, performing performance tuning, workload management, data modeling, PL/SQL development, partitioning, indexing, and query optimization to ensure efficient, high-volume data processing
  • Three (3) years of experience with Developing and maintaining Microsoft SQL Server, designing ETL workflows with SSIS, implementing indexing strategies, high availability, T-SQL analytics, and ensuring secure, scalable, and high-performance data solutions
PythonPandasNumPySQLAlchemyPySparkSQLPL/SQLT-SQLMicrosoft SQL ServerOracleTeradataSnowflakeHadoopHDFSHiveMapReduce
Full posting text

Lead advanced analysis of petabyte-scale, multi-source datasets-including streaming, cloud, and legacy systems-to support reporting, predictive modeling, and Al-driven solutions.

DESCRIPTION: Duties: Lead advanced analysis of petabyte-scale, multi-source datasets-including streaming, cloud, and legacy systems-to support reporting, predictive modeling, and Al-driven solutions. Translate business objectives into scalable technical strategies and deliver end-to-end data solutions that maximize product value. Design and implement AI-enabled data transformation pipelines using NLP and generative AI to automate metadata tagging, enhance data quality, and streamline documentation. Oversee multiple data initiatives by managing priorities, milestones, KPIs, and stakeholder communication while mitigating risks and inefficiencies. Partner cross-functionally with Product, Marketing, Operations, Technology, and Data teams to strengthen the data ecosystem and ensure foundational data needs are met. Apply domain expertise in Home Lending analytics, mentor junior team members, and promote best practices in data management and innovation. QUALIFICATIONS: Minimum education and experience required: Bachelor's degree in Computer Engineering, Computer Science, Management Information Systems, Data & Analytics or related field of study plus 7 years of experience in the job offered or as Data Scientist, Business Intelligence Developer, Software Engineer, or related occupation. Skills Required: This position requires three (3) years of experience with the following: developing production-grade data pipelines and ETL workflows using Python including libraries Pandas, NumPy, and SQLAlchemy and PySpark for distributed data processing of datasets; developing, optimizing, and maintaining SQL stored procedures with dynamic parameterization, error handling, and performance tuning for operational reporting systems in Microsoft SQL Server, Oracle database, Teradata and Snowflake architecting and implementing big data solutions using Hadoop ecosystem components including HDFS, Hive, and MapReduce for processing multi-terabyte datasets; conducting PySpark optimization techniques including partitioning strategies, broadcast joins, and memory management for cluster computing environments; designing and deploying end-to-end data solutions on AWS cloud infrastructure, including S3 for data lake architecture with lifecycle policies and versioning and RDS and Aurora for relational database management with high availability configurations; Using Amazon Redshift for data warehousing, optimizing distribution keys and sort keys for performance and scalability; Conducting serverless SQL querying and analysis of petabyte-scale datasets ssing Amazon Athena; orchestrating and automating data workflows using AWS services including Glue, Lambda, EMR, and Step Functions; Administering and optimizing Snowflake for enterprise data warehousing, including virtual warehousing, time travel, zero-copy cloning, building data pipelines, and implementing streams and tasks for automated and real-time data processing; Designing and optimizing Teradata and Oracle databases, performing performance tuning, workload management, data modeling, PL/SQL development, partitioning, indexing, and query optimization to ensure efficient, high-volume data processing; Developing and maintaining Microsoft SQL Server, designing ETL workflows with SSIS, implementing indexing strategies, high availability, T-SQL analytics, and ensuring secure, scalable, and high-performance data solutions; designing normalized and denormalized database schemas supporting OLTP and OLAP workloads; developing interactive dashboards and reports using Tableau including calculated fields, parameters, and LOD expressions; developing interactive dashboards and reports using Power BI including DAX, Power Query, and custom visuals; data quality management processes including profiling, cleansing, validation, and monitoring with measurable quality metrics including accuracy, completeness, and consistency using Tableau and Power BI; participating in sprint planning, daily standups, retrospectives, and delivering iterative data solutions in Agile Scrum environments; applying project management techniques including scope definition, resource allocation, risk management, and stakeholder communication for data initiatives. Job Location: 7255 Baymeadows Way, Jacksonville, FL 32256.

Data & AnalyticsPredictive Science
Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…

Opportunity details

Deadline
Closing in 16d · 17 Oct

As stated by the source. Anything not shown was not stated.

JPMorgan Chase

Financial Services

With a history tracing its roots to 1799 in New York City, JPMorganChase is one of the world's oldest, largest, and best-known financial institutions—carrying forth the innovative spirit of our heritage firms in global operations across 100 markets. We serve millions of customers and many of the world’s most prominent corporate, institutional, and government clients daily, managing assets and investments, offering business advice and strategies, and providing innovative banking solutions and services. Social Media Terms and Conditions: https://bit.ly/JPMCSocialTerms JPMorgan Chase & Co. is an

Company pageWebsite

More at JPMorgan Chase

JPMorgan ChaseProduct ManagementPreferredFunded

On-site · Wilmington, DE, United States / Columbus, OH, United States / Plano, TX, United States

Senior · 11h ago

Chase Auto Product Manager, Vice President
Agile Methodologies+8 more
View role
JPMorgan ChaseProduct ManagementPreferredFunded

On-site · Mumbai, Maharashtra, India

Senior · 11h ago

Technical Product Manager - ETL, Data Reporting, SQL
SQLTableauThoughtSpot+6 more
View role
JPMorgan ChaseProduct ManagementPreferredFunded

On-site · LONDON, United Kingdom

Senior · 11h ago

Product Manager - Global Banking - Deal Lifecycle Vice President
Product Management+7 more
View role
JPMorgan ChasePredictive SciencePreferredFunded

On-site · GLASGOW, LANARKSHIRE, United Kingdom

Senior · 11h ago

Lead Data Scientist -Platform AI Acceleration
View role
Apply