Data Engineer III - Databricks, Pyspark, Python, AWS
JPMorgan
Make my CV for this job, freeView job and applyYour CV, rewritten for this role using only your real experience. Sign in with Google and upload your CV. Nothing to install.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
120 open databricks roles across 26 companies are on ApplySarthi right now, most of them in Bengaluru (6), Hyderabad (3), Chennai (3).
- Data Architect mit Expertise in Databricks (m/w/d) | Join our Talent Pool - SwitzerlandCallista Group AG
- Data Engineer-Data Platforms-DatabricksIBM
- Azure Databricks Platform EngineerFiserv
- Senior Manager Data Engineer (Databricks, Pyspark, Snowflake)Capitalone
- Engenheiro de Dados GCP Sênior (GCP / Databricks)Jobgether
What databricks roles keep asking for: Databricks (38%), Agile (22%), Azure (20%), SQL (20%), AWS (19%), Spark (18%), Python (18%), CI/CD (16%) — counted across their open postings here.
Data Engineer jobs in Bengaluru · Data Engineer jobs in India · Remote Data Engineer jobs · AWS jobs · Data modelling jobs · Data warehousing jobs · Databricks jobs
JPMorgan has 7,498 open roles listed here.
- Director of Software Engineering - Data & Wealth Managementbengaluru
- Lead Software Engineer - Javabengaluru
- Lead Software Engineer- Javabengaluru
- Payment Experiences and Services - Senior Product Associate
- Senior Associate, Relationship Manager
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for databricks roles keep coming back to Databricks, Agile, Azure, SQL. Practise those questions before you sit with JPMorgan.
Questions you are likely to be asked
- Why do you want to join JPMorgan?
- What is your experience with AWS? Tell me one thing you learned the hard way.
- How would you explain your model's result to someone who is not technical?
- What would you check first if a model's accuracy dropped after going live?
- When would you not use machine learning for a problem?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Data Engineer III - Databricks, Pyspark, Python, AWS at JPMorgan interview free →Be part of a dynamic team where your distinctive skills will contribute to a winning culture and team. As a Data Engineer III - Databricks, Pyspark, Python, AWS at JPMorgan Chase within the Commercial & Investment Bank, you'll serve as a seasoned member of an agile team to design and deliver trusted data collection, storage, access, and analytics solutions in a secure, stable, and scalable way. You are responsible for developing, testing, and maintaining critical data pipelines and architectures across multiple technical areas within various business functions in support of the firm’s business objectives. Job responsibilities Design, develop, and maintain big data pipelines ( batch and streaming ) using PySpark/Spark and Databricks . Lead/own data modelling and solution design for data products, including defining target-state architecture, data flows, and transformation patterns. Build scalable ingestion and transformation workflows for high-volume datasets , ensuring reliability , quality , and performance . Develop and optimize complex SQL transformations, reconciliation queries, and analytical datasets; perform query tuning for large-scale workloads. Apply strong data warehousing concepts ( dimensional modeling , SCDs , partitioning strategies , etc.) to build well-structured, analytics-ready data layers and Leverage common AWS services , with strong emphasis on S3 and AWS data processing capabilities , to support scalable storage and processing. Perform advanced debugging and troubleshooting across distributed Spark workloads ( data skew , shuffle tuning , memory/compute optimization ). Implement engineering best practices: modular design , efficient coding , code reviews , and CI-friendly development approaches. Use GitHub/Bitbucket and standard version control workflows to manage codebase, peer reviews, and releases. Partner with cross-functional stakeholders to convert requirements into robust big data solutions . Uses enterprise-authorized AI capabilities within the work environment to accelerate data pipeline/design analysis and documentation, validating outputs and handling data according to sensitivity and security requirements. Applies reuse-first, AI-assisted practices to strengthen SDLC-quality routines for data pipelines (e.g., test generation and control validation), ensuring traceability/auditability and alignment to resiliency and security expectations. Required qualifications, capabilities, and skills Formal training or certification on data engineering concepts and 3+ years applied experience Experience in data engineering / big data engineering , with strong hands-on delivery and Expert-level SQL skills (must be extremely strong): complex joins, window functions , CTEs , optimization, and analytical problem solving at scale. Strong hands-on coding experience with Python in production environments; demonstrated ability to write efficient , maintainable code. Deep expertise in Apache Spark (in depth) and PySpark , including performance tuning and distributed processing fundamentals . Strong experience with Databricks for large-scale data processing and pipeline development. Strong understanding of data warehousing concepts and best practices; proven capability in data modelling and solution design for scalable, maintainable data platforms/products. Experience implementing both batch and streaming data processing solutions. Familiarity with AWS S3 and common AWS services used in data platforms and processing Excellent debugging , troubleshooting , problem-solving skills and experience with GitHub , Bitbucket , and version control best practices. Demonstrated experience using enterprise-authorized AI capabilities within the work environment to support data engineering workflows with strong validation habits and awareness of data sensitivity. Ability to review and validate AI-assisted outputs (e.g., query suggestions, test ideas, or model change summaries) before use, escalating when uncertain and following data handling requirements. Preferred qualifications, capabilities, and skills Good to have: infrastructure provisioning in AWS using Infrastructure as Code (IaC) (e.g., Terraform , AWS CloudFormation ).
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Technology Support III - Software Asset Optimization Owner - ESMJPMorgan · hyderabad
- External Reporting - AnalystJPMorgan · bengaluru
- Securities Services - Data & AI Product Manager - Vice PresidentJPMorgan · bengaluru
- Markets- Commodities Systematic Trading-Associate-MumbaiJPMorgan · mumbai
- Lead Software Engineer - Java Full StackJPMorgan · hyderabad
- Senior Manager of Infrastructure Engineering -Cisco Call Manager, Telephony, VOIP (Team Management Experience)JPMorgan · bengaluru
- Associate - Product ControllerJPMorgan · mumbai
- Senior Thunderhead/SmartDX DeveloperJPMorgan · bengaluru
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on oraclehcm · posted 2026-09-23. ApplySarthi collects openings and links to application pages; the role is advertised by JPMorgan, not by us.