Data Engineer / Data Scientist
Jobgether
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
2,051 open scientist roles across 293 companies are on ApplySarthi right now, most of them in Bengaluru (147), Hyderabad (71), Delhi NCR (31).
- Senior Scientist in Process Chemistry & CatalysisRoche
- Senior Principal ScientistPfizer
- IN_Manager_AI Data Scientist Engineer_GCC_Advisory_BangalorePwc · bengaluru
- Staff Data ScientistGeneralmotors
- Lead Applied Scientist, Search - NLP/GenAIThomsonreuters
What scientist roles keep asking for: Python (52%), Machine learning (40%), SQL (25%), C++ (18%), Java (17%), R (15%), Deep learning (15%), LLMs (14%) — counted across their open postings here.
Data Engineer jobs in India · Remote Data Engineer jobs · Azure jobs · CI/CD jobs · Databricks jobs · Kafka jobs
Jobgether has 3,899 open roles listed here.
- (Technical Targeter- Virtual Operations) Cyber Technical Analyst Principal (TS/SCI with Poly Required)
- Advogado(a) | Bancário
- Agentic Workforce Adoption Manager
- AI Engineer
- AI Marketing Project Manager
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for scientist roles keep coming back to Python, Machine learning, SQL, C++. Practise those questions before you sit with Jobgether.
Questions you are likely to be asked
- Why do you want to join Jobgether?
- What is your experience with Machine learning? Tell me one thing you learned the hard way.
- Walk me through a model you built, from the data to how it was used.
- How did you know your model was actually good, and not just good on your test set?
- Tell me about a time the data was messy or wrong. What did you do?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Data Engineer / Data Scientist at Jobgether interview free →Accountabilities:: You will develop, optimize, and support modern data processing and machine learning solutions across cloud-based environments. The role requires strong technical ownership, practical problem-solving, and collaboration across engineering, data science, and business teams. Develop scalable data processing solutions using Python, PySpark, and Azure Databricks. Build, maintain, and optimize batch and real-time streaming data pipelines. Develop Spark DataFrame-based transformations and data processing workflows. Debug, troubleshoot, and optimize Spark applications and Databricks jobs. Implement Delta Lake solutions to improve data reliability, versioning, and query performance. Develop APIs using Python or Scala for data and machine learning applications. Support machine learning initiatives, MLOps workflows, and model deployment activities. Work with Azure services for data ingestion, storage, security, integration, and processing. Configure and manage Databricks job clusters, compute environments, and notebook workflows. Build and execute DataFrame-based data validation and quality checks. Develop pipelines using Event Hubs, Kafka, IoT sources, or other real-time data technologies. Support data quality monitoring and production troubleshooting. Implement secure integrations between Azure services using managed identities and secrets. Contribute to CI/CD practices for data engineering and machine learning workloads. Collaborate with technical and business stakeholders while independently managing assigned deliverables. Apply performance tuning techniques to Spark applications and Databricks workloads. Requirements The ideal candidate brings 5–8 years of relevant experience across data engineering, data science, machine learning, or cloud analytics, with strong hands-on capabilities in Python, PySpark, Azure, and Databricks. You should be comfortable developing production-ready data solutions, troubleshooting distributed processing workloads, and contributing to machine learning and MLOps initiatives. Bachelor’s or Master’s degree in Computer Science, Data Science, Engineering, Information Technology, or a related discipline. 5–8 years of relevant professional experience in data engineering, data science, machine learning, or cloud analytics. Strong hands-on expertise in Python and PySpark. Good knowledge of Microsoft Azure and Azure Databricks. Hands-on experience with MLOps practices and tools. Practical experience supporting machine learning projects. Basic understanding of machine learning model deployment. Strong experience developing and debugging Spark-based applications. Hands-on experience with Databricks notebook development. Strong knowledge of Spark DataFrames using PySpark or Scala. Experience optimizing Spark jobs and Databricks workloads. Experience developing APIs using Python or Scala. Working knowledge of Azure Event Hubs, Storage Accounts, Key Vault, Service Bus, Azure Functions, and Azure Data Lake Storage. Understanding of Databricks job clusters and compute configurations. Experience implementing cloud-based data solutions on Azure. Knowledge of real-time streaming technologies such as Kafka. Experience developing batch and streaming pipelines using Event Hubs, Kafka, or IoT data sources. Hands-on experience implementing Delta Lake solutions. Working knowledge of GitHub or similar version-control platforms. Exposure to MLflow or comparable tools for experiment tracking and model lifecycle management is beneficial. Experience with CI/CD for data and machine learning workloads is a plus. Knowledge of data quality validation, monitoring, and production support is advantageous. Strong analytical and problem-solving abilities. Ability to work independently while collaborating effectively with cross-functional project teams. Benefits Full-time position. Remote work arrangement. Immediate requirement with an opportunity to join a technology-focused data and AI environment. Opportunity to work with modern cloud technologies including Microsoft Azure and Azure Databricks. Hands-on exposure to Python, PySpark, Spark, Delta Lake, streaming, and MLOps. Opportunity to contribute to machine learning projects and model deployment initiatives. Exposure to real-time data technologies such as Kafka, Event Hubs, and IoT data sources. Opportunities to work across data engineering, machine learning, and cloud analytics. Collaboration with technical and business teams on impactful data initiatives. Scope for continued development in cloud, data engineering, and machine learning technologies.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- AI Infrastructure System EngineerJobgether
- B2B Paid Advertising SpecialistJobgether
- Business Development Rep (AI Cloud)Jobgether
- Business Development Rep (AI Cloud)Jobgether
- Business Development Rep (AI Cloud)Jobgether
- Business Development Rep (AI Cloud)Jobgether
- Business Development Representative, PressableJobgether
- Business Development Representative, PressableJobgether
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on lever · posted 2026-10-02. ApplySarthi collects openings and links to application pages; the role is advertised by Jobgether, not by us.