Data Engineer
Century Health
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
7,322 open data roles across 585 companies are on ApplySarthi right now, most of them in Bengaluru (433), Hyderabad (325), Mumbai (157).
- Data Scientist Internship (November, 6 months)shifttechnology
- Member of Technical Staff, Data AIHandshake
- Finance Technology Data Solutions EngineerJobgether
- Data Center Chief Engineer, DCEOAmazon
What data roles keep asking for: AWS (25%), SQL (23%), Python (23%), Machine learning (12%) — counted across their open postings here.
Data Engineer jobs in India · Remote Data Engineer jobs · AWS jobs · Airflow jobs · CI/CD jobs · Data warehousing jobs
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for data roles keep coming back to AWS, SQL, Python, Machine learning. Practise those questions before you sit with Century Health.
Questions you are likely to be asked
- Why do you want to join Century Health?
- What is your experience with ETL? Tell me one thing you learned the hard way.
- Walk me through a model you built, from the data to how it was used.
- How did you know your model was actually good, and not just good on your test set?
- Tell me about a time the data was messy or wrong. What did you do?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Data Engineer at Century Health interview free →**Century Health** is a clinical data intelligence company turning messy real-world clinical data into structured, analysis-ready insights for researchers and pharma teams. --- ## The Role We're looking for a ** Data Engineer (3–5 years)** to be a core builder on our data platform — designing ETL pipelines, profiling raw clinical datasets, and ensuring data flowing through our systems is clean and reliable. --- ## What You'll Do - Design, build, and maintain **ETL/ELT pipelines** ingesting data from CSVs, Parquet, XLSX, APIs, and databases - **Optimize pipeline performance** — tune queries, manage compute, reduce latency and cost - Profile raw datasets to identify quality issues: missing values, duplicates, schema drift, outliers - Build and maintain **dbt models** for clean, documented, analysis-ready data layers - Orchestrate workflows using **Apache Airflow on AWS MWAA** - Process large-scale data using **PySpark on EMR** - Collaborate with ML engineers and GTM teams on downstream use cases - Document pipelines, models, assumptions, and known issues clearly --- ## What We're Looking For **Must-Have** - 3–5 years of professional data engineering experience - Strong **SQL** — window functions, CTEs, query optimization - Solid **Python** — clean, modular, production-grade code - Hands-on **PySpark** for large-scale processing - Experience with **dbt**, **Snowflake**, and cloud-based ETL/ELT - Familiarity with **AWS**: S3, MWAA, ECS Fargate, EMR, RDS, Bedrock - Strong data intuition and ability to work independently in ambiguous environments - Experience with test suites — pytest, Great Expectations, dbt tests - Active use of **AI coding tools** (Cursor, Claude Code, etc.) - Immediate Joiners only (1 month notice period). We can be flexible for exceptional candidates. **Nice to Have** - Healthcare data experience (HIPAA, FHIR, HL7, OMOP) - Data science background (statistical modeling, ML pipelines, feature engineering) - DevOps exposure (CI/CD, Docker, Terraform/CDK) --- ## Tech Stack - **Cloud:** AWS (MWAA, ECS Fargate, EMR, RDS, S3, Bedrock) - **Data Warehouse:** Snowflake - **Transformation:** dbt - **Processing:** PySpark, Python - **Orchestration:** Apache Airflow (MWAA) - **AI Tools:** Cursor, Claude Code --- ## Hiring Process 1. Application Review 2. Take-Home Assignment (~3 hours, real clinical data) 3. Technical Interview 4. Managerial Interview --- ## Why Century Health - Hard data problems with real clinical impact - Small, high-ownership team — your work ships and matters - Modern cloud-native stack with architectural influence - AI-first engineering culture
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Data EngineerAlan
- Data EngineerBookboost
- Data EngineerOpenai
- Data EngineerElevenlabs
- Data EngineerGalaxy
- Data EngineerOmnicom Media · pune
- Data EngineerProdigal · bengaluru
- Data EngineerProlific
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on wellfound · posted 2026-09-07. ApplySarthi collects openings and links to application pages; the role is advertised by Century Health, not by us.