Data & Machine Learning Engineer
The Product Highway
Tailor my CV for this job, freeView job and applyYour CV rewritten for this role, from your real experience. Sign in with Google, nothing to install.
Got this interview? Our apps help you get the job.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
1,109 open learning roles across 230 companies are on ApplySarthi right now, most of them in Bengaluru (63), Hyderabad (25), Delhi NCR (14).
- Principal Machine Learning Developer, AI/ML PlatformAutodesk
- Director, Machine LearningJobgether
- Manager - Learning & Development (Frontline Capability Development)SWIGGY · bengaluru
- ETIC, Skills & Learning Operations- Graduate ProgramPwc
- Principal Machine Learning EngineerHubSpot
What learning roles keep asking for: Machine learning (49%), Python (36%), LLMs (23%), PyTorch (23%), Deep learning (16%), AWS (13%), Generative AI (13%) — counted across their open postings here.
AWS jobs · Airflow jobs · Machine learning jobs · PostgreSQL jobs
The Product Highway has 5 open roles listed here.
- Full Stack Developer
- Forward-Deployed Software Engineer (FDSE)bengaluru
- Frontend Engineerbengaluru
- AI Engineerbengaluru
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for learning roles keep coming back to Machine learning, Python, LLMs, PyTorch. Practise those questions before you sit with The Product Highway.
Questions you are likely to be asked
- Why do you want to join The Product Highway?
- What is your experience with Python? Tell me one thing you learned the hard way.
- When would you not use machine learning for a problem?
- Walk me through a model you built, from the data to how it was used.
- How did you know your model was actually good, and not just good on your test set?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Data & Machine Learning Engineer at The Product Highway interview free →**Role: Data & Machine Learning Engineer** AI-Native Product Engineering | WFO Bangalore | [CTC band] | 4+ years **Apply Link**: https://theproducthighway.keka.com/careers/jobdetails/412 **About TPH** The Product Highway is an AI-native product strategy and engineering firm that partners with businesses from the earliest spark of an idea all the way through to enterprise scale. We grew from zero to multi-million ARR within our first year, working with clients across India, APAC, Europe, and North America. The conventional product/software industry is coming to an end. The line between product development and business development is disappearing. The best talent won't spend a career maintaining one product. They'll operate as forward-deployed product managers and engineers, building businesses end to end in small, high-leverage teams, one after another. That's the vision we're building toward. AI got very good at the how: shipping code faster than ever. But almost nothing has changed in the what: deciding what to build, why, and in what order. That's where we live. We believe software is an approximation of the real world, and what matters isn't lines of code or sprint velocity. It's whether the solution actually maps to how a business works, how customers think, and how value gets created. From in-house AI project managers to near-universal adoption of tools like Cursor and Claude, we've rethought every conventional process from first principles. AI isn't a feature we offer. It's how we think, build, and deliver. The result: legacy systems rebuilt in under 6 months, enterprise platforms built in 3, mobile applications shipped in under 1. We only take on problems we find genuinely interesting and worthy of being solved, and we refuse to ship anything we wouldn't stand behind. Every hire gets us closer to that standard. **Who we're looking for** Regardless of role, every person at TPH shares these traits: **End-to-end owners.** You own outcomes, not tasks. If something you're responsible for falls through a crack, that's on you. Not because someone assigned it, but because you wouldn't have it any other way. **Clear, direct communicators**. Bad news doesn't get better with age. Context doesn't transfer through vague Slack messages. You surface problems early and communicate with precision. Uncompromising on quality. You have a quality bar that's yours, not your manager's. You won't ship something you wouldn't stand behind, even under deadline pressure. **AI-obsessed.** You see AI as how work gets done, not a nice-to-have. If you're still doing something manually that AI could handle, you feel that as friction, not normalcy. Structured thinkers. When faced with an ambiguous problem, you break it down, reason through the trade-offs, and arrive at a position. You don't wait for someone to tell you the answer. **Experienced enough to use AI systematically**. You have enough depth in your craft that AI makes you dangerous, not dependent. You direct it, evaluate its output, and know when it's wrong. **Why we're hiring** The AI systems we ship are only as good as the data underneath them: clean pipelines, well-engineered features, labelled datasets that hold up under scrutiny, and evaluation sets that actually mean something. That layer is now the bottleneck on most of what we build - recommender systems, moderation pipelines, retrieval-heavy products, and vision and vision-language workflows. We need an engineer who owns it end to end: someone who has moved data from raw to production-ready, worked shoulder to shoulder with annotation teams, and knows the difference between a notebook that runs once and a pipeline that survives real volume. **The work** You'll build the data and model layer across TPH's client portfolio: ingestion and transformation pipelines, feature engineering on unstructured text, annotation workflows that feed training and evaluation, recommender systems for e-commerce and content surfaces, and inference pipelines for vision and vision-language models. The projects change. The bar doesn't: production quality, measured behaviour, no vibes-based shipping. **Stack exposure:** Python, SQL, Postgres, pandas and Polars, pipeline orchestration (Airflow, Dagster, dbt), PyTorch, embeddings and similarity search (pgvector, FAISS), recommender and collaborative-filtering libraries, vision and vision-language models (Grounding DINO, CLIP-family VLMs), experiment tracking and eval harnesses, AWS. **What you'll achieve** 1. Own the data layer behind AI systems real businesses depend on, across multiple domains, not one product for years 2. Build genuine depth in the parts of ML engineering that matter in production: data quality, feature design, evaluation, reproducibility, inference cost and latency 3. Work directly with founders, annotation teams, and client stakeholders. Your judgment on data shapes what the model can do, not just how it gets trained 4. Learn how dataset design, labelling quality, and evaluation connect to business outcomes, because on every project the client is paying for outcomes **Key Responsibilities** 1. Build and maintain data pipelines using Python and SQL. 2. Transform raw text into categorical variables and structured features. 3. Work closely with annotators to understand labelling guidelines, edge cases and data-quality issues. 4. Convert annotated data into formats suitable for model training, inference and evaluation. 5. Implement model workflows using predefined architectures and training methods. 6. Support recommender systems, including collaborative-filtering approaches. 7. Build and optimize inference pipelines for vision and vision-language models such as Grounding DINO and VLMs. **Required Skills** 1. 4+ years of overall engineering experience with strong software fundamentals. ML engineering on top of weak engineering doesn't work. 2. Strong hands-on experience with Python, SQL and data engineering. 3. Experience processing unstructured text and engineering categorical features. 4. Experience working with annotated datasets and data-labelling teams. 5. Practical knowledge of recommender systems and collaborative filtering. 6. Understanding of computer-vision and vision-language models from an inference perspective. 7. Familiarity with data validation, model evaluation and reproducible ML workflows. 8. Strong analytical, problem-solving and communication skills. **Strong pluses** 1. Hands-on experience with PyTorch: training loops, fine-tuning, and debugging model behaviour 2. Working with embeddings and similarity search at scale (pgvector, FAISS, or a dedicated vector database) 3. Experience with multimodal datasets: paired image-text data, annotation tooling, and quality control at scale 4. Optimizing model inference: batching, quantization, GPU utilization, and latency budgets 5. Built internal tooling for data and ML work: annotation dashboards, dataset versioning, evaluation harnesses **AI-native expectations** 1. Uses Claude Code or Cursor as the primary development environment, shipping production code through AI-assisted workflows daily 2. Maintains CLAUDE.md and project context files so AI tools know your conventions, architecture, and constraints 3. Plan-first for multi-file changes: AI reads the codebase, you approve the approach, then execute 4. Uses AI to accelerate the data work itself: drafting transformation logic, generating test fixtures and edge cases, stress-testing pipelines, triaging model errors 5. Tests after every AI-generated change. AI writes fast; you keep it honest
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- Full Stack DeveloperThe Product Highway
- Forward-Deployed Software Engineer (FDSE)The Product Highway · bengaluru
- AI EngineerThe Product Highway · bengaluru
- Frontend EngineerThe Product Highway · bengaluru
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on wellfound · posted 2026-08-15. ApplySarthi collects openings and links to application pages; the role is advertised by The Product Highway, not by us.