Principal Data Engineer, LLM/AI Platforms
Jobgether
Make my CV for this job, freeView job and applyYour CV, rewritten for this role using only your real experience. Sign in with Google and upload your CV. Nothing to install.
Skills named in this job
Read from the description itself, not inferred.
This role on the market
383 open platforms roles across 86 companies are on ApplySarthi right now, most of them in Hyderabad (11), Bengaluru (6), Pune (1).
- Data Engineer-Data Platforms-SnowFlakeIBM
- Product Director - Employee Platforms Data Product ManagementJPMorgan
- Platforms and Infrastructure EngineerEnactintelligence
- Full-stack Engineer 5 (AI Platform & Knowledge Library) (Enterprise Platforms Technology)Capitalone
- Data Advisory & Insights Specialist - Data PlatformsRoche
What platforms roles keep asking for: AWS (12%) — counted across their open postings here.
Data Engineer jobs in the United States · Remote Data Engineer jobs · AWS jobs · Airflow jobs · BigQuery jobs · Data modelling jobs
Jobgether has 3,942 open roles listed here.
- AI Researcher — Distillation
- AI Researcher — Distillation
- Art Director
- Applied ML Engineer
- AI Science Writer, Nebius Academy (Contract)
Counted across 14 company job boards, updated as roles open and close.
Preparing for this interview
Interviews for platforms roles keep coming back to AWS. Practise those questions before you sit with Jobgether.
Questions you are likely to be asked
- Why do you want to join Jobgether?
- What is your experience with LLMs? Tell me one thing you learned the hard way.
- How would you explain your model's result to someone who is not technical?
- What would you check first if a model's accuracy dropped after going live?
- When would you not use machine learning for a problem?
Prep Sarthi gives you a free mock interview: an AI interviewer asks you questions like these out loud, from your own CV and this job, and shows your score and your weakest answer.
Practise the Principal Data Engineer, LLM/AI Platforms at Jobgether interview free →Accountabilities:: Architect, implement, and optimize data platforms and pipelines designed for LLMs, RAG, and advanced AI agentic systems at Exabyte scale. Drive the adoption and deployment of agentic workflows and agent-harnessing techniques to support autonomous, data-driven capabilities. Design highly scalable, fault-tolerant, secure, and cost-effective data solutions that enable rapid iteration without compromising engineering quality. Develop production-ready code with strong attention to performance, maintainability, testing, and operational reliability. Provide technical leadership in data modeling, normalization, semantic cataloging, and data architecture for AI and machine learning workloads. Establish MLOps and DataOps best practices for LLM platforms, including monitoring, observability, automated recovery, and service reliability. Own the end-to-end lifecycle of critical data services, including development, testing, deployment, monitoring, and continuous optimization. Collaborate with data scientists, product managers, and engineering teams to transform research prototypes into robust, production-ready services. Lead technical workshops, design reviews, and knowledge-sharing initiatives while mentoring engineers and strengthening organizational expertise in AI platform technologies. Champion DevSecOps practices and engineering standards across large-scale distributed data environments. Identify opportunities to improve platform performance, reliability, scalability, and developer productivity through new technologies and engineering practices. Requirements Master’s degree or PhD in Computer Science, Data Engineering, or a related STEM discipline, or equivalent practical experience. 10+ years of progressive experience in Data Engineering or Platform Engineering, including at least 3 years architecting and building AI/ML or Data Science platforms at massive scale. 3+ years of experience in a Principal or Staff-level engineering capacity, with demonstrated technical leadership and mentorship experience. Hands-on expertise with LLM engineering, including fine-tuning, prompt engineering, deployment, RAG, and agentic workflow development. Proven experience designing and delivering large-scale distributed systems, including sharding, partitioning, concurrency, and fault-tolerant architectures. Expert-level proficiency in Python or JVM-based technologies, with a strong ability to write clean, performant, maintainable, and well-tested production code. Deep experience with distributed data processing frameworks such as Spark, Dask, or Flink. Strong knowledge of cloud platforms such as AWS, GCP, or OCI and their associated data services. Expertise with containerization and orchestration technologies including Docker and Kubernetes. Experience with messaging and streaming technologies such as Kafka or Pulsar. Familiarity with data warehousing and orchestration platforms such as Snowflake, BigQuery, Airflow, and Kubeflow. Experience with MLOps technologies such as MLflow, SageMaker, or Vertex AI. Familiarity with agentic AI frameworks such as LangChain or LlamaIndex. Strong understanding of engineering practices including peer code reviews, resilient architecture, comprehensive testing, and secure development methodologies. Demonstrated ability to use AI technologies to improve decision-making, automate workflows, increase efficiency, and support measurable business outcomes. Strong communication and collaboration skills, with the ability to influence technical direction and mentor engineers across teams. Direct experience deploying and managing LLMs in production is a plus. Experience in cybersecurity, intelligence, or highly regulated industries is a plus. Contributions to open-source data or AI/ML projects are a plus. Benefits CAD $210,000–$320,000 annual base salary for Canadian-based employment, plus variable/incentive compensation, equity, and benefits. Flexible remote work opportunities. Comprehensive health and wellness programs supporting physical and mental wellbeing. Competitive vacation and holiday programs to support time away and recharge. Paid parental and adoption leave. Professional development and continuous learning opportunities at all career levels. Employee networks, geographic communities, and volunteer opportunities to build professional connections. Opportunities to work on large-scale AI, data engineering, and cybersecurity technologies. Equity opportunities as part of the overall compensation package. Retirement and financial benefits as applicable. Inclusive workplace practices and support for employees with disabilities. Canadian employment requires legal entitlement to work in Canada and may include applicable background checks.
Match this job to your CV
ApplySarthi scores your CV against this role, shows the skills you are missing, and writes a tailored version for the application.
Check my match →Similar open roles
- .Net Software DeveloperJobgether
- Account DirectorJobgether
- Account Director, Renewals & GrowthJobgether
- Advisor, BMO SmartFolio WFHJobgether
- Agentic AI DeveloperJobgether
- AI Graphic Designer + Video EditorJobgether
- AI/ML Data ScientistJobgether
- Analista de Automação e IA com N8NJobgether
Need answers during your interview? Try Live Sarthi.
Live Sarthi, an Interview Sarthi app, shows answer suggestions during the call.
- Hidden from supported screen sharingThe overlay stays out of supported Windows screen captures.
- Answers start in about 1.5 secondsResponse time varies with your connection and model.
- From your own CVYour projects and your experience, not a generic script.
- 30 minutes freeThen ₹99 for a 2-day pass with unlimited calls — you pay for the days you are interviewing, not a subscription.
A Windows app, from the same team as ApplySarthi.
Listed on lever · posted 2026-09-26. ApplySarthi collects openings and links to application pages; the role is advertised by Jobgether, not by us.