Base Career helps you apply smarter for this job.
Key skills for this role
Build end-to-end data pipelines: raw trace ingestion → dedup → format conversion → quality gating → training-ready datasets
Process large-scale JSONL data on AWS S3 (tens of thousands of traces per batch)
Convert between chat-completion formats (e.g., OpenAI → Llama 3.1 tool-calling format)
Implement smart deduplication and sampling to balance training distribution
Design identity-aware train/test splits that measure true generalization
Build data validation gates to detect schema drift and format anomalies
Create a continuous pipeline that auto-processes new production traces for retraining
Skip the repetitive application forms
Install the Base Career Chrome Extension and autofill job applications across major job boards with your profile.
Trusted by over 500,000 job seekers on Base Career
More from this employer
Hyderabad, IND
DATAECONOMY is seeking an AI/ML MLOps Engineer to fine-tune, deploy, evaluate, and monitor self-hosted large language models across the full machine learning lifecycle. The role requires 5–8 years of experience, strong P
Hyderabad, IND
DATAECONOMY is seeking an experienced Java Full Stack Developer to build and maintain enterprise applications across backend, frontend, cloud, and deployment layers. The role requires strong Java, Spring Boot, Spring Sec
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Hyderabad, IND
Cloud-first data and AI consultancy helping enterprises modernize data platforms, build AI solutions, and improve business decisions.
Visit company websiteJobs and hiring trendsFull-time
Senior · 6+ years experience
Hybrid
Apply faster on company sites with our extension.