FULLTIME
Data & AI Engineer
Webbtree
Not specified · onsite · Posted 1d ago
Your match
Sign in to see your match score, skill gaps & tailored resume.
Section · 01
About this role
Company Profile Webbtree makes software to help companies find and hire great people. We get recruiting and its role in building healthy workplaces. And while we take recruiting seriously, we don't take ourselves too seriously. At Webbtree, you'll find smart people who have fun, learn and innovate, and help others do the same. We brainstorm, we laugh, and, occasionally, we party (there's a lot to celebrate), but we also appreciate people's need for quiet time and focused work. We respect everyone, we hire the best, and make sure every experience is special. If that’s not enough to join, we’re backed by a multi-national Executive Search firm - WhiteCrow Research that has been working with Global Fortune 100 companies for more than 12 years. DATA & AI ENGINEER
Mumbai, India | Hybrid | Full-time | 2+ years About The Role We are looking for a hands-on Data & AI Engineer to work on the data, search, analytics and AI systems powering our recruitment platform. This is not a traditional data analyst/reporting role. You will work across large-scale search, data engineering, analytics, NLP and Generative AI, including OpenSearch, vector search, embeddings, reranking, LLMs and AI-driven retrieval/screening systems.
What You'll Work On
- Design, build and optimise large-scale search/retrieval pipelines using OpenSearch / Elasticsearch.
- Work with keyword, vector, semantic and hybrid search; embeddings, ranking, relevance and filtering.
- Design scalable MySQL data models, schemas, indexes and queries; optimise database performance.
- Build ingestion, transformation, indexing/re-indexing and data pipelines supporting AI, analytics and reporting.
- Build BI-ready datasets and Tableau dashboards; ensure data quality and consistency.
- Experiment with and evaluate embeddings, rerankers, LLMs and ML models using accuracy, relevance, latency and cost.
- Build LLM/GenAI workflows using LangChain and Langfuse; support tracing, observability and evaluation.
- Take AI/data experiments from POC → production with backend and product teams; optimise latency, throughput, reliability and cost.
- Work with AWS infrastructure including OpenSearch, S3, EC2 and RDS. MUST HAVE
- 2+ years of hands-on experience in OpenSearch/Elasticsearch, ML/AI/GenAI, Search/Information Retrieval or Data Engineering.
- Strong Python programming and extremely strong SQL skills.
- Production experience with MySQL: schema design, query optimisation, indexing and performance tuning.
- Strong data modelling/database design and experience with large-scale data pipelines/datasets.
- Practical ML/AI model evaluation experience and understanding of experimentation metrics.
- Experience with at least one of: LLMs, NLP, Generative AI, embedding models, reranking models or recommendation/search systems.
- Experience integrating external ML/AI models through APIs. GOOD TO HAVE
- Tableau/BI, data warehouses or analytics layers; Langfuse or similar observability/evaluation platforms.
- RAG/retrieval pipelines, vector databases, rerankers/cross-encoders, Hugging Face/open-source models.
- AWS (OpenSearch, S3, EC2, RDS), high-volume search, NLP/information retrieval/recommendation systems.
- Familiarity with MRR, NDCG, Recall@K and Precision@K; Docker and CI/CD.
What We're Looking For Someone extremely hands-on with data who thinks in SQL first, is deeply comfortable with MySQL and data modelling, and can move seamlessly between data engineering, analytics and AI experimentation. You should be comfortable going from data → experiment → evaluation → optimisation → production, while also building reporting layers for business users. WHY THIS ROLE IS INTERESTING Own the full data lifecycle: how data is stored, queried, used for analytics, powers search and powers AI. This is a rare opportunity where data engineering, analytics and AI engineering converge in one system, with direct influence over how data is structured, search works and AI systems behave in production. EDUCATION Bachelor’s or Master’s degree in Computer Science, Data Science, AI/ML, Engineering or a related field. Equivalent practical experience is also considered. Powered by Webbtree
Sourced from linkedin · view original
Let the agent run this one for you.
Tailored resume, auto-apply, and referral lookup — in under 2 minutes.
Section · 02