Skip to main content
Top expert badge
Recommended expert
Profile header background

Ayush Rai-Senior AI Engineer | RLHF & LLM Evaluation Specialist | Bhopal, India

Ayush Rai - Senior AI Engineer | RLHF & LLM Evaluation Specialist | Bhopal, India - profile avatar
Profile header overlay
Available
Bhopal, India

Check rate

Experience

Jan 2026 - Dec 2026
Bhopal, India
Remote

Adaptive RAG : Agentic Research Automation System

TDP

Position summary
Adaptive RAG : Agentic Research Automation System at TDP
Industries
Information Technology
Business areas
Information Technology
Product Development
Research and Development

Python, LangGraph, FastAPI, Qdrant, MongoDB, Streamlit

  • Built an agentic RAG system that dynamically routes queries across indexed document retrieval, general LLM reasoning, and real-time web search via a modular LangGraph orchestration pipeline.
  • Implemented query analysis, route classification, relevance grading, query rewriting, answer generation, and Tavily search fallback integrated with Qdrant vector search, MongoDB session memory, and OpenAI-compatible LLMs.
Apr 2025 - Present

AI Evaluation Contributor & Reviewer

Scale AI, Turing, Snorkel AI, Airdawgs Labs : Partner Programs

Position summary
AI Evaluation Contributor & Reviewer at Scale AI, Turing, Snorkel AI, Airdawgs Labs : Partner Programs
Industries
Information Technology
Business areas
Information Technology
Quality Assurance
Research and Development
  • Contributed structured preference ranking, pairwise comparison, and model evaluation data across code, reasoning, and STEM domains for Scale AI; maintained high inter-annotator agreement and evaluation consistency.
  • Evaluated LLM-generated code and reasoning outputs on the Turing / OpenAI Feather Platform; contributed human-feedback data for prompt evaluation, model alignment, and AI quality assessment across coding tasks.
  • Performed data annotation, preference ranking, and quality review for Snorkel AI; ensured guideline compliance and dataset integrity across multiple evaluation task batches before submission.
  • Served as Senior Contributor & Reviewer at Airdawgs Labs (Terminus, Marlin, OpenDrain): audited annotation quality, calibrated contributors, identified edge cases, and maintained RLHF evaluation consistency across projects.
Mar 2024 - Apr 2025
Remote

Generative AI Engineer

Outlier AI

Position summary
Generative AI Engineer at Outlier AI
Industries
Information Technology
Business areas
Information Technology
Quality Assurance
Research and Development
Explore AI Engineer
  • Executed RLHF and SFT workflows for frontier LLM training: evaluated model responses for correctness, reasoning quality, and instruction adherence; ranked competing outputs using structured quality criteria.
  • Delivered advanced prompt engineering and code-generation datasets across Python and JavaScript, improving contextual accuracy by 25% and reducing iteration cycles by 10%.
  • Conducted systematic error analysis identifying hallucinations, factual inaccuracies, and logical inconsistencies and authored detailed rationales to calibrate preference selections and maintain dataset quality.
Jan 2024 - Dec 2025
Bhopal, India
Remote

LLM Code Evaluation & Preference Annotation Corpus

WootzApp

Position summary
LLM Code Evaluation & Preference Annotation Corpus at WootzApp
Industries
Information Technology
Business areas
Information Technology
Quality Assurance
Research and Development

Python, unittest, Selenium, PyTorch, OpenAI API, LangChain

  • Architected a 1,400+ file LLM code evaluation repository with 193 automated unittest-backed eval tasks and 65 Colosseum A/B multi-turn preference comparison tasks spanning Python, ML, web, finance.
  • Engineered end-to-end evaluation infrastructure TASKS.json manifest generator, batch test runner, structure validator, and Colosseum→JSONL exporter indexing 260+ tasks.
  • Authored 14 technical guides covering architecture, onboarding, API reference, and subsystem documentation; curated cross-domain reference solutions across finance (yield curves, arbitrage), ML/DL (Keras, PyTorch, XGBoost), and algorithms.
Sep 2022 - Mar 2024
Bhopal, India

Senior AI Engineer

FoCDoT Technologies Pvt. Ltd.

Position summary
Senior AI Engineer at FoCDoT Technologies Pvt. Ltd.
Industries
Information Technology
Business areas
Information Technology
Operations
Product Development
Research and Development
Strategy
Explore AI Engineer
  • Designed and shipped production AI systems spanning LLM evaluation pipelines, RAG architectures, and AI-assisted decision platforms for startup and enterprise clients, contributing to 40L+ in cumulative client revenue.
  • Built and operated RLHF and human-feedback workflows for frontier AI evaluation programs, delivering structured datasets, grading rubrics, error taxonomies, and alignment feedback loops that improved AI output reliability by 30%.
  • Led AI readiness audits and maturity assessments covering data constraints, model risks, and deployment feasibility for 10+ client engagements; worked directly with founders and CXOs to align AI strategy with business outcomes.

Industry experience

See where this freelancer has spent most of their professional time.

Experienced in Information Technology.

Information Technology
Profile match chart

Business area experience

See which departments and functions this freelancer has contributed to most.

Experienced in Information Technology, Research and Development, Quality Assurance, Product Development, Operations, and Strategy.

Information Technology
Research and Development
Quality Assurance
Product Development
Operations
Strategy
Profile match chart

Summary

AI Engineer with 4+ years building production LLM systems, RAG pipelines, and human-feedback evaluation infrastructure in fast-moving startup environments. Deep hands-on experience in RLHF workflows, preference ranking, AI evaluation, and model quality assessment through partner programs with Scale AI, Turing, Snorkel AI, and OpenAI. Proven track record shipping reliable, latency-optimised AI products and operating structured evaluation pipelines at scale.

Skills

  • Rlhf & Ai Evaluation: Rlhf, Sft, Preference Ranking, Pairwise Comparison, Human Feedback Systems, Model Evaluation, Agent Evaluation, Benchmarking, Error Analysis, Data Annotation

  • Backend Engineering: Python, Fastapi, Node.Js, Nestjs, Rest Apis, Oauth 2.0, Docker

  • Data & Infrastructure: Postgresql, Sqlite, Mongodb, Chromadb, Qdrant, Vector Dbs, Aws S3

  • Llm & Ai Frameworks: Langgraph, Langchain, Crewai, Openai Api, Anthropic Api, Ollama, Sentence-Transformers

  • Frontend & Interfaces: React, Next.Js, Typescript, Tailwind Css, Streamlit, Radix Ui

  • Engineering & Delivery: System Design, Prompt Engineering, Ci/Cd, Git, Mcp, Cloud Fundamentals

Languages

Hindi
Native
English
Advanced

Education

Sep 2020 - Jul 2024

Lakshmi Narain College of Technology and Science

B.Tech · Computer Science and Engineering · Bhopal, India · 8.47 / 10.0

Statistics

Experience

Total positions 5
Experience in Information Technology 4 y
Avg length 1 y 3 m
Longest experience 1 y 11 m

Global experience

Countries worked in 1 (India)
Primary country India

Expertise

Recent roles Adaptive RAG : Agentic Research Automation System, AI Evaluation Contributor & Reviewer, Generative AI Engineer
Main industries Information Technology
Main business areas Information Technology, Research and Development, Quality Assurance

Qualifications

Highest degree Bachelor

Profile

Member since
Last update
Need a freelancer? Find your match in seconds.
Try FRATCH GPT
More actions

Frequently asked questions

Have questions? Find more information here.

Ayush is based in Bhopal, India and can operate in on-site, hybrid, and remote work models.

Ayush speaks the following languages: Hindi (Native), English (Advanced).

Ayush has at least 4 years of experience. During this time, Ayush has worked in at least 5 different roles and for 5 different companies. The average length of individual experience is 1 year and 9 months. Note that Ayush may not have shared all experience and actually has more experience.

Based on recent experience, Ayush would be well-suited for roles such as: Adaptive RAG : Agentic Research Automation System, AI Evaluation Contributor & Reviewer, Generative AI Engineer.

Ayush's most recent position is Adaptive RAG : Agentic Research Automation System at TDP.

In recent years, Ayush has worked for TDP, Scale AI, Turing, Snorkel AI, Airdawgs Labs : Partner Programs, Outlier AI, WootzApp, and FoCDoT Technologies Pvt. Ltd..

Ayush is most experienced in industries like Information Technology.

Ayush is most experienced in business areas like Information Technology, Research and Development, and Quality Assurance. Ayush also has some experience in Product Development, Operations, and Strategy.

Ayush holds a Bachelor in Computer Science and Engineering from Lakshmi Narain College of Technology and Science.

Ayush is immediately available full-time for suitable projects.

Daily rate distribution

0 1 2 3 4
<€320 €320-​480 €480-​640 €640-​800 €960+

The rates shown represent the typical market range for freelancers in this position based on recent contracts on our platform.

Average rates for similar positions

Rates are based on recent contracts and do not include FRATCH margin.

600
450
300
150
Rate comparison chart
Daily rate avg. 407 €

The average daily rate is the mean of all daily rates from recent contracts of comparable freelancers on our platform.

600
450
300
150
Rate comparison chart
Median rate 372 €

The median daily rate is the middle value of all daily rates — half of comparable freelancers charge less, half charge more. Unlike the average, it is barely affected by outliers.

Calculated based on our freelancers’ daily rates as of 28 Aug 2026. Actual rates may vary depending on seniority level, experience, skill specialization, project complexity, and engagement length.