0 viewsTalent
Sandeep D. — Senior AI Alignment Specialist from India

Sandeep D.

Senior AI Alignment Specialist

India 3-6 years
Open to offersNew to Platform
Languages
EnglishHindiUrduBhojpuri
Video Introduction
No video introduction yet
The candidate has not added a video.
Contact information and social networks are private. Connect to unlock.
Hidden

About

Sandeep D. is a seasoned AI Alignment & Multilingual Model Training Specialist based in Varanasi, India, with extensive hands-on experience in Reinforcement Learning from Human Feedback (RLHF) and Supervised Fine-Tuning (SFT). With over 16 years in linguistic and editorial realms, Sandeep artfully employs cultural-linguistic reasoning to enhance semantic alignment and instruction adherence. His expertise in hallucination & bias detection and red teaming ensures robust safety evaluations and cultural context precision across Hindi-English datasets. At Outlier.ai, he crafted bilingual prompts for semantic and contextual accuracy in multilingual AI systems, excelling in structured preference ranking and factual evaluations. His tenure includes pivotal roles at CrowdGen and Alignerr, focusing on AI content quality and multimodal evaluation. Sandeep's academic pursuits include a Bachelor in Tourism Studies and diplomas in Computer Applications and Urdu Language, complementing his professional mastery in multilingual model alignment.

Experience

  • AI Content Quality Reviewer

    CrowdGen · 2023 — 2026
    Assessed digital outputs for accuracy in regional representation, cultural sensitivity, and compliance with established policies. Utilized structured scoring frameworks to aid in the refinement of datasets.
  • AI Alignment Program Participant

    Mercor · 2025 — Present
    Chosen for the Prompt Academy Fellowship, concentrating on advanced prompt creation and structured reasoning evaluations. Successfully completed the Advanced AI Red Teaming Trial in both English and Hindi, identifying patterns related to jailbreaks and instruction bypassing.
  • AI Evaluator & Reviewer

    Outlier.ai · 2024 — Present
    Performed over 500 multilingual AI training tasks within 15 alignment programs, including SFT and RLHF. Promoted to Reviewer shortly after beginning a speech-alignment project for ensuring consistent evaluation and precise analysis. Created bilingual prompts aimed at enhancing semantic alignment and cultural context accuracy in instruction-tuned LLMs. Conducted evaluations focused on structured ranking, rewriting, and factual accuracy to boost instruction adherence, safety compliance, and quality of responses. Recognized patterns of hallucinations, context drift, and cultural misalignment in Hindi-English model outputs.
  • AI Evaluation Specialist

    Alignerr · 2024 — Present
    Reviewed outputs from generative image and video models focusing on contextual grounding and indicators of bias. Provided structured feedback aimed at enhancing multilingual reliability and maintaining quality consistency.
  • Senior Linguistic Consultant & Editorial Specialist

    Company not specified · 2009 — Present
    Directed editorial projects, including a bilingual poetry collection and a multi-author anthology. Ensured high-accuracy academic translations by maintaining terminological and contextual integrity, along with semantic fidelity. Managed the publishing process, which included editing, typography coordination, and production oversight.

Skills & Expertise

Education

  • Bachelor in Tourism Studies
    Indira Gandhi National Open University
  • Diploma in Computer Applications
    Indira Gandhi National Open University
  • Diploma in Urdu Language
    Indira Gandhi National Open University

Interested in this professional?

Sign in as an employer to save this profile or invite Sandeep D. to a job.

Sign in as an employer