Summary
Overview
Work History
Education
Skills
Timeline
Generic
Kashif  Zahoor

Kashif Zahoor

Senior AI Trainer | Senior LLM Evaluator | RLHF Specialist

Summary

Results-oriented Senior AI Trainer and LLM Evaluator with over nine years of experience in AI initiatives, including data annotation, LLM evaluation, and quality assurance. Expertise in benchmark-testing conversational AI and STEM reasoning models using structured rubrics. Strong analytical background in electrical engineering and information technology, with proficiency in Python, SQL, and data validation.

Overview

2
2
Languages
2027
2027
years of professional experience

Work History

Senior AI Trainer & LLM Evaluator

Turing
2024 - Current
  • Evaluated state-of-the-art LLMs and agentic workflows to enhance multi-step reasoning, ensure safety protocols, verify factual accuracy, and improve effective tool calling.
  • Performed rigorous RLHF-style ranking, prompt refinement, and established high-density feedback loops to drive model alignment and behavioral tuning.
  • Authored comprehensive evaluation reports detailing model deficiencies, edge-case failures, and providing actionable optimization strategies for improvement.

AI Trainer & LLM Evaluator

Handshake.ai
12.2025 - 01.2026
  • Annotated and audited large-scale multimodal datasets, ensuring data quality and timely delivery.
  • Detected and mitigated labeling inconsistencies and systemic biases, enhancing dataset reliability and classification accuracy.
  • Developed and delivered training programs for AI models, enhancing understanding of machine learning concepts.
  • Collaborated with cross-functional teams to refine AI algorithms and improve user experience.
  • Analyzed training data to assess effectiveness, implementing adjustments to optimize instruction methods.

AI Trainer & LLM Evaluator

Pareto
08.2025 - 11.2025
  • Engineered and evaluated advanced STEM-centric prompts, improving assessment accuracy through structured, rubric-based scoring frameworks.
  • Validated analytical correctness of model-generated reasoning chains, ensuring mathematical and logical validity through Python scripts.
  • Developed and delivered comprehensive training programs on AI technologies and applications.
  • Facilitated workshops to enhance understanding of machine learning concepts among diverse audiences.
  • Assessed trainee performance and provided tailored feedback to improve competency in AI tools.

AI Trainer & Search Engine Evaluator

Oneforma
2021 - 2023
  • Evaluated search relevance, classified user intent, and reviewed content quality to enhance search algorithm performance.
  • Adapted search criteria for localized nuances to improve user-intent mapping across regional variations.
  • Developed and implemented training programs for AI model optimization.
  • Evaluated trainee performance to enhance learning outcomes and effectiveness.
  • Collaborated with cross-functional teams to integrate AI solutions into business processes.

AI Annotation Specialist

Telus International
2019 - 2021
  • Produced high-quality NLP annotations, intent mapping, and entity recognition to enhance conversational AI functionality.
  • Collaborated with QA teams to implement quality assurance pipelines, ensuring accuracy and reliability of core training datasets.
  • Led cross-functional teams to enhance customer experience through process optimization.
  • Developed training materials and conducted workshops to upskill staff on new technologies.
  • Analyzed performance metrics to identify trends and drive strategic initiatives for operational improvement.

LLM Evaluator & Data Labeling Contractor

Alignerr
2016 - 2017
  • Facilitated multilingual data annotation, text classification, and foundational prompt preparation to enhance model training accuracy.
  • Conducted comprehensive evaluations of program effectiveness to inform strategic decisions.
  • Collaborated with cross-functional teams to implement process improvements and enhance service delivery.
  • Developed and maintained evaluation frameworks aligned with organizational goals and objectives.

Education

Master of Science - Information Technology

University of The Punjab
Lahore, Punjab, Pakistan
04.2001 -

Bachelor of Science - Electrical Engineering (Computer Science)

University of Engineering & Technology (UET)
Lahore, Punjab, Pakistan
04.2001 -

Skills

AI Evaluation & Alignment: RLHF, Human Preference Ranking, Comparative Response Evaluation, Hallucination Detection, Bias & Safety Mitigation

Prompt Engineering: Prompt Optimization, Advanced Instruction Following, Few-Shot Prompting, Adversarial Testing (Red Teaming)

Advanced AI Capabilities: Agentic System Evaluation, Tool Use Validation, Multimodal Annotation, NLP, Search Relevance, Query Classification

Model Evaluation

Domain Expertise: STEM Evaluation, Rubric-Based Scoring, Model Benchmarking, Quality Assurance Reporting

Data Annotation: Video Annotation, Image Annotation, Text Labeling

Technical Proficiencies: Python, SQL, JSON, Microsoft Excel, Data Validation Pipelines

Timeline

AI Trainer & LLM Evaluator

Handshake.ai
12.2025 - 01.2026

AI Trainer & LLM Evaluator

Pareto
08.2025 - 11.2025

Master of Science - Information Technology

University of The Punjab
04.2001 -

Bachelor of Science - Electrical Engineering (Computer Science)

University of Engineering & Technology (UET)
04.2001 -

Senior AI Trainer & LLM Evaluator

Turing
2024 - Current

AI Trainer & Search Engine Evaluator

Oneforma
2021 - 2023

AI Annotation Specialist

Telus International
2019 - 2021

LLM Evaluator & Data Labeling Contractor

Alignerr
2016 - 2017
Kashif Zahoor Senior AI Trainer | Senior LLM Evaluator | RLHF Specialist