Mario A. K.
AI Data, LLM Evaluation QA Specialist
Compétences

Voir mes services


Portfolio
Expérience professionnelle
AI Preference Ranking Evaluator (RLHF)
SuperAnnotate
Apr 2026 - Present • 4 mos
• Perform cross-lingual pairwise ranking for a major LLM. • Evaluate instruction execution, JSON/Markdown formatting compliance and tone. • Provide structured judgments distinguishing quality across competing model responses.
Greek LLM Evaluator (RLHF)
Welocalize
May 2026 - Aug 2026 • 3 mos
• Conduct multi-stage pairwise preference evaluations of Greek-language AI responses. • Assess outputs for factuality, safety, linguistic naturalness and overall response quality. • Apply detailed evaluation criteria consistently across large AI-generated response sets.
LLM Output Evaluator & Prompt Engineer
Mindrift
Dec 2025 - Apr 2026 • 4 mos
• Designed prompts for model capability benchmarking and evaluated outputs against client specifications. • Assessed instruction adherence and supported AI training/evaluation workflows through structured feedback.