Airton

Airton

Zurich, Switzerland

I will evaluate and improve your LLM outputs with RLHF feedback

(0)
Remote 5 months ago
 75 $  - Per hour
I help AI teams improve their Large Language Model outputs through structured RLHF (Reinforcement Learning from Human Feedback) evaluation.

With 3+ years of hands-on experience at TaskUs and Translated, I identify weaknesses in model responses, annotate outputs systematically, and develop actionable feedback that measurably improves model quality.

I specialize in Portuguese-language LLM evaluation (PT-PT, PT-BR, PALOP) as well as English, making me especially valuable for teams targeting Lusophone markets. I've built annotation libraries, established quality frameworks, and trained other annotators on best practices.

What I deliver: Structured evaluation reports, annotated datasets, RLHF feedback cycles, and quality improvement recommendations — using annotation tools, quality metric frameworks, and pattern recognition to surface recurring issues.

Available for ongoing evaluation contracts or one-off model audits.
0.0 (0)
0
0
0
0
0
⚠️
Please sign in as a customer to give your feedback