Evaluating Large Language Models responses

January 30, 2025 -
  January 30, 2025

On line

ORGANISERS

GALA/ EUATC

The final session of the 3-part prompt design webinar focuses on evaluating the output of prompts, diagnosing issues, and refining the approach. You’ll learn to measure LLM performance, choose the right evaluation data set, and mitigate risks. Topics include:
  • Methods for prompt evaluation.
  • Understanding the F1 score.
  • Types of hallucinations and strategies for managing them.
  • Recognizing and addressing data contamination and other risks.
  • Best practices for robust prompt design

Please note that each session in the series is free to EUATC and GALA members. There is a charge of $75 per session for non-members of both organisations.

EUATC Network members should apply for the discount code from your national association to access this webinar FREE-OF-CHARGE

 

 

 

Share this event:

ELIS SURVEY 2021