Update ai assistants evals version 2 - #673
Conversation
Revise AI Evaluations documentation with details of version 2 release
Updated the document to reflect changes in the AI Assistants creation and modification process, including new versioning details and improved structure.
Updated images in the documentation for AI Assistants integration steps, ensuring correct display and context for users.
|
Important
This repository does not receive automatic reviews because it has fewer than 10 stars. ⚙️ Run configurationConfiguration used: Organization UI Review profile: ASSERTIVE Plan: Team Run ID: Thanks for using CodeRabbit! It's free for OSS, and your support helps us grow. If you like it, consider giving us a shout-out. Comment |
|
🚀 Deployed on https://deploy-preview-673--glific-docs.netlify.app |
Clarified instructions and added images for better understanding of the AI Assistant creation and editing process.
Refactor evaluation metrics section for clarity and consistency. Update instructions for setting up Golden Q&A sets and running evaluations.
| ### Step 3: Select an AI Assistant | ||
| Click the "Search or select an AI assistant" dropdown under AI Assistant and choose the specific assistant (and its version) you want to evaluate. | ||
| The overall score is a weighted average of three checks, each scored 0–5: | ||
| 1. Adherence to ground truth (weight 50%, how well answers match the golden answers), |
There was a problem hiding this comment.
Here - It would be helpful to add what drives the score—whether the metric is assessing if the core answer is correct, or how closely the response aligns with the ground truth in terms of content and scope.
There was a problem hiding this comment.
updated the definition
Corrected capitalization and phrasing for clarity in the AI Assistant documentation.
Updated the wording for clarity and consistency in the evaluation metrics and scoring description.
No description provided.