enhancementhelp wanted
Repository metrics
- Stars
- (16,939 stars)
- PR merge metrics
- (PR metrics pending)
Description
Evaluation dataset generation is coming to deepeval by end of this week. For this feature, we're looking at the following:
- Allow users to generate test cases based on their knowledge base
- Allow users to choose how to chunk their knowledge base
- Allow users to specify how many test cases to generate
- Allow users to complicate test cases to make them more realistic (https://arxiv.org/pdf/2304.12244.pdf, https://mlabonne.github.io/blog/notes/Large%20Language%20Models/phi1.htm)
Feedback, suggestions, and contributions for any of the points above are welcomed 😊