LLM Prompt Evaluation Tracker
Record and assess prompts used with large language models to ensure quality and effectiveness.
Evaluator Name
*
First Name
Last Name
Date of Evaluation
*
 -
Month
 -
Day
Year
2 digit month, 2 digit day, 4 digit year
Date
Prompt Text
*
Task or Context Description
*
LLM Model Used
*
Please Select
GPT-4
GPT-3.5
Claude 3
Llama 2
Other
LLM Output (copy/paste the generated response)
*
Prompt Evaluation Criteria
*
Rows
Clarity
Relevance
Completeness
Bias/Fairness
Creativity
Excellent
1
2
3
4
5
Good
6
7
8
9
10
Fair
11
12
13
14
15
Poor
16
17
18
19
20
Overall Prompt Quality
*
1
2
3
4
5
What are the strengths of this prompt?
What are the weaknesses or areas for improvement?
Would you recommend this prompt for future use?
*
Yes
No
With Modifications
Prompt Category
Please Select
Question Answering
Summarization
Creative Writing
Instruction Following
Other
Submit Evaluation
Should be Empty: