🚀NEW LABGetting Started with Claude AgentsStart lab
Evaluation

Can LLMs Design Good Questions?

First page
Can LLMs Design Good Questions?
Paper summary

systematically evaluates the quality of questions generated with LLMs; here are the main findings: 1) there is a strong preference for asking about specific facts and figures in both LLaMA and GPT models, 2) the question lengths tend to be around 20 words but different LLMs tend to exhibit distinct preferences for length, 3) LLM-generated questions typically require significantly longer answers, and 4) human-generated questions tend to concentrate on the beginning of the context while LLM-generated questions exhibit a more balanced distribution, with a slight decrease in focus at both ends.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack