BenchmarkPromptsWithResponses. Every prompt engineering paper should provide not only on-average performance of the prompting strategy, but should also release the responses to facilitate future research and avoid repeatedly calling the LLMs for the same queries+prompts.

github.com/holarissun/BenchmarkPromptsWithResponses

Vaya's read on this project

Problem, audience, market, and the verdict — sign in to see it.

Updates

No recent activity.