dipankarsarkar/llm-evaluation-self-audit download history
dipankarsarkar/llm-evaluation-self-audit is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 293 times (293 in the last 7 days), and 293 times in total. It ranks #50,050 among datasets by monthly downloads.
LLM Evaluation Self-Audit Data for How Reproducible Are Evaluation Conclusions? A Self-Audit of LLM-Inferred Prompt Structure, Dipankar Sarkar, Skelf Research. Code and paper source: github.com/sarkar-dipankar/llm-evaluation-self-audit (this package is built from commit 687ee8a).
Open dipankarsarkar/llm-evaluation-self-audit on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.