BAEM1N/Korean-RAG-LLM-Judge-Benchmark download history
BAEM1N/Korean-RAG-LLM-Judge-Benchmark is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 72 times (28 in the last 7 days), and 1,518 times in total. It ranks #143,699 among datasets by monthly downloads.
Korean RAG LLM-as-Judge Benchmark allganize/RAG-Evaluation-Dataset-KO의 한국어 RAG 300 Q&A 위에 46개 답변 생성 모델의 답변과 LLM-as-Judge 평가 결과를 얹은 데이터셋입니다. 원문 문항·기준 답변·출처 PDF 메타데이터는 allganize 데이터셋을 그대로 따릅니다. 🔗 분석 코드·단계별 실험 보고서: https://github.com/BAEM1N/RAG-Evaluation · 📊 요약 대시보드: https://rag.baeum.
Open BAEM1N/Korean-RAG-LLM-Judge-Benchmark on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.