ShawnLi02/FORTIS_Agent_Skill_Safety download history

ShawnLi02/FORTIS_Agent_Skill_Safety is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 73 times (19 in the last 7 days), and 1,168 times in total. It ranks #142,324 among datasets by monthly downloads.

FORTIS: Benchmarking Agent Skill Safety FORTIS is a benchmark for evaluating AI agent safety in skill and tool selection. It measures whether LLM agents select minimally-privileged capabilities when multiple valid options exist. Overview Modern LLM agents operate through

Open ShawnLi02/FORTIS_Agent_Skill_Safety on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.