OpenDataArena/ODA-Fin-RL-8B download history

OpenDataArena/ODA-Fin-RL-8B is an 8.2B-parameter reinforcement learning model by OpenDataArena. In the last 30 days it was downloaded 275 times (18 in the last 7 days), and 1,650 times in total.

It ranks #111,760 on the Hub by monthly downloads and #383 among reinforcement learning models.

It has 4 likes.

It is a fine-tune of OpenDataArena/ODA-Fin-SFT-8B.

1 models build on ODA-Fin-RL-8B: 1 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 395 times in the last 30 days. See the ODA-Fin-RL-8B galaxy.

Most downloaded derivatives

Open OpenDataArena/ODA-Fin-RL-8B on Hugging Face