Bityuno Zero Qwen2.5-3B Countdown
Bityuno Zero is an implementation inspired by TinyZero, designed to develop self-verification and search skills through reinforcement learning. This model is based on Qwen2.5-3B and has been specifically trained for the "Countdown" task, its so experimental, check the repo for more information!
- Downloads last month
- 2
This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social
visibility and check back later, or deploy to Inference Endpoints (dedicated)
instead.
Model tree for JackCloudman/bityuno-zero-qwen2.5-3B-countdown
Base model
Qwen/Qwen2.5-3B