Bityuno Zero Qwen2.5-3B Countdown

Bityuno Zero is an implementation inspired by TinyZero, designed to develop self-verification and search skills through reinforcement learning. This model is based on Qwen2.5-3B and has been specifically trained for the "Countdown" task, its so experimental, check the repo for more information!

image/png

Downloads last month
2
Safetensors
Model size
3.09B params
Tensor type
FP16
·
Inference Examples
This model does not have enough activity to be deployed to Inference API (serverless) yet. Increase its social visibility and check back later, or deploy to Inference Endpoints (dedicated) instead.

Model tree for JackCloudman/bityuno-zero-qwen2.5-3B-countdown

Base model

Qwen/Qwen2.5-3B
Finetuned
(37)
this model