Pradheep1647/kibitzer-s2-shaw-142m-comp

Kibitzer clean-rebuild checkpoint.

  • Training objective: policy_value_competition_continuation
  • Checkpoint: S2_shaw_142M_comp.pt
  • Architecture: 32.1M-parameter position-only transformer/SSM hybrid
  • Policy output: 4,672 AlphaZero-style move logits
  • Value output: side-to-move scalar in [-1, 1]
  • Source: https://github.com/Mantissagithub/kibitzer/tree/clean-rebuild
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Pradheep1647/kibitzer-s2-shaw-142m-comp

Finetunes
1 model