Pradheep1647/kibitzer-s2-shaw-142m-comp
Kibitzer clean-rebuild checkpoint.
- Training objective:
policy_value_competition_continuation - Checkpoint:
S2_shaw_142M_comp.pt - Architecture: 32.1M-parameter position-only transformer/SSM hybrid
- Policy output: 4,672 AlphaZero-style move logits
- Value output: side-to-move scalar in
[-1, 1] - Source: https://github.com/Mantissagithub/kibitzer/tree/clean-rebuild
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support