Reinforcement Learning
PEFT
Safetensors
English
Chinese
grpo
lora
game-agent
snake
system-one
jev
kev
Instructions to use ZhangPY/kev-snake-text-grpo with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- PEFT
How to use ZhangPY/kev-snake-text-grpo with PEFT:
from peft import PeftModel from transformers import AutoModel base_model = AutoModel.from_pretrained("Qwen/Qwen3.5-4B-Base") model = PeftModel.from_pretrained(base_model, "ZhangPY/kev-snake-text-grpo") - Notebooks
- Google Colab
- Kaggle
Welcome to the community
The community tab is the place to discuss and collaborate with the HF community!