Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up
sergiopaniegoย 
posted an update 10 days ago
Post
2522
we've just added several example scripts to TRL showing how to train models with GRPO using some of the new OpenEnv environments

train a model to interact with a browser (๐ŸŽฎ BrowserGym Env), play Wordle (๐ŸŽฎ Wordle Env) and moooore!

TRL (GRPO + vLLM) + OpenEnv! โšก๏ธ

๐Ÿ“ go play with them: https://github.com/huggingface/trl/tree/main/examples/scripts/openenv

๐Ÿ“ examples list: https://huggingface.co/docs/trl/main/en/example_overview#scripts
In this post