Use when the user wants to build, validate, or submit RL training environments on Prime Intellect Lab using the verifiers library. Covers environment creation from HuggingFace datasets, reward function authoring and validation, pushing to the Environments…
2026年2月24日