Dataset Description: The SWE RL dataset provides GitHub issues for training and validating real world software engineering agents using the OpenHands environment in NeMo Gym. The dataset is a refactored version of the SWE Gym, R2E Gym, and SWE Bench Verified datasets to support the NeMo Gym input format. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments to train large language models. NeMo Gym contains a growing collection of training environments and datasets to enable Reinforcement Learning from Verifiable Reward (RLVR). This dataset was utilized in the development of the NVIDIA Nemotron family of models. NeMo Gym is an open source library within the NVIDIA NeMo framework, NVIDIA's GPU accelerated, end to end training framework for large language models (LLMs), multi modal models, and speech models. This dataset is part of the https://huggingface.co/collections/nvidia/nemo gym/ collection This dataset is ready for commercial use. Dataset Owner(s): NVIDIA Corporation Dataset Creation Date: 03/11/2026 License/Terms of Use: This dataset is licensed under Creative Commons Attribution 4.0 International (CC BY 4.0). Intend…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy