This repository contains the core data, checkpoints, and training records for the paper: "ReinFlow: Fine-Tuning Flow Matching Policy with Online Reinforcement Learning".
data-offline: This directory includes the train.npz and normalization.npz files for OpenAI Gym tasks, derived and normalized from the official D4RL datasets. An exception is the Humanoid-v3 environment, where the data was collected from our pre-trained SAC… See the full description on the dataset page:
https://huggingface.co/datasets/ReinFlow/ReinFlow-data-checkpoints-logs.