CoolFace
20 results

verl

Juvenilecris /waterbird-verlimage10K<n<100K0 likes1.9k downloads1y agoHugging FaceMiical /verl_vla_libero_collectedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "panda", "total_episodes": 32, "total_frames": 2990, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 1e-06, "fps": 10, "splits": { "train": "0:32" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Miical/verl_vla_libero_collected.tabularrobotics1K<n<10K0 likes1.7k downloads3mo agoHugging FaceSeanWang0027 /verl_mask_training 👋 Hi, everyone! verl is a RL training library initiated by ByteDance Seed team and maintained by the verl community. verl: Volcano Engine Reinforcement Learning for LLMs verl is a flexible, efficient and production-ready RL training library for large language models (LLMs). verl is the open-source version of HybridFlow: A Flexible and Efficient RLHF Framework paper. verl is flexible and easy to use with: Easy extension of diverse RL algorithms: The… See the full description on the dataset page: https://huggingface.co/datasets/SeanWang0027/verl_mask_training.0 likes1k downloads5mo agoHugging Facesungyub /deepscaler-preview-verl DeepScaleR-Preview VERL 📊 Dataset Summary This dataset contains 35,789 mathematical reasoning problems in VERL format, processed from agentica-org/DeepScaleR-Preview-Dataset. Key Features: 35,789 high-quality math problems Converted to VERL format for reward modeling Verified ground truth answers Ready for reinforcement learning training 🔗 Source Dataset Original Repository Repository:… See the full description on the dataset page: https://huggingface.co/datasets/sungyub/deepscaler-preview-verl.texttext-generation10K<n<100K0 likes886 downloads3mo agoHugging Facesliuau /DeepScaleR-Preview-Dataset-verl-formattext10K<n<100K0 likes743 downloads11mo agoHugging Faceverl-team /gsm8k-v0.4.1The dataset is generated based on verl 0.4.1 with command: python3 examples/data_preprocess/gsm8k.py text1K<n<10K0 likes605 downloads1y agoHugging Face