datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
wb-feedbacks
Dataset Card for Wildberries products
Dataset Summary
The dataset contains product reviews from the Russian marketplace Wildberries, collected by mining about The dataset was collected by bruteforcing possible product identifiers (about 230 million) and querying all available feedbacks for them. The data are stored in zstd-archives containing jsonl-files. The 'nmId' in the dataset usually corresponds to the valid product article on the site, but sometimes reviews are… See the full description on the dataset page: https://huggingface.co/datasets/nyuuzyou/wb-feedbacks.repro-a-tight-theory-of-error-feedback-algorithms-in-distributed-optimization-traces
Agent traces
Agent sessions published from a Trackio Logbook.
ko_ultrafeedback_gemini_feedbackmaywell/ko_Ultrafeedback_binarized 중 12000 여개의 chosen 을 Google Gemini Pro를 이용해서 피드백하고 점수를 평가
Self_Ref_Feedbackfeedback_qesconvfeedback_qesconvThe Feedback-ESConv dataset.
Paper: Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors.
GitHub repo
@inproceedings{chaszczewicz2024multilevel,
title={Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors},
author={ Chaszczewicz, Alicja and
Shah, Raj Sanjay and
Louie, Ryan and
Arnow, Bruce A. and
Kraut, Robert and… See the full description on the dataset page: https://huggingface.co/datasets/SALT-NLP/feedback_qesconv.han-humanoid-adaptive-action-feedback-v1
Humanoid Adaptive Action Feedback
Overview
This dataset contains structured feedback
on executed humanoid actions and subsequent adjustments.
It supports reinforcement and adaptive behavior modeling.
Data Fields
action_id
executed_action
environment_state
human_feedback_score
adjustment_applied
post_adjustment_score
Intended Use
Reinforcement learning research
Continuous improvement systems
Adaptive humanoid frameworks
License
MIT
humanoid-learning-feedback-dataset
Humanoid Learning Feedback Dataset
Dataset for tracking humanoid performance feedback
after task execution.
Description
Includes reward scores, execution time,
and error signals for adaptive learning.
File
learning_feedback_dataset.json
License
MIT
ultrafeedback_binarized_feedbackfeedback_qesconv_badareas_questions_reflectionsValidation Results:
PAIR-Reflections:
• "Precision": 0.8955223880597015
• "Recall": 0.2553191489361703
• "Accuracy": 0.77
• "F1-score": 0.3973509933774836
Rule-based HasQuestions:
• "Precision": 1.0
• "Recall": 0.9629629629629629
• "Accuracy": 0.74
• "F1-score": 0.9811320754716981
llm-matcher-feedback
