pranavvmurthy26/synthetic-financial-tool-calling-grpo-rlvr-1k
🤖 Synthetic Financial Tool Calling Dataset for GRPO and RLVR This is a synthetic dataset designed for training language models on financial tool calling using GRPO (Group Relative Policy Optimization) with verifiable rewards (RLVR). The dataset contains ~1.1K examples of financial planning queries paired with expected tool calls and answers. Dataset sample schema, { "prompt": [ { "role": "system", "content": "You are a financial planning assistant with tools… See the full description on the dataset page: https://huggingface.co/datasets/pranavvmurthy26/synthetic-financial-tool-calling-grpo-rlvr-1k.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face