CoolFace
20 results

assistant

AssistantBench /AssistantBench Bibtex citation @misc{yoran2024assistantbenchwebagentssolve, title={AssistantBench: Can Web Agents Solve Realistic and Time-Consuming Tasks?}, author={Ori Yoran and Samuel Joseph Amouyal and Chaitanya Malaviya and Ben Bogin and Ofir Press and Jonathan Berant}, year={2024}, eprint={2407.15711}, archivePrefix={arXiv}, primaryClass={cs.CL}, url={https://arxiv.org/abs/2407.15711}, } textquestion-answeringn<1K23 likes7.5k downloads2y agoHugging Facenvidia /Nemotron-RL-agent-workplace_assistant Dataset Description: The Nemotron-RL-agent-workplace_assistant is a tool use - multi step agentic environment that tests the agent’s ability to execute tasks in a workplace setting. Workbench contains a sandbox environment with five databases, 26 tools, and 690 tasks. These tasks represent common business activities, such as sending emails, scheduling meetings, etc. This dataset is released as part of NVIDIA NeMo Gym, a framework for building reinforcement learning environments… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/Nemotron-RL-agent-workplace_assistant.text1K<n<10K31 likes3.3k downloads7mo agoHugging Facelu-christina /assistant-axis-vectors The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models This repository contains pre-computed axes and persona vectors for Gemma 2 27B, Qwen 3 32B, and Llama 3.3 70B, as described in the paper The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models. Paper | Code | Demo The Assistant Axis is a direction in activation space that captures how "Assistant-like" a model's current persona is. It can be used to: Monitor persona… See the full description on the dataset page: https://huggingface.co/datasets/lu-christina/assistant-axis-vectors.other12 likes2.4k downloads8mo agoHugging FaceGIZ /audit_assistant_reportsn<1K0 likes1.7k downloads1y agoHugging Faceglaiveai /glaive-code-assistant-v3 Glaive-code-assistant-v3 Glaive-code-assistant-v3 is a dataset of ~1M code problems and solutions generated using Glaive’s synthetic data generation platform. This is built on top of the previous version of the dataset that can be found here. This already includes v1 and v2 of the dataset. To report any problems or suggestions in the data, join the Glaive discord text100K<n<1M62 likes1.7k downloads2y agoHugging Faceglaiveai /glaive-code-assistant Glaive-code-assistant Glaive-code-assistant is a dataset of ~140k code problems and solutions generated using Glaive’s synthetic data generation platform. The data is intended to be used to make models act as code assistants, and so the data is structured in a QA format where the questions are worded similar to how real users will ask code related questions. The data has ~60% python samples. To report any problems or suggestions in the data, join the Glaive discord text100K<n<1M105 likes1.6k downloads3y agoHugging Face