CoolFace
21 results

Net

code-search-net /code_search_net Dataset Card for CodeSearchNet corpus Dataset Summary CodeSearchNet corpus is a dataset of 2 milllion (comment, code) pairs from opensource libraries hosted on GitHub. It contains code and documentation for several programming languages. CodeSearchNet corpus was gathered to support the CodeSearchNet challenge, to explore the problem of code retrieval using natural language. Supported Tasks and Leaderboards language-modeling: The dataset can be used to… See the full description on the dataset page: https://huggingface.co/datasets/code-search-net/code_search_net.texttext-generation1M<n<10M338 likes41k downloads7mo agoHugging Facewayslab /llm-network-study-data LLM-Network-Study-Data Per-request network captures (.pcapng) collected by the LLM-Network-Study benchmark harness (benchmark.py and the per-workload test scripts). Each directory holds one capture file per request, named request_<id>_run<n>_<timestamp>.pcapng. A directory name encodes four dimensions: <capture-env>_<provider/model>_<workload>[_<dataset/variant>]_results Dimension legend Dimension Values Meaning Capture env ethernet Wired connection to… See the full description on the dataset page: https://huggingface.co/datasets/wayslab/llm-network-study-data.tabularn<1K0 likes30k downloads29d agoHugging Facenetflix /Vera-Layered-Video-Dataset Dataset for Vera: A Layered Diffusion Model for Content-Preserving Video Editing Hongkai Zheng¹²* &nbsp;·&nbsp; Ta-Ying Cheng² &nbsp;·&nbsp; Benjamin Klein² &nbsp;·&nbsp; Yisong Yue¹ &nbsp;·&nbsp; Zhuoning Yuan²† ¹California Institute of Technology &nbsp;&nbsp; ²Netflix, Inc. *Work done during an internship at Netflix &nbsp; †Project Lead TL;DR: A layered diffusion framework for video editing. Vera jointly generates an edit layer, an alpha… See the full description on the dataset page: https://huggingface.co/datasets/netflix/Vera-Layered-Video-Dataset.videotext-to-video10K<n<100K59 likes14k downloads2mo agoHugging Faceliuhyuu /NetEaseCrowd 🧑‍🤝‍🧑 NetEaseCrowd: A Dataset for Long-term and Online Crowdsourcing Truth Inference View it in GitHub Introduction We introduce NetEaseCrowd, a large-scale crowdsourcing annotation dataset based on a mature Chinese data crowdsourcing platform of NetEase Inc.. NetEaseCrowd dataset contains about 2,400 workers, 1,000,000 tasks, and 6,000,000 annotations between them, where the annotations are collected in about 6 months. In this dataset, we provide ground truths for… See the full description on the dataset page: https://huggingface.co/datasets/liuhyuu/NetEaseCrowd.tabular1M<n<10M1 likes5.7k downloads2y agoHugging Facerasoul-nikbakht /NetSpec-LLM 📁 Network Spec for LLM Understanding 📄 Overview This repository houses a comprehensive collection of ETSI (European Telecommunications Standards Institute) documents, systematically downloaded, processed, and organized for streamlined access and analysis. Each ETSI deliverable is paired with its corresponding metadata to ensure thorough information management. 🔍 Data Processing Workflow The data processing involves two main scripts that automate the… See the full description on the dataset page: https://huggingface.co/datasets/rasoul-nikbakht/NetSpec-LLM.text100K<n<1M5 likes5k downloads2y agoHugging FaceNan-Do /code-search-net-python Dataset Card for "code-search-net-python" Dataset Description Homepage: None Repository: https://huggingface.co/datasets/Nan-Do/code-search-net-python Paper: None Leaderboard: None Point of Contact: @Nan-Do Dataset Summary This dataset is the Python portion of the CodeSarchNet annotated with a summary column.The code-search-net dataset includes open source functions that include comments found at GitHub.The summary is a short description of what the… See the full description on the dataset page: https://huggingface.co/datasets/Nan-Do/code-search-net-python.texttext-generation100K<n<1M30 likes4.4k downloads3y agoHugging Face

Projects