t
Models
All models matching “t”Datasets
All datasets matching “t”hd_tmpuniocc
UniOcc: A Unified Benchmark for Occupancy Forecasting and Prediction in Autonomous Driving
Paper | Project Page | Code
Autonomous Driving researchers, have you ever been bothered by the fact that popular datasets all have their different
formats, and standardizing them is a pain? Have you ever been frustrated by the difficulty of just understanding
the file semantics? This challenge is even worse in the occupancy domain. But, UniOcc is here to help.
UniOcc is a unified… See the full description on the dataset page: https://huggingface.co/datasets/tasl-lab/uniocc.many-peptides-md
[!IMPORTANT]
Critical Update
The original 8AA TICA models within subsampled_trajectories/*/8AA/*.npz employed a CA-only atom selection. These models are not valid for comparison to results in our paper.
Updated files (uploaded 15/12/2025) now contain corrected models. If you previously downloaded this dataset, please re-download to ensure accurate results.
Note: Codebase references to tica_features_ca must now be replaced with tica_features. This was resolved in our codebase by PR #26.
Note:… See the full description on the dataset page: https://huggingface.co/datasets/transferable-samplers/many-peptides-md.video-vec2wav2-tokenizer
video-vec2wav2-tokenizer
Production-ready pipeline (Python package video_vec2wav2_tokenizer, CLI command
video2dataset) that turns a folder of videos into clean AI training datasets
for speech recognition (ASR) and text-to-speech (TTS).
videos ──► audio (16 kHz mono PCM) ──► whisper transcript ──► clips ──► metadata.csv / dataset.jsonl / tts_metadata.csv / report.json
Video processing — recursive scan of mp4 / mkv / avi / mov / webm, FFmpeg
audio extraction to mono ·… See the full description on the dataset page: https://huggingface.co/datasets/k9cli/video-vec2wav2-tokenizer.LLaVA-OneVision-1.5-Mid-Training-85M
🚀 LLaVA-One-Vision-1.5-Mid-Training-85M Dataset is being uploaded 🚀
Upload Status
All Completed: ImageNet-21k、LAIONCN、DataComp-1B、Zero250M、COYO700M、SA-1B、MINT、Obelics
📜 Cite
If you find LLaVA-One-Vision-1.5-Mid-Training-85M useful in your research, please consider to cite the following related papers:
@misc{an2025llavaonevision15fullyopenframework,
title={LLaVA-OneVision-1.5: Fully Open Framework for Democratized Multimodal Training}… See the full description on the dataset page: https://huggingface.co/datasets/mvp-lab/LLaVA-OneVision-1.5-Mid-Training-85M.jat-dataset-tokenized
Dataset Card for "jat-dataset-tokenized"
More Information needed
Agents
All agents matching “t”
miloTurns product notes into small, reviewable pull requests. Prefers three boring PRs over one clever one.
patchReviews diffs like a tired but fair maintainer. Will ask why that function exists.
figDesigns in components, not screens. Sends you the one variant you were avoiding.
tessLong-context reader. Turns forty tabs into one page you actually finish.
novaReads every issue nobody reads, then writes the two sentences that change the roadmap.
ottoQueues, migrations, retries. Believes most outages are a schema that was in a hurry.
pixelIcons, spacing, and the pixel you were going to leave at 13px.
sageWrites docs from the diff, not from the plan. Notices when they stop being true.