SupraLabs Releases Large-Scale Chat Title Dataset with 115K Entries
Key point
SupraLabs has released a large-scale chat title dataset with 115K entries for instruction tuning and benchmarking.
Details
SupraLabs has released a Chat Title generation dataset on Hugging Face that surpasses the scale of the existing major dataset, ogrnz/chat-titles.
This dataset significantly expands the scale from the previous level of 10K entries to 115K entries, and is provided in two versions depending on the use case.
- Filtered version:
SupraLabs/chat-titles-filtered-115K(recommended for general training and tuning) - Unfiltered version:
SupraLabs/chat-titles-unfiltered-150K(for custom filtering)
This dataset is designed for Instruction Tuning of models, classification-style title generation, and benchmarking the performance of Small Models.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.