Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Sanster 's Collections
LLM Training Dataset
multimodal

LLM Training Dataset

updated Mar 14, 2024
Upvote
1

  • teknium/OpenHermes-2.5

    Viewer • Updated Apr 15, 2024 • 1M • 18.4k • 883

  • Open-Orca/SlimOrca-Dedup

    Viewer • Updated May 19, 2025 • 363k • 599 • 93

  • argilla/ultrafeedback-binarized-preferences-cleaned

    Viewer • Updated Dec 11, 2023 • 60.9k • 12.6k • 163

  • argilla/ultrafeedback-multi-binarized-preferences-cleaned

    Viewer • Updated Dec 11, 2023 • 158k • 128 • 7

  • argilla/distilabel-intel-orca-dpo-pairs

    Viewer • Updated Aug 7, 2025 • 12.9k • 9.87k • 182

  • openchat/openchat_sharegpt4_dataset

    Updated Jul 1, 2023 • 765 • 173

  • rombodawg/LosslessMegaCodeTrainingV3_1.6m_Evol

    Viewer • Updated Oct 19, 2023 • 1.56M • 41 • 27

  • OpenAssistant/oasst2

    Viewer • Updated Jan 11, 2024 • 135k • 21.7k • 298

  • WizardLMTeam/WizardLM_evol_instruct_V2_196k

    Viewer • Updated Mar 10, 2024 • 143k • 4.18k • 250

  • lmsys/lmsys-chat-1m

    Viewer • Updated Jul 27, 2024 • 1M • 6.93k • 954

  • Hello-SimpleAI/HC3-Chinese

    Viewer • Updated Jan 21, 2023 • 25.7k • 2.22k • 175

  • argilla/dpo-mix-7k

    Viewer • Updated Jul 16, 2024 • 7.5k • 595 • 175
Upvote
1
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs