WithinUsAI/claude_mythos_distilled_25k
Dataset com 10 mil – 100 mil exemplos — 3.0 mil downloads no Hugging Face. Claude Mythos Distilled 25K A high-quality synthetic supervised fine-tuning (SFT) dataset designed to train and fine-tune any LLM to mirror the capabi…
O dataset WithinUsAI/claude_mythos_distilled_25k está entre os destaques do Hugging Face — dados que alimentam o treinamento e a avaliação dos modelos do momento.
Ficha do dataset
- Tamanho: 10 mil – 100 mil exemplos
- Idiomas: inglês
- Licença: Apache 2.0
- Downloads: 3.0 mil · Curtidas: 166
Sobre o dataset
Claude Mythos Distilled 25K A high-quality synthetic supervised fine-tuning (SFT) dataset designed to train and fine-tune any LLM to mirror the capabilities, reasoning style, agentic behavior, and technical depth of Anthropic's Claude Mythos (distilled frontier model).
Como carregar
Use a biblioteca datasets do Hugging Face:
pip install -U datasets
from datasets import load_dataset
ds = load_dataset("WithinUsAI/claude_mythos_distilled_25k")
print(ds)
print(ds["train"][0])Tags
synthetic claude mythos distillation cybersecurity coding reasoning agentic
Leia também
Unicorn, pelican, Middle-earth: OpenAI co-founder Karpathy is looking for the next AI vibe test
CAPA: o benchmark que mede se o assistente de código aprende com você — ou repete a mesma pergunta
Why biological data matters more in AI drug discovery