Dataset library
Every dataset we train on is public. Reasoning traces with full thinking, multi-turn agent sessions with tool calls, and plain chat data, generated from frontier models and formatted for supervised fine-tuning.
Claude Fable 5
Fable-5-Cursor-Traces
244 Fable 5 Cursor agent sessions for training & research. This dataset has 244 Cursor sessions with Fable 5 at High/xHigh/Max effort levels for distillation. [!IMPORTANT] This dataset is compatible with Teich! Use it directly in your…
Ox Alpha
Ox-Alpha-Pi-Traces
This directory contains raw agent trace files generated by teich. JSONL files: 2247 Model metadata: stealth/ox-alpha Domains and prompt distribution Topic Traces Games & simulation (headless) 196 Frontend & Node-testable web 159 Health…
Ox Alpha
Ox-Alpha-10k
10,005 single-turn prompts for text-response teacher generation Each row carries id, category, subcategory All data was gathered using stealth/ox-alpha via OpenRouter (reasoning effort high) Topic distribution Category Rows Share Coding…
DeepSeek V4 Pro
DeepSeek-v4-Pro-Agent
This directory contains raw agent trace files generated by teich. All assistant responses were generated by deepseek/deepseek-v4-pro. JSONL files: 4006 Training-ready tools A complete configured tools schema snapshot is embedded in the…
DeepSeek V4 Flash
DeepSeek-v4-Flash-Chat
Teich Test This directory contains newline-delimited JSON training examples generated by teich. All assistant responses were generated by deepseek/deepseek-v4-flash. Rows: 6313 Format Each file is newline-delimited JSON where every line…
Claude Opus 4.7
lordx64-claude-opus-4.7-max-cleaned
reasoning-distill-claude-opus-4-7-max-cleaned Cleaned version of lordx64/reasoning-distill-claude-opus-4-7-max. See the original dataset for full provenance, collection methodology, and terms of use. Cleaning steps Step Filter Reason…
Claude Sonnet 4.6
Claude-Sonnet-4.6-Reasoning-1100x
Claude Sonnet 4.6 - High Reasoning 1096 conversations, all single-turn user → assistant pairs created using Claude Sonnet 4.6 with reasoning effort set to high. This is a pure reasoning/critical-thinking distillation dataset. Heavily…
Claude Opus 4.6
Claude-Opus-4.6-Reasoning-887x
Claude Opus 4.6 - High Reasoning This is a reasoning dataset generated using Claude Opus 4.6 with high reasoning effort It contains distilled reasoning traces from Bullshit Bench for bullshit detection, legal and life decisions data for…
Hunter Alpha
Hunter-Alpha-UIGEN-T3-Agent-SFT
Hunter Alpha UIGEN T3 All of the prompts for this dataset were sourced from Tesslate/UIGEN-T3-Dataset-Extended-Reasoning, and the rest were generated. Unfortunately the model was taken down from openrouter and revealed as…
Hunter Alpha
Hunter-Alpha-Coding-Agent-SFT
200 of the prompts for this dataset were sourced from MiniMaxAI/VIBE, and the rest were generated. Each prompt was given to Hunter-Alpha (The stealth model recently revealed to be xiaomi/mimo-v2-pro) with the follow tools and system…
Healer Alpha
Healer-Alpha-16k
This is a reasoning dataset generated using the stealth model Healer Alpha. As the largest dataset we have made yet. The prompts from this dataset were almost all generated by Opus 4.5/4.6, GPT 5.1 and Gemini 3 (flash and pro). The…
Hunter Alpha
Hunter-Alpha-16k
This is a reasoning dataset generated using the stealth model Hunter Alpha, which was revealed to be xiaomi/mimo-v2-pro. As the largest dataset we have made yet. The prompts from this dataset were almost all generated by Opus 4.5/4.6,…
Claude Opus
Claude-Opus-Dataclaw-Unredacted
How this dataset was built Collected the local Petromallet raw export plus selected public Dataclaw uploads. Filtered to the supported Opus-family source rows. Deduplicated by session_id and first user message. Converted raw assistant…
Aurora Alpha
Aurora-Alpha-15.5k
This is a non-reasoning dataset generated using the stealth model Aurora Alpha. The prompts from this dataset were almost all generated by GPT 5.1 and Gemini 3 (flash and pro). The categories covered include academia, multi-lingual…
Pony Alpha
Pony-Alpha-15k
This is a reasoning dataset generated using the stealth model Pony Alpha, which ended up being GLM-5. As the largest dataset we have made yet. The prompts from this dataset were almost all generated by GPT 5.1 and Gemini 3 (flash and…
Step 3.5 Flash
Step-3.5-Flash-2600x
Step 3.5 Flash - 2,600x This is a reasoning dataset created using Step 3.5 Flash with a reasoning depth set to high. The dataset is meant for creating distilled versions of Step 3.5 FLash by fine-tuning already existing open-source…
Mistral Small
mistral-small-creative-500x
This is a non-reasoning dataset created using Mistral Small Creative. The dataset is meant for creating distilled versions of Mistral Small Creative by fine-tuning already existing open-source LLMs. This dataset only covers generating…
MiniMax M2.1
MiniMax-M2.1-Code-SFT
200 of the prompts for this dataset were sourced from MiniMaxAI/VIBE. The rest were generated. Each prompt was given to MiniMax M2.1 with the follow tools and system prompt: read_file - Read file contents from workspace write_file -…
Gemini 3 Flash
Gemini-3-Flash-Preview-VIBE
This dataset is our first attempt at an agentic coding SFT dataset. All of the prompts for this dataset were sourced from MiniMaxAI/VIBE. Each prompt was given to Gemini 3 Flash Preview with the follow tools and system prompt: read_file…
Mixed sources
convo-v1
A conversation dataset generated using MiniMax M2.1 to help teach models how to talk to users.
MiniMax M2.1
MiniMax-M2.1-8800x
8,800x This is a reasoning dataset created using MiniMax M2.1 with reasoning effort set to high (not sure if that flag does anything for this model though). The dataset is meant for creating distilled versions of MiniMax M2.1 by…
MiMo V2 Flash
MiMo-V2-Flash-2300x
MiMo V2 Flash - 2,300x This is a reasoning dataset created using MiMo V2 Flash with a reasoning depth set to high. The dataset is meant for creating distilled versions of MiMo V2 Flash by fine-tuning already existing open-source LLMs.…
Gemini 3 Flash
gemini-3-flash-preview
This is a reasoning dataset created using Gemini 3 Flash Preview with a reasoning depth set to high. The dataset is meant for creating distilled versions of Gemini 3 Flash Preview by fine-tuning already existing open-source LLMs. This…
GLM 4.7
glm-4.7-350x
This is a reasoning dataset created using GLM 4.7. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of 4.7 by fine-tuning already existing open-source LLMs.…
MiniMax M2.1
minimax-m2.1-1000x
This is a reasoning dataset created using MiniMax M2.1 with a reasoning depth set to high. The dataset is meant for creating distilled versions of MiniMax M2.1 by fine-tuning already existing open-source LLMs. Some of these prompts are…
GLM 4.7
glm-4.7-2000x
This is a reasoning dataset created using GLM 4.7 with reasoning effort set to high (not sure if that flag does anything for this model though). The dataset is meant for creating distilled versions of GLM 4.7 by fine-tuning already…
Claude Haiku 4.5
claude-haiku-4.5-high-reasoning-1700x
This is a reasoning dataset created using Claude Haiku 4.5 with reasoning effort set to high. The dataset is meant for creating distilled versions of Claude Haiku 4.5 by fine-tuning already existing open-source LLMs. This dataset…
Claude Haiku 4.5
claude-haiku-4.5-1700x
This is a non-reasoning dataset created using Claude Haiku 4.5. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of Claude Haiku 4.5 by fine-tuning already…
GPT-5.1 Codex Max
gpt-5.1-codex-max-1000x
GPT 5.1 Codex Max - 1,000x This is a reasoning dataset created using GPT 5.1 Codex Max with a reasoning depth set to high. The dataset is meant for creating distilled versions of GPT 5.1 Codex Max by fine-tuning already existing…
Gemini 3 Flash
gemini-3-flash-preview-1000x
Gemini 3 Flash Preview - 1,000x This is a reasoning dataset created using Gemini 3 Flash Preview with a reasoning depth set to high. The dataset is meant for creating distilled versions of Gemini 3 Flash Preview by fine-tuning already…
Gemini 3 Flash
gemini-3-flash-preview-standalone-html-1k
This dataset was made by querying Gemini 3 Flash Preview to train models on web development for games, websites, and web apps in html + inline css and javascript. Prompt breakdown: First 800 targeted at standalone html websites, games,…
Mixed sources
open-moderator-v1
Open Moderator Summary Open Moderator is an English moderation dataset of ~11,000 chat-style examples derived from publicly submitted posts on Confess Your Sins. Each example is labeled into one of several safety categories to support…
GPT-5.2
gpt-5.2-high-reasoning-250x
Generated using DataGen by TeichAI This is a reasoning dataset created using GPT 5.2 with a reasoning depth set to high. The dataset is meant for creating distilled versions of GPT 5.2 by fine-tuning already existing open-source LLMs.…
DeepSeek V3.2 Speciale
deepseek-v3.2-speciale-OpenCodeReasoning-3k
The questions for this dataset were all sourced from the first 3k prompts in nvidia/OpenCodeReasoning Dataset Stats (provided by OpenRouter): Cost: $ 19.2 (USD) Tokens (input + output): 47 M
DeepSeek V3.2 Speciale
deepseek-v3.2-speciale-openr1-math-3k
Inspired by @OpenR1 The questions for this dataset were all sourced from the first 3.3k prompts in open-r1/OpenR1-Math-220k Dataset Stats (provided by OpenRouter): Cost: $ 21.1 (USD) Tokens (input + output): 52.3 M
DeepSeek V3.2 Speciale
deepseek-v3.2-speciale-1000x
This is a reasoning dataset created using Deepseek v3.2 Speciale with a reasoning depth set to high. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of…
Sherlock Alpha
sherlock-thinking-alpha-11000x
This dataset is unique in the sense that it is a non-reasoning dataset that was generated by a reasoning model (the stealth model that turned into grok 4.1 fast) The prompts from this dataset were generated by multiple models across the…
Claude Opus 4.5
claude-4.5-opus-high-reasoning-250x
This is a reasoning dataset created using Claude Opus 4.5 with a reasoning depth set to high. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of Claude Opus…
Gemini 3 Pro
gemini-3-pro-preview-high-reasoning-1000x
This is a reasoning dataset created using Gemini 3 Pro Preview with a reasoning depth set to high. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of Gemini 3…
GPT-5.1
gpt-5.1-high-reasoning-1000x
This is a reasoning dataset created using GPT 5.1 (high reasoning) from OpenAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development, Logic, Math,…
Gemini 3 Pro
gemini-3-pro-preview-high-reasoning-250x
This is a reasoning dataset created using Gemini 3 Pro Preview with a reasoning depth set to high. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of Gemini 3…
Sherlock Alpha
sherlock-dash-alpha-1000x
Sherlock Alpha
sherlock-think-alpha-1000x
Gemini 2.5 Flash
gemini-2.5-flash-11000x
Gemini 2.5 Flash - 11,000x This dataset was created by querying Gemini 2.5 Flash over 11,000 times with the goal of compiling the behavior, reasoning traces, output style, and (most importantly) knowledge of the model into a single…
GPT-5 Codex
gpt-5-codex-1000x
This is a reasoning dataset created using GPT 5 Codex from OpenAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development, Logic, Math, Embedded…
Kimi K2 Thinking
kimi-k2-thinking-1000x
This is a reasoning dataset created using Kimi k2 thinking from MoonshotAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development, Logic, Math,…
Gemini 2.5 Flash Lite
gemini-2.5-flash-lite-2509-preview-1000x
This is a reasoning dataset created using Gemini 2.5 Flash Lite Preview 09-2025 from Google. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development,…
Kimi K2 Thinking
kimi-k2-thinking-250x
This is a reasoning dataset created using Kimi k2 thinking from MoonshotAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development, Logic, Math,…
Polaris Alpha
polaris-alpha-1000x
This is a non-reasoning dataset created using Polaris Alpha (likely an alpha version of a new ChatGPT) from OpenAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following…
Grok 4 Fast
brainstorm-v3.1-grok-4-fast-200x
The responses for this dataset were generated with grok 4 fast. Original dataset: https://huggingface.co/datasets/DevQuasar/brainstorm-v3.1_vicnua_1k
Claude Sonnet 4.5
claude-sonnet-4.5-high-reasoning-250x
This is a reasoning dataset created using Claude Sonnet 4.5 with a high reasoning effort. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of Claude Sonnet 4.5…
GPT-5 Codex
gpt-5-codex-250x
GLM 4.6
glm-4.6-250x
This is a reasoning dataset created using GLM 4.6. Some of these questions are from reedmayhew and the rest were generated. The dataset is meant for creating distilled versions of 4.6 by fine-tuning already existing open-source LLMs.…
Grok Code Fast 1
grok-code-fast-1-1000x
This is a reasoning dataset created using Grok Code Fast 1 from xAI. Some of these questions are from reedmayhew and the rest were generated. Most of the questions cover the following topics: Web Development, Logic, Math, Embedded…