claude-opus-4.6-4.7-reasoning-8.7k angrygiraffe

Background Ended up with some tokens to burn on a Claude Max plan. Assembly began during 4.6 and moved to 4.7. Model is tagged. The development evolved as it went along. The dataset has not been manually reviewed. It's entirely Claude developed. Clarification on Reasoning The reasoning is not Claude's actual chain-of-thought (cot) and is not summarized cot. It's a fully synthetic cot created as part of the Assistant response to mimic the type of "thinking"… See the full description on the dataset page: https://huggingface.co/datasets/angrygiraffe/claude-opus-4.6-4.7-reasoning-8.7k.

类型
dataset
许可
apache-2.0
语言
en
下载量
2,177
点赞
439
访问
public
文件
0

标签

  • 指令微调
  • 合成数据
  • 多语言语料
  • 多教师
  • 推理
  • 文本生成

摘要

这是一个由Claude Opus 4.6/4.7生成的8,706条合成思维链(CoT)指令微调数据集,覆盖编码、数学、科学、人文、艺术、金融、法律、医学等28个类别。每条assistant回复均包含150-500字的真实思维推理块,旨在教模型"如何思考"而非仅"说什么",并提供了完整版、指令版、角色扮演版和代码版四个子集,适合用于思维链和推理能力微调。

README

--- license: apache-2.0 task_categories: - text-generation - question-answering language: - en tags: - sft - chain-of-thought - coding - math - roleplay - science - humanities - art - multi-turn - text - json pretty_name: Claude Opus 4.6/4.7 Reasoning Dataset size_categories: - 1K<n<10K --- # 背景 最终…

查看完整页面 · 查看原文