한국어 NLP 데이터셋 · 텍스트
⚖
LLMCoT추론
HAE-RAE CoT 1.5M
한국어 NLP 데이터셋·Korean NLP Dataset
한국어 연쇄 사고(Chain-of-Thought) 학습용 대규모 코퍼스입니다. HF 카드 기준 1,586,688행입니다.
Source VerifiedLicense Needs review
Source
Hugging Face
License
CC BY 4.0
Catalog hint — re-check original
Language
Korean
Needs review
Format
Text
Needs review
Size
1.05GB
1,586,688 샘플
Year
2024
이 데이터셋이 맞는 경우
HAE-RAE CoT 1.5M은 Hugging Face에서 제공하는 한국어 NLP 데이터셋입니다. 태그: LLM, CoT, 추론. 라이선스 표기: CC BY 4.0. 카탈로그 기준으로는 상업 이용 가능으로 표시되어 있습니다. GearDel은 탐색 카탈로그이며 파일을 호스팅하지 않습니다.
- 과제 / 카테고리
- 텍스트 · 텍스트 · LLM, CoT, 추론
- Language
- Korean
- Format · Size
- Text · 1.05GB (1,586,688 샘플)
- License
- CC BY 4.0(Catalog hint — re-check original) · License document
- Source & original link
- Hugging Face · Original dataset link(Verified)
- Verification status
- Source: Verified · License: Needs review · Last checked: 2026-09-03
HAE-RAE CoT 1.5M: Korean NLP Dataset from Hugging Face. License CC BY 4.0. Open the original link to download.
비슷한 데이터셋
같은 유형·태그·제공처를 공유하는 카탈로그 항목입니다.
No samples in catalog
Check the original provider page for examples.
Trust & metadata
GearDel does not host datasets. Status below is catalog metadata and may differ from the original provider.
- Source
- Hugging FaceVerified
- Source URL
- https://huggingface.co/datasets/HAERAE-HUB/HAE-RAE-COT-1.5MVerified
- License
- CC BY 4.0Needs review
- Verification
- License needs review
- Language
- KoreanNeeds review
- Format
- TextNeeds review
- Size
- 1.05GBNeeds review
- Last checked
- 2026-09-03
Always re-check license and terms on the original page. Standard license document