29 lines
1.3 KiB
Markdown
29 lines
1.3 KiB
Markdown
|
# EmoLLM's datasets
|
||
|
|
||
|
* Category of dataset: **General** and **Role-play**
|
||
|
* Type of data: **QA** and **Conversation**
|
||
|
* Summary: General(**6 datasets**), Role-play(**3 datasets**)
|
||
|
|
||
|
## Category
|
||
|
* **General**: generic dataset, including psychological Knowledge, counseling technology, etc.
|
||
|
* **Role-play**: role-playing dataset, including character-specific conversation style data, etc.
|
||
|
|
||
|
## Type
|
||
|
* **QA**: question-and-answer pair
|
||
|
* **Conversation**: multi-turn consultation dialogue
|
||
|
|
||
|
## Summary
|
||
|
|
||
|
| Category | Dataset | Type | Total |
|
||
|
| :---------: | :-------------------: | :----------: | :-----: |
|
||
|
| *General* | data | Conversation | 5600+ |
|
||
|
| *General* | data_pro | Conversation | 36500+ |
|
||
|
| *General* | multi_turn_dataset_1 | Conversation | 36,000+ |
|
||
|
| *General* | multi_turn_dataset_2 | Conversation | 27,000+ |
|
||
|
| *General* | single_turn_dataset_1 | QA | 14000+ |
|
||
|
| *General* | single_turn_dataset_2 | QA | 18300+ |
|
||
|
| *Role-play* | aiwei | Conversation | 4000+ |
|
||
|
| *Role-play* | SoulStar | QA | 11200+ |
|
||
|
| *Role-play* | tiangou | Conversation | 3900+ |
|
||
|
| …… | …… | …… | …… |
|