000491f1be
* Add files via upload * 新增ENmd文档 * Update README.md * Update README_EN.md * Update LICENSE * [docs] update lmdeploy file * add ocr.md * Update tutorial.md * Update tutorial_EN.md * Update General_evaluation_EN.md * Update General_evaluation_EN.md * Update README.md Add InternLM2_7B_chat_full's professional evaluation results * Update Professional_evaluation.md * Update Professional_evaluation.md * Update Professional_evaluation.md * Update Professional_evaluation.md * Update Professional_evaluation_EN.md * Update README.md * Update README.md * Update README_EN.md * Update README_EN.md * Update README_EN.md * [DOC] update readme * Update LICENSE * Update LICENSE * update personal info and small format optimizations * update personal info and translations for contents in a table * Update RAG README * Update demo link in README.md * Update xlab app link * Update xlab link * add xlab model * Update web_demo-aiwei.py * add bitex --------- Co-authored-by: xzw <62385492+aJupyter@users.noreply.github.com> Co-authored-by: এ許我辞忧࿐♡ <127636623+Smiling-Weeping-zhr@users.noreply.github.com> Co-authored-by: Vicky <vicky_3021@163.com> Co-authored-by: MING_X <119648793+MING-ZCH@users.noreply.github.com> Co-authored-by: Nobody-ML <1755309985@qq.com> Co-authored-by: 8baby8 <3345710651@qq.com> Co-authored-by: chaoke <101492509+8baby8@users.noreply.github.com> Co-authored-by: aJupyter <ajupyter@163.com> Co-authored-by: HongCheng <kwchenghong@gmail.com> Co-authored-by: santiagoTOP <“1537211712top@gmail.com”>
44 lines
1.8 KiB
Markdown
44 lines
1.8 KiB
Markdown
# EmoLLM's datasets
|
||
|
||
* Category of dataset: **General** and **Role-play**
|
||
* Type of data: **QA** and **Conversation**
|
||
* Summary: General(**6 datasets**), Role-play(**3 datasets**)
|
||
|
||
## Category
|
||
* **General**: generic dataset, including psychological Knowledge, counseling technology, etc.
|
||
* **Role-play**: role-playing dataset, including character-specific conversation style data, etc.
|
||
|
||
## Type
|
||
* **QA**: question-and-answer pair
|
||
* **Conversation**: multi-turn consultation dialogue
|
||
|
||
## Summary
|
||
|
||
| Category | Dataset | Type | Total |
|
||
| :---------: | :-------------------: | :----------: | :-----: |
|
||
| *General* | data | Conversation | 5600+ |
|
||
| *General* | data_pro | Conversation | 36500+ |
|
||
| *General* | multi_turn_dataset_1 | Conversation | 36,000+ |
|
||
| *General* | multi_turn_dataset_2 | Conversation | 27,000+ |
|
||
| *General* | single_turn_dataset_1 | QA | 14000+ |
|
||
| *General* | single_turn_dataset_2 | QA | 18300+ |
|
||
| *Role-play* | aiwei | Conversation | 4000+ |
|
||
| *Role-play* | SoulStar | QA | 11200+ |
|
||
| *Role-play* | tiangou | Conversation | 3900+ |
|
||
| …… | …… | …… | …… |
|
||
|
||
|
||
## Source
|
||
**General**:
|
||
* dataset `data` from this repo
|
||
* dataset `data_pro` from this repo
|
||
* dataset `multi_turn_dataset_1` from [Smile](https://github.com/qiuhuachuan/smile)
|
||
* dataset `multi_turn_dataset_2` from [CPsyCounD](https://github.com/CAS-SIAT-XinHai/CPsyCoun)
|
||
* dataset `single_turn_dataset_1` from this repo
|
||
* dataset `single_turn_dataset_2` from this repo
|
||
|
||
**Role-play**:
|
||
* dataset `aiwei` from this repo
|
||
* dataset `tiangou` from this repo
|
||
* dataset `SoulStar` from [SoulStar](https://github.com/Nobody-ML/SoulStar)
|