* [pre-commit.ci] pre-commit autoupdate updates: - [github.com/pre-commit/pre-commit-hooks: v2.3.0 → v4.4.0](https://github.com/pre-commit/pre-commit-hooks/compare/v2.3.0...v4.4.0) - [github.com/astral-sh/ruff-pre-commit: v0.0.280 → v0.0.290](https://github.com/astral-sh/ruff-pre-commit/compare/v0.0.280...v0.0.290) - [github.com/psf/black: 22.12.0 → 23.9.1](https://github.com/psf/black/compare/22.12.0...23.9.1) - [github.com/codespell-project/codespell: v2.2.4 → v2.2.5](https://github.com/codespell-project/codespell/compare/v2.2.4...v2.2.5) * [pre-commit.ci] auto fixes from pre-commit.com hooks for more information, see https://pre-commit.ci --------- Co-authored-by: pre-commit-ci[bot] <66853113+pre-commit-ci[bot]@users.noreply.github.com>
58 lines
2.0 KiB
Markdown
58 lines
2.0 KiB
Markdown
---
|
||
language:
|
||
- zh
|
||
tags:
|
||
- bert
|
||
license: "apache-2.0"
|
||
---
|
||
|
||
# Please use 'Bert' related functions to load this model!
|
||
|
||
## Chinese BERT with Whole Word Masking
|
||
For further accelerating Chinese natural language processing, we provide **Chinese pre-trained BERT with Whole Word Masking**.
|
||
|
||
**[Pre-Training with Whole Word Masking for Chinese BERT](https://arxiv.org/abs/1906.08101)**
|
||
Yiming Cui, Wanxiang Che, Ting Liu, Bing Qin, Ziqing Yang, Shijin Wang, Guoping Hu
|
||
|
||
This repository is developed based on:https://github.com/google-research/bert
|
||
|
||
You may also interested in,
|
||
- Chinese BERT series: https://github.com/ymcui/Chinese-BERT-wwm
|
||
- Chinese MacBERT: https://github.com/ymcui/MacBERT
|
||
- Chinese ELECTRA: https://github.com/ymcui/Chinese-ELECTRA
|
||
- Chinese XLNet: https://github.com/ymcui/Chinese-XLNet
|
||
- Knowledge Distillation Toolkit - TextBrewer: https://github.com/airaria/TextBrewer
|
||
|
||
More resources by HFL: https://github.com/ymcui/HFL-Anthology
|
||
|
||
## Citation
|
||
If you find the technical report or resource is useful, please cite the following technical report in your paper.
|
||
- Primary: https://arxiv.org/abs/2004.13922
|
||
```
|
||
@inproceedings{cui-etal-2020-revisiting,
|
||
title = "Revisiting Pre-Trained Models for {C}hinese Natural Language Processing",
|
||
author = "Cui, Yiming and
|
||
Che, Wanxiang and
|
||
Liu, Ting and
|
||
Qin, Bing and
|
||
Wang, Shijin and
|
||
Hu, Guoping",
|
||
booktitle = "Proceedings of the 2020 Conference on Empirical Methods in Natural Language Processing: Findings",
|
||
month = nov,
|
||
year = "2020",
|
||
address = "Online",
|
||
publisher = "Association for Computational Linguistics",
|
||
url = "https://www.aclweb.org/anthology/2020.findings-emnlp.58",
|
||
pages = "657--668",
|
||
}
|
||
```
|
||
- Secondary: https://arxiv.org/abs/1906.08101
|
||
```
|
||
@article{chinese-bert-wwm,
|
||
title={Pre-Training with Whole Word Masking for Chinese BERT},
|
||
author={Cui, Yiming and Che, Wanxiang and Liu, Ting and Qin, Bing and Yang, Ziqing and Wang, Shijin and Hu, Guoping},
|
||
journal={arXiv preprint arXiv:1906.08101},
|
||
year={2019}
|
||
}
|
||
```
|