Metadata-Version: 2.1
Name: WeTextProcessing
Version: 0.0.1
Summary: WeTextProcessing, including TN & ITN
Home-page: https://github.com/wenet-e2e/WeTextProcessing
Author: Zhendong Peng, Xingchen Song
Author-email: pzd17@tsinghua.org.cn, sxc19@tsinghua.org.cn
License: UNKNOWN
Platform: UNKNOWN
Classifier: Programming Language :: Python :: 3
Classifier: Operating System :: OS Independent
Classifier: Topic :: Scientific/Engineering :: Artificial Intelligence
Description-Content-Type: text/markdown
License-File: LICENSE
Requires-Dist: pynini
Requires-Dist: importlib-resources

## Text Normalization & Inverse Text Normalization

### 1. How To Use

``` bash
$ git clone https://github.com/wenet-e2e/WeTextProcessing.git
$ cd WeTextProcessing
$ python normalize.py --text "text to be normalized"
$ python inverse_normalize.py --text "text to be denormalized"
```

### 2. TN Pipeline

Please refer to [TN.README](tn/README.md)

### 3. ITN Pipeline (Coming soon...)

Please refer to [ITN.README](itn/README.md)

## Acknowledge

1. Thank the authors of foundational libraries like [OpenFst](https://www.openfst.org/twiki/bin/view/FST/WebHome) & [Pynini](https://www.openfst.org/twiki/bin/view/GRM/Pynini).
3. Thank [NeMo](https://github.com/NVIDIA/NeMo) team & NeMo open-source community.
2. Thank [Zhenxiang Ma](https://github.com/mzxcpp), [Jiayu Du](https://github.com/dophist), and [SpeechColab](https://github.com/SpeechColab) organization.
3. Referred [Pynini](https://github.com/kylebgorman/pynini) for reading the FAR, and printing the shortest path of a lattice in the C++ runtime.
4. Referred [TN of NeMo](https://github.com/NVIDIA/NeMo/tree/main/nemo_text_processing/text_normalization/zh) for the data to build the tagger graph.
5. Referred [ITN of chinese_text_normalization](https://github.com/speechio/chinese_text_normalization/tree/master/thrax/src/cn) for the data to build the tagger graph.


