Token Classification
Transformers
PyTorch
TensorBoard
Safetensors
xlm-roberta
punctuation prediction
punctuation
Instructions to use oliverguhr/fullstop-punctuation-multilingual-sonar-base with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use oliverguhr/fullstop-punctuation-multilingual-sonar-base with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("token-classification", model="oliverguhr/fullstop-punctuation-multilingual-sonar-base")# Load model directly from transformers import AutoTokenizer, AutoModelForTokenClassification tokenizer = AutoTokenizer.from_pretrained("oliverguhr/fullstop-punctuation-multilingual-sonar-base") model = AutoModelForTokenClassification.from_pretrained("oliverguhr/fullstop-punctuation-multilingual-sonar-base", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Commit ·
f72bf70
1
Parent(s): 0a7f17f
citation added
Browse files
README.md
CHANGED
|
@@ -115,4 +115,34 @@ model = PunctuationModel(model = "oliverguhr/fullstop-dutch-punctuation-predicti
|
|
| 115 |
```
|
| 116 |
|
| 117 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 118 |
|
|
|
|
| 115 |
```
|
| 116 |
|
| 117 |
|
| 118 |
+
## How to cite us
|
| 119 |
+
|
| 120 |
+
```
|
| 121 |
+
@article{guhr-EtAl:2021:fullstop,
|
| 122 |
+
title={FullStop: Multilingual Deep Models for Punctuation Prediction},
|
| 123 |
+
author = {Guhr, Oliver and Schumann, Anne-Kathrin and Bahrmann, Frank and Böhme, Hans Joachim},
|
| 124 |
+
booktitle = {Proceedings of the Swiss Text Analytics Conference 2021},
|
| 125 |
+
month = {June},
|
| 126 |
+
year = {2021},
|
| 127 |
+
address = {Winterthur, Switzerland},
|
| 128 |
+
publisher = {CEUR Workshop Proceedings},
|
| 129 |
+
url = {http://ceur-ws.org/Vol-2957/sepp_paper4.pdf}
|
| 130 |
+
}
|
| 131 |
+
|
| 132 |
+
```
|
| 133 |
+
|
| 134 |
+
```
|
| 135 |
+
@misc{https://doi.org/10.48550/arxiv.2301.03319,
|
| 136 |
+
doi = {10.48550/ARXIV.2301.03319},
|
| 137 |
+
url = {https://arxiv.org/abs/2301.03319},
|
| 138 |
+
author = {Vandeghinste, Vincent and Guhr, Oliver},
|
| 139 |
+
keywords = {Computation and Language (cs.CL), Artificial Intelligence (cs.AI), FOS: Computer and information sciences, FOS: Computer and information sciences, I.2.7},
|
| 140 |
+
title = {FullStop:Punctuation and Segmentation Prediction for Dutch with Transformers},
|
| 141 |
+
publisher = {arXiv},
|
| 142 |
+
year = {2023},
|
| 143 |
+
copyright = {Creative Commons Attribution Share Alike 4.0 International}
|
| 144 |
+
}
|
| 145 |
+
|
| 146 |
+
```
|
| 147 |
+
|
| 148 |
|