|
Download README.md from toktik-pgx/markuplm-large: direct link, hf CLI and curl.
- Browser
- Download file 887 Bytes
-
https://huggingface.co/toktik-pgx/markuplm-large/resolve/main/README.md
- Command line
-
hf download hf://toktik-pgx/markuplm-large/README.md
-
curl -L -o README.md https://huggingface.co/toktik-pgx/markuplm-large/resolve/main/README.md
887 Bytes
| language: | |
| - en | |
| # MarkupLM | |
| **Multimodal (text +markup language) pre-training for [Document AI](https://www.microsoft.com/en-us/research/project/document-ai/)** | |
| ## Introduction | |
| MarkupLM is a simple but effective multi-modal pre-training method of text and markup language for visually-rich document understanding and information extraction tasks, such as webpage QA and webpage information extraction. MarkupLM archives the SOTA results on multiple datasets. For more details, please refer to our paper: | |
| [MarkupLM: Pre-training of Text and Markup Language for Visually-rich Document Understanding](https://arxiv.org/abs/2110.08518) Junlong Li, Yiheng Xu, Lei Cui, Furu Wei | |
| ## Usage | |
| We refer to the [docs](https://huggingface.co/docs/transformers/main/en/model_doc/markuplm) and [demo notebooks](https://github.com/NielsRogge/Transformers-Tutorials/tree/master/MarkupLM). |