Update README.md
Browse files
README.md
CHANGED
|
@@ -15,6 +15,14 @@ license: apache-2.0
|
|
| 15 |
|
| 16 |
The model only takes images as document-side inputs and produce vectors representing document pages. `minicpm-visual-embedding-v0` is trained with over 200k query-visual document pairs, including textual document, visual document, arxiv figures, plots, charts, industry documents, textbooks, ebooks, and openly-available PDFs, etc. The performance of `minicpm-visual-embedding-v0` is on a par with our ablation text embedding model on text-oriented documents, and an advantages on visually-intensive documents.
|
| 17 |
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 18 |

|
| 19 |
|
| 20 |
# News
|
|
|
|
| 15 |
|
| 16 |
The model only takes images as document-side inputs and produce vectors representing document pages. `minicpm-visual-embedding-v0` is trained with over 200k query-visual document pairs, including textual document, visual document, arxiv figures, plots, charts, industry documents, textbooks, ebooks, and openly-available PDFs, etc. The performance of `minicpm-visual-embedding-v0` is on a par with our ablation text embedding model on text-oriented documents, and an advantages on visually-intensive documents.
|
| 17 |
|
| 18 |
+
Our model is capable of:
|
| 19 |
+
|
| 20 |
+
- Help you read a long visually-intensive or text-oriented PDF document and find the pages that answer your question.
|
| 21 |
+
|
| 22 |
+
- Help you build a personal library and retireve book pages from a large collection of books.
|
| 23 |
+
|
| 24 |
+
- It works like human: read and comprehend with **vision** and remember **multimodal** information in hippocampus.
|
| 25 |
+
|
| 26 |

|
| 27 |
|
| 28 |
# News
|