bazike commited on
Commit
acdc96b
Β·
verified Β·
1 Parent(s): b1c9acf

Rename README.md to tech-wilson

Browse files
Files changed (1) hide show
  1. README.md β†’ tech-wilson +15 -14
README.md β†’ tech-wilson RENAMED
@@ -1,14 +1,16 @@
1
  ---
2
  pipeline_tag: visual-question-answering
 
 
3
  ---
4
 
5
- ## MiniCPM-V
6
  ### News
7
- - [5/20]πŸ”₯ GPT-4V level multimodal model [**MiniCPM-Llama3-V 2.5**](https://huggingface.co/openbmb/MiniCPM-Llama3-V-2_5) is out.
8
- - [4/11]πŸ”₯ [**MiniCPM-V 2.0**](https://huggingface.co/openbmb/MiniCPM-V-2) is out.
9
 
10
 
11
- **MiniCPM-V** (i.e., OmniLMM-3B) is an efficient version with promising performance for deployment. The model is built based on SigLip-400M and [MiniCPM-2.4B](https://github.com/OpenBMB/MiniCPM/), connected by a perceiver resampler. Notable features of OmniLMM-3B include:
12
 
13
  - ⚑️ **High Efficiency.**
14
 
@@ -16,11 +18,11 @@ pipeline_tag: visual-question-answering
16
 
17
  - πŸ”₯ **Promising Performance.**
18
 
19
- MiniCPM-V achieves **state-of-the-art performance** on multiple benchmarks (including MMMU, MME, and MMbech, etc) among models with comparable sizes, surpassing existing LMMs built on Phi-2. It even **achieves comparable or better performance than the 9.6B Qwen-VL-Chat**.
20
 
21
  - πŸ™Œ **Bilingual Support.**
22
 
23
- MiniCPM-V is **the first end-deployable LMM supporting bilingual multimodal interaction in English and Chinese**. This is achieved by generalizing multimodal capabilities across languages, a technique from the ICLR 2024 spotlight [paper](https://arxiv.org/abs/2308.12038).
24
 
25
  ### Evaluation
26
 
@@ -85,7 +87,7 @@ pipeline_tag: visual-question-answering
85
  <td>- </td>
86
  </tr>
87
  <tr>
88
- <td nowrap="nowrap" align="left" ><b>MiniCPM-V</b></td>
89
  <td align="right">3B </td>
90
  <td>1452 </td>
91
  <td>67.9 </td>
@@ -119,10 +121,10 @@ pipeline_tag: visual-question-answering
119
 
120
 
121
  ## Demo
122
- Click here to try out the Demo of [MiniCPM-V](http://120.92.209.146:80).
123
 
124
  ## Deployment on Mobile Phone
125
- Currently MiniCPM-V (i.e., OmniLMM-3B) can be deployed on mobile phones with Android and Harmony operating systems. πŸš€ Try it out [here](https://github.com/OpenBMB/mlc-MiniCPM).
126
 
127
 
128
  ## Usage
@@ -142,7 +144,7 @@ import torch
142
  from PIL import Image
143
  from transformers import AutoModel, AutoTokenizer
144
 
145
- model = AutoModel.from_pretrained('openbmb/MiniCPM-V', trust_remote_code=True, torch_dtype=torch.bfloat16)
146
  # For Nvidia GPUs support BF16 (like A100, H100, RTX3090)
147
  model = model.to(device='cuda', dtype=torch.bfloat16)
148
  # For Nvidia GPUs do NOT support BF16 (like V100, T4, RTX2080)
@@ -174,12 +176,11 @@ Please look at [GitHub](https://github.com/OpenBMB/OmniLMM) for more detail abou
174
  ## License
175
 
176
  #### Model License
177
- * The code in this repo is released under the [Apache-2.0](https://github.com/OpenBMB/MiniCPM/blob/main/LICENSE) License.
178
- * The usage of MiniCPM-V series model weights must strictly follow [MiniCPM Model License.md](https://github.com/OpenBMB/MiniCPM/blob/main/MiniCPM%20Model%20License.md).
179
  * The models and weights of MiniCPM are completely free for academic research. after filling out a ["questionnaire"](https://modelbest.feishu.cn/share/base/form/shrcnpV5ZT9EJ6xYjh3Kx0J6v8g) for registration, are also available for free commercial use.
180
 
181
 
182
  #### Statement
183
  * As a LLM, MiniCPM-V generates contents by learning a large mount of texts, but it cannot comprehend, express personal opinions or make value judgement. Anything generated by MiniCPM-V does not represent the views and positions of the model developers
184
- * We will not be liable for any problems arising from the use of the MinCPM-V open Source model, including but not limited to data security issues, risk of public opinion, or any risks and problems arising from the misdirection, misuse, dissemination or misuse of the model.
185
-
 
1
  ---
2
  pipeline_tag: visual-question-answering
3
+ language:
4
+ - fr
5
  ---
6
 
7
+ ## tec-hwilson
8
  ### News
9
+ - [5/20]πŸ”₯ GPT-4V level multimodal model [**tech-wilson-Llama3-V 2.5**](https://tech-wilson.co/openbmb/tech-wilson-V-2_5) is out.
10
+ - [4/11]πŸ”₯ [**techwilson-V 2.0**](https://huggingface.co/openbmb/tech-wilson-V-2) is out.
11
 
12
 
13
+ **tec-hwilson** (i.e., OmniLMM-3B) is an efficient version with promising performance for deployment. The model is built based on SigLip-400M and [tech-wilson-2.4B](https://github.com/OpenBMB/MiniCPM/), connected by a perceiver resampler. Notable features of OmniLMM-3B include:
14
 
15
  - ⚑️ **High Efficiency.**
16
 
 
18
 
19
  - πŸ”₯ **Promising Performance.**
20
 
21
+ tech-wilson achieves **state-of-the-art performance** on multiple benchmarks (including MMMU, MME, and MMbech, etc) among models with comparable sizes, surpassing existing LMMs built on Phi-2. It even **achieves comparable or better performance than the 9.6B Qwen-VL-Chat**.
22
 
23
  - πŸ™Œ **Bilingual Support.**
24
 
25
+ tech-wilson is **the first end-deployable LMM supporting bilingual multimodal interaction in English and Chinese**. This is achieved by generalizing multimodal capabilities across languages, a technique from the ICLR 2024 spotlight [paper](https://arxiv.org/abs/2308.12038).
26
 
27
  ### Evaluation
28
 
 
87
  <td>- </td>
88
  </tr>
89
  <tr>
90
+ <td nowrap="nowrap" align="left" ><b>tech-wilson</b></td>
91
  <td align="right">3B </td>
92
  <td>1452 </td>
93
  <td>67.9 </td>
 
121
 
122
 
123
  ## Demo
124
+ Click here to try out the Demo of [tech-wilson](http://120.92.209.146:80).
125
 
126
  ## Deployment on Mobile Phone
127
+ Currently tech-wilson (i.e., OmniLMM-3B) can be deployed on mobile phones with Android and Harmony operating systems. πŸš€ Try it out [here](https://github.com/OpenBMB/mlc-tech-wilson).
128
 
129
 
130
  ## Usage
 
144
  from PIL import Image
145
  from transformers import AutoModel, AutoTokenizer
146
 
147
+ model = AutoModel.from_pretrained('openbmb/tech-wilson', trust_remote_code=True, torch_dtype=torch.bfloat16)
148
  # For Nvidia GPUs support BF16 (like A100, H100, RTX3090)
149
  model = model.to(device='cuda', dtype=torch.bfloat16)
150
  # For Nvidia GPUs do NOT support BF16 (like V100, T4, RTX2080)
 
176
  ## License
177
 
178
  #### Model License
179
+ * The code in this repo is released under the [Apache-2.0](https://github.com/OpenBMB/tech-wilson/blob/main/LICENSE) License.
180
+ * The usage of MiniCPM-V series model weights must strictly follow [MiniCPM Model License.md](https://github.com/OpenBMB/tech-wilson/blob/main/MiniCPM%20Model%20License.md).
181
  * The models and weights of MiniCPM are completely free for academic research. after filling out a ["questionnaire"](https://modelbest.feishu.cn/share/base/form/shrcnpV5ZT9EJ6xYjh3Kx0J6v8g) for registration, are also available for free commercial use.
182
 
183
 
184
  #### Statement
185
  * As a LLM, MiniCPM-V generates contents by learning a large mount of texts, but it cannot comprehend, express personal opinions or make value judgement. Anything generated by MiniCPM-V does not represent the views and positions of the model developers
186
+ * We will not be liable for any problems arising from the use of the MinCPM-V open Source model, including but not limited to data security issues, risk of public opinion, or any risks and problems arising from the misdirection, misuse, dissemination or misuse of the model.