Feature Extraction
Transformers
Safetensors
English
Korean
multilingual
qwen3_vl
vision-language
embedding
multimodal-embedding
mmeb
digital-forensics
custom_code
Instructions to use Urock-AI/Eddy-vl_embedding_1.9B_v1 with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use Urock-AI/Eddy-vl_embedding_1.9B_v1 with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("feature-extraction", model="Urock-AI/Eddy-vl_embedding_1.9B_v1", trust_remote_code=True)# pip install -U transformers accelerate # Load model directly from transformers import AutoProcessor, AutoModel processor = AutoProcessor.from_pretrained("Urock-AI/Eddy-vl_embedding_1.9B_v1", trust_remote_code=True) model = AutoModel.from_pretrained("Urock-AI/Eddy-vl_embedding_1.9B_v1", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
Upload README.md with huggingface_hub
Browse files
README.md
CHANGED
|
@@ -10,6 +10,7 @@ tags:
|
|
| 10 |
- multimodal-embedding
|
| 11 |
- mmeb
|
| 12 |
- digital-forensics
|
|
|
|
| 13 |
library_name: transformers
|
| 14 |
pipeline_tag: feature-extraction
|
| 15 |
base_model:
|
|
@@ -24,6 +25,8 @@ base_model:
|
|
| 24 |
|
| 25 |
[Urock-AI](https://huggingface.co/Urock-AI) · [urock.kr](https://urock.kr/) · License: Apache 2.0
|
| 26 |
|
|
|
|
|
|
|
| 27 |
> **Eddy-VL is a multimodal embedding model light enough to run on edge devices.** It keeps the retrieval quality of a 2B-class vision-language embedder in a lighter, faster package — built at Urock-AI Lab for real-world multimodal search over images, video, and documents.
|
| 28 |
|
| 29 |
---
|
|
@@ -228,18 +231,24 @@ Curated by Urock-AI Lab from public sources — including **MS COCO**, **ko-coco
|
|
| 228 |
|
| 229 |
## Citation & license
|
| 230 |
|
| 231 |
-
Eddy-VL is derived from **[Qwen3-VL-Embedding-2B](https://huggingface.co/Qwen/Qwen3-VL-Embedding-2B)**; please honor its license terms (Apache 2.0) and cite
|
|
|
|
|
|
|
| 232 |
|
| 233 |
```bibtex
|
| 234 |
-
@misc{
|
| 235 |
-
title
|
| 236 |
-
author
|
| 237 |
-
year
|
| 238 |
-
|
| 239 |
-
|
|
|
|
|
|
|
| 240 |
}
|
| 241 |
```
|
| 242 |
|
|
|
|
|
|
|
| 243 |
---
|
| 244 |
|
| 245 |
**Urock-AI Lab** — Digital Forensic AI · [huggingface.co/Urock-AI](https://huggingface.co/Urock-AI) · [urock.kr](https://urock.kr/)
|
|
|
|
| 10 |
- multimodal-embedding
|
| 11 |
- mmeb
|
| 12 |
- digital-forensics
|
| 13 |
+
- arxiv:2607.16316
|
| 14 |
library_name: transformers
|
| 15 |
pipeline_tag: feature-extraction
|
| 16 |
base_model:
|
|
|
|
| 25 |
|
| 26 |
[Urock-AI](https://huggingface.co/Urock-AI) · [urock.kr](https://urock.kr/) · License: Apache 2.0
|
| 27 |
|
| 28 |
+
**Technical report:** [arXiv:2607.16316](https://arxiv.org/abs/2607.16316) · [PDF](https://arxiv.org/pdf/2607.16316)
|
| 29 |
+
|
| 30 |
> **Eddy-VL is a multimodal embedding model light enough to run on edge devices.** It keeps the retrieval quality of a 2B-class vision-language embedder in a lighter, faster package — built at Urock-AI Lab for real-world multimodal search over images, video, and documents.
|
| 31 |
|
| 32 |
---
|
|
|
|
| 231 |
|
| 232 |
## Citation & license
|
| 233 |
|
| 234 |
+
Eddy-VL is derived from **[Qwen3-VL-Embedding-2B](https://huggingface.co/Qwen/Qwen3-VL-Embedding-2B)**; please honor its license terms (Apache 2.0) and cite the technical report when you use this model. Evaluation uses the MMEB benchmark (Jiang et al., *VLM2Vec*, ICLR 2025).
|
| 235 |
+
|
| 236 |
+
If you find Eddy-VL useful, please cite:
|
| 237 |
|
| 238 |
```bibtex
|
| 239 |
+
@misc{cho2026eddyvl,
|
| 240 |
+
title = {Eddy-VL 1.9B: Structural Pruning and Layered Distillation for Edge-Deployable Multimodal Embedding},
|
| 241 |
+
author = {Cho, HanYeong and Kim, Changwoo and Chu, Taeuk and Park, Jimin},
|
| 242 |
+
year = {2026},
|
| 243 |
+
eprint = {2607.16316},
|
| 244 |
+
archivePrefix = {arXiv},
|
| 245 |
+
primaryClass = {cs.CV},
|
| 246 |
+
url = {https://arxiv.org/abs/2607.16316}
|
| 247 |
}
|
| 248 |
```
|
| 249 |
|
| 250 |
+
Model weights and inference code: [Urock-AI/Eddy-vl_embedding_1.9B_v1](https://huggingface.co/Urock-AI/Eddy-vl_embedding_1.9B_v1).
|
| 251 |
+
|
| 252 |
---
|
| 253 |
|
| 254 |
**Urock-AI Lab** — Digital Forensic AI · [huggingface.co/Urock-AI](https://huggingface.co/Urock-AI) · [urock.kr](https://urock.kr/)
|