Image-to-Text
Transformers
PyTorch
textract
feature-extraction
ocr
vision-language
qwen2-vl
custom-model
text-extraction
document-ai
high-accuracy
custom_code
Instructions to use BabaK07/textract-ai with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use BabaK07/textract-ai with Transformers:
# Use a pipeline as a high-level helper # Warning: Pipeline type "image-to-text" is no longer supported in transformers v5. # You must load the model directly (see below) or downgrade to v4.x with: # pip install "transformers<5.0.0" from transformers import pipeline pipe = pipeline("image-to-text", model="BabaK07/textract-ai", trust_remote_code=True)# pip install -U transformers accelerate # Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("BabaK07/textract-ai", trust_remote_code=True, device_map="auto") - Notebooks
- Google Colab
- Kaggle
Download chat_template.jinja from BabaK07/textract-ai: direct link, hf CLI and curl.
- Browser
- Download file 1.02 kB
-
https://huggingface.co/BabaK07/textract-ai/resolve/main/chat_template.jinja
- Command line
-
hf download hf://BabaK07/textract-ai/chat_template.jinja
-
curl -L -o chat_template.jinja https://huggingface.co/BabaK07/textract-ai/resolve/main/chat_template.jinja
1.02 kB
| {% set image_count = namespace(value=0) %}{% set video_count = namespace(value=0) %}{% for message in messages %}{% if loop.first and message['role'] != 'system' %}<|im_start|>system | |
| You are a helpful assistant.<|im_end|> | |
| {% endif %}<|im_start|>{{ message['role'] }} | |
| {% if message['content'] is string %}{{ message['content'] }}<|im_end|> | |
| {% else %}{% for content in message['content'] %}{% if content['type'] == 'image' or 'image' in content or 'image_url' in content %}{% set image_count.value = image_count.value + 1 %}{% if add_vision_id %}Picture {{ image_count.value }}: {% endif %}<|vision_start|><|image_pad|><|vision_end|>{% elif content['type'] == 'video' or 'video' in content %}{% set video_count.value = video_count.value + 1 %}{% if add_vision_id %}Video {{ video_count.value }}: {% endif %}<|vision_start|><|video_pad|><|vision_end|>{% elif 'text' in content %}{{ content['text'] }}{% endif %}{% endfor %}<|im_end|> | |
| {% endif %}{% endfor %}{% if add_generation_prompt %}<|im_start|>assistant | |
| {% endif %} |