Salesforce/blip-image-captioning-large

bsd-3-clause

BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...

image to textBy Salesforce

Salesforce/blip-image-captioning-base

bsd-3-clause

BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...

image to textBy Salesforce

microsoft/trocr-base-handwritten

mit

TrOCR (base-sized model, fine-tuned on IAM) TrOCR model fine-tuned on the IAM dataset( It was introduced in the paper TrOCR: Transformer-bas...

image to textBy microsoft

jinhybr/OCR-Donut-CORD

mit

Donut (base-sized model, fine-tuned on CORD) Donut model fine-tuned on CORD. It was introduced in the paper OCR-free Document Understanding ...

image to textBy jinhybr

microsoft/trocr-base-printed

TrOCR (base-sized model, fine-tuned on SROIE) TrOCR model fine-tuned on the SROIE dataset( It was introduced in the paper TrOCR: Transformer...

image to textBy microsoft

microsoft/trocr-large-printed

TrOCR (large-sized model, fine-tuned on SROIE) TrOCR model fine-tuned on the SROIE dataset( It was introduced in the paper TrOCR: Transforme...

image to textBy microsoft

kha-white/manga-ocr-base

apache-2.0

Manga OCR Optical character recognition for Japanese text, with the main focus being Japanese manga. It uses Vision Encoder Decoder( framewo...

image to textBy kha-white

Salesforce/blip-image-captioning-large

bsd-3-clause

BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...

image to textBy Salesforce

nlpconnect/vit-gpt2-image-captioning

apache-2.0

nlpconnect/vit-gpt2-image-captioning This is an image captioning model trained by @ydshieh in flax ( this is pytorch version of this( The Il...

image to textBy nlpconnect

Salesforce/blip-image-captioning-base

bsd-3-clause

BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...

image to textBy Salesforce
Showing 10 of 15