Byte Goose AIBeta
Salesforce/blip-image-captioning-large
bsd-3-clauseBLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation Model card for image captioning pre...
image to textBy Salesforce
nlpconnect/vit-gpt2-image-captioning
apache-2.0nlpconnect/vit-gpt2-image-captioning This is an image captioning model trained by @ydshieh in flax ( this is pytorch version of this( The Il...
image to textBy nlpconnect
naver-clova-ix/donut-base
mitDonut (base-sized model, pre-trained only) Donut model pre-trained-only. It was introduced in the paper OCR-free Document Understanding Tran...
image to textBy naver-clova-ix
blip-image-captioning-base
bsd-3-clauseOfficial image-to-text model by Salesforce.
image to textBy Salesforce
NuExtract3
apache-2.0Official image-to-text model by numind.
image to textBy numind
Showing 5 of 15