image2text 🚀
Text from images / PDFs? Whether you're parsing lab reports, invoices, or scanned documents. image2text is a modular OCR pipeline
Clarity|Reproducibility|Control
📂 github.com/eddsosa/image2text
#Python #OCR #Tesseract #OpenSource #ScientificSoftware #image2text #PDFtoText
image2text 🚀
Text from images / PDFs? Whether you're parsing lab reports, invoices, or scanned documents. image2text is a modular OCR pipeline
Clarity|Reproducibility|Control
📂 github.com/eddsosa/image2text
#Python #OCR #Tesseract #OpenSource #ScientificSoftware #image2text #PDFtoText
¿Texto desde imágenes o PDFs? Ya sea que analices reportes, facturas o escaneos, image2text es un pipeline OCR modular Claridad | Reproducibilidad | Control
📂 github.com/eddsosa/image2text
#Python #OCR #Tesseract #OpenSource #SoftwareCientífico #image2text #PDFaTexto
¿Texto desde imágenes o PDFs? Ya sea que analices reportes, facturas o escaneos, image2text es un pipeline OCR modular Claridad | Reproducibilidad | Control
📂 github.com/eddsosa/image2text
#Python #OCR #Tesseract #OpenSource #SoftwareCientífico #image2text #PDFaTexto
#AIart #AIartwork #AIimage #AIimagegeneration #AIgenerated #GenerativeAI #digitalart #digitalartwork #promptart #diffusionmodel #ComfyUI #workflow #Node #image2text
#AIart #AIartwork #AIimage #AIimagegeneration #AIgenerated #GenerativeAI #digitalart #digitalartwork #promptart #diffusionmodel #ComfyUI #workflow #Node #image2text
I was wondering recently about that since I wanted to collect some photos for Image2Text benchmarking for Danish and wanted to start a (small scale) crowd sourcing - how to best organize that in terms of consent to share the image/content filtering
I was wondering recently about that since I wanted to collect some photos for Image2Text benchmarking for Danish and wanted to start a (small scale) crowd sourcing - how to best organize that in terms of consent to share the image/content filtering
Like this has to be ready on mac/windows like this year, right? Seems so simple/obvious, am I missing something?
Like this has to be ready on mac/windows like this year, right? Seems so simple/obvious, am I missing something?
ちなみに推論にかかる時間はRTX3070Ti使用でキャプションタスクにおいて3秒くらいっすね。
ちなみに推論にかかる時間はRTX3070Ti使用でキャプションタスクにおいて3秒くらいっすね。
Check the prompt wizard:
bsky.app/profile/youn...
You can also generate your image from the description text.
Try here with $30 credits:
media.nurie.ai/ref?code=9VW...
#MediaSage #PromptWizard #Image2Text #Image2Prompt #NURIEAI #NURIE
Then type keyword and click 'Help to Write'
media.nurie.ai/gen-image
Check what I generated here:
media.nurie.ai/forum
#MediaSage #PromptWizard #NURIEAI #NURIE #Easter
Check the prompt wizard:
bsky.app/profile/youn...
You can also generate your image from the description text.
Try here with $30 credits:
media.nurie.ai/ref?code=9VW...
#MediaSage #PromptWizard #Image2Text #Image2Prompt #NURIEAI #NURIE
A Novel Evaluation Framework for Image2Text Generation
https://arxiv.org/abs/2408.01723
A Novel Evaluation Framework for Image2Text Generation
https://arxiv.org/abs/2408.01723
Por ejemplo Florence2 es un modelo que hace image2text puede describir cualquier imagen y detectar que "yogur en la cara" parece semen, seria facil implementarlo
Por ejemplo Florence2 es un modelo que hace image2text puede describir cualquier imagen y detectar que "yogur en la cara" parece semen, seria facil implementarlo