zsxkib / uform-gen

🖼️ Super fast 1.5B Image Captioning/VQA Multimodal LLM (Image-to-Text) 🖋️

  • Public
  • 2.1K runs
  • GitHub
  • License