Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

voxreality
/
rgb_language_vqa

Image-to-Text
Transformers
PyTorch
English
blip
visual-question-answering
text-generation-inference
Model card Files Files and versions
xet
Community
1

Instructions to use voxreality/rgb_language_vqa with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • Transformers

    How to use voxreality/rgb_language_vqa with Transformers:

    # Use a pipeline as a high-level helper
    # Warning: Pipeline type "image-to-text" is no longer supported in transformers v5.
    # You must load the model directly (see below) or downgrade to v4.x with:
    # 'pip install "transformers<5.0.0'
    from transformers import pipeline
    
    pipe = pipeline("image-to-text", model="voxreality/rgb_language_vqa")
    # Load model directly
    from transformers import AutoProcessor, AutoModelForVisualQuestionAnswering
    
    processor = AutoProcessor.from_pretrained("voxreality/rgb_language_vqa")
    model = AutoModelForVisualQuestionAnswering.from_pretrained("voxreality/rgb_language_vqa", device_map="auto")
  • Notebooks
  • Google Colab
  • Kaggle

You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

Gated model
You can list files but not access them

Preview of files found in this repository
  • .gitattributes
    1.52 kB
    initial commit almost 2 years ago
  • README.md
    2.86 kB
    update readme almost 2 years ago
  • config.json
    638 Bytes
    Upload 12 files almost 2 years ago
  • generation_config.json
    136 Bytes
    Upload 12 files almost 2 years ago
  • markdown_apache-2.0.md
    12.6 kB
    Upload 12 files almost 2 years ago
  • optimizer_state.pt
    3.08 GB
    xet
    Upload 12 files almost 2 years ago
  • preprocessor_config.json
    471 Bytes
    Upload 12 files almost 2 years ago
  • pytorch_model.bin
    1.54 GB
    xet
    Upload 12 files almost 2 years ago
  • scheduler_state.pt
    575 Bytes
    xet
    Upload 12 files almost 2 years ago
  • special_tokens_map.json
    125 Bytes
    Upload 12 files almost 2 years ago
  • tokenizer.json
    711 kB
    Upload 12 files almost 2 years ago
  • tokenizer_config.json
    1.35 kB
    Upload 12 files almost 2 years ago
  • vocab.txt
    232 kB
    Upload 12 files almost 2 years ago