Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

voxreality
/
rgb_language_cap

Image-to-Text
Transformers
PyTorch
English
vision-encoder-decoder
image-text-to-text
text-generation-inference
Model card Files Files and versions
xet
Community
1

Instructions to use voxreality/rgb_language_cap with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • Transformers

    How to use voxreality/rgb_language_cap with Transformers:

    # Use a pipeline as a high-level helper
    # Warning: Pipeline type "image-to-text" is no longer supported in transformers v5.
    # You must load the model directly (see below) or downgrade to v4.x with:
    # 'pip install "transformers<5.0.0'
    from transformers import pipeline
    
    pipe = pipeline("image-to-text", model="voxreality/rgb_language_cap")
    # Load model directly
    from transformers import AutoTokenizer, AutoModelForMultimodalLM
    
    tokenizer = AutoTokenizer.from_pretrained("voxreality/rgb_language_cap")
    model = AutoModelForMultimodalLM.from_pretrained("voxreality/rgb_language_cap", device_map="auto")
  • Notebooks
  • Google Colab
  • Kaggle

You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

Gated model
You can list files but not access them

Preview of files found in this repository
  • .gitattributes
    1.52 kB
    initial commit almost 2 years ago
  • README.md
    1.73 kB
    Update README.md almost 2 years ago
  • config.json
    4.85 kB
    Upload 10 files almost 2 years ago
  • generation_config.json
    149 Bytes
    Upload 10 files almost 2 years ago
  • merges.txt
    456 kB
    Upload 10 files almost 2 years ago
  • preprocessor_config.json
    374 Bytes
    Upload 10 files almost 2 years ago
  • pytorch_model.bin
    957 MB
    xet
    Upload 10 files almost 2 years ago
  • special_tokens_map.json
    131 Bytes
    Upload 10 files almost 2 years ago
  • tokenizer.json
    2.11 MB
    Upload 10 files almost 2 years ago
  • tokenizer_config.json
    476 Bytes
    Upload 10 files almost 2 years ago
  • training_args.bin
    6.14 kB
    xet
    Upload 10 files almost 2 years ago
  • vocab.json
    798 kB
    Upload 10 files almost 2 years ago