Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

thangylvp
/
stcc

Audio-Text-to-Text
Transformers
Safetensors
Vietnamese
qwen3_asr
automatic-speech-recognition
qwen3-asr
speech-to-command
function-calling
tool-calling
vllm
custom_code
Model card Files Files and versions
xet
Community

Instructions to use thangylvp/stcc with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

  • Libraries
  • Transformers

    How to use thangylvp/stcc with Transformers:

    # Load model directly
    from transformers import AutoProcessor, AutoModelForMultimodalLM
    
    processor = AutoProcessor.from_pretrained("thangylvp/stcc", trust_remote_code=True)
    model = AutoModelForMultimodalLM.from_pretrained("thangylvp/stcc", trust_remote_code=True, device_map="auto")
  • Notebooks
  • Google Colab
  • Kaggle

You need to agree to share your contact information to access this model

This repository is publicly accessible, but you have to accept the conditions to access its files and content.

Log in or Sign Up to review the conditions and access this model content.

Gated model
You can list files but not access them

Preview of files found in this repository
  • README.md
    1.37 kB
    Add verified 1-9 second audio test suite about 16 hours ago
  • expected_outputs.json
    4.54 kB
    Add verified 1-9 second audio test suite about 16 hours ago
  • non_tool_01s_news_prefix.wav
    48.1 kB
    Add verified 1-9 second audio test suite about 16 hours ago
  • non_tool_news.wav
    430 kB
    xet
    Add tool-call and non-tool audio smoke tests 1 day ago
  • tool_call_03s_driver_seat_heat_level_2.wav
    144 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_04s_rear_fan_face_feet.wav
    192 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_05s_front_rainbow_ambient_light.wav
    240 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_06s_front_rainbow_ambient_light.wav
    288 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_07s_unsubscribe_news_podcast.wav
    336 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_08s_next_am_station.wav
    384 kB
    xet
    Add verified 1-9 second audio test suite about 16 hours ago
  • tool_call_set_fog_lights.wav
    96 kB
    Add tool-call and non-tool audio smoke tests 1 day ago