techsword's picture
Add converted checkpoint_best (85k updates, main branch)
4b88e67 verified
|
Raw
History Blame Contribute Delete
596 Bytes
metadata
library_name: transformers
license: apache-2.0
pipeline_tag: feature-extraction
tags:
  - audio
  - wav2vec2

Fairseq wav2vec2-base pretraining checkpoint (checkpoint_best, ~85000 updates), converted to HuggingFace format with the official transformers converter and verified (weight-level spot check + forward-pass comparison against the fairseq model; see conversion log). Pretrained on MAGICDATA (Mandarin Chinese).

Raw fairseq checkpoints: techsword/wav2vec2-base-mandarin-magicdata-checkpoints Used in: tone-probe experiments, https://github.com/techsword/tone-encoding-in-speech-model