-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 291k • 205 -
microsoft/xclip-base-patch32
Video Classification • 0.2B • Updated • 116k • 114 -
MCG-NJU/videomae-base
Video Classification • 94.2M • Updated • 461k • 56 -
OpenGVLab/VideoMAEv2-Base
Video Classification • 86.2M • Updated • 26.9k • 19
Alban NYANTUDRE
anyantudre
AI & ML interests
ML Engineer 👨🏾💻| Deep Learning (Vision, Language, Speech)
Recent Activity
published a model about 15 hours ago
anyantudre/waxal-w2vbert_multi_s42 published a model about 15 hours ago
anyantudre/waxal-whisper_small_multi_s42 published a model about 15 hours ago
anyantudre/spark-tts-0.5B-loraOrganizations
OCR
- Running661
MinerU Document Extraction Tools
📚661Embedded MinerU document extraction demo
- Running on ZeroAgentsFeatured491
DeepSeek OCR 2 Demo
🚀491Try out DeepSeek-OCR-2 on your PDFs or images
- Running on ZeroAgentsFeatured282
granite-docling-258M demo
📝282Convert and query documents from images with AI
- Running on ZeroAgents42
Multimodal RAG with Granite Vision
🚀42RAG example using Granite [vision, embedding, instruct]
Video-models
-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 291k • 205 -
microsoft/xclip-base-patch32
Video Classification • 0.2B • Updated • 116k • 114 -
MCG-NJU/videomae-base
Video Classification • 94.2M • Updated • 461k • 56 -
OpenGVLab/VideoMAEv2-Base
Video Classification • 86.2M • Updated • 26.9k • 19
OCR
- Running661
MinerU Document Extraction Tools
📚661Embedded MinerU document extraction demo
- Running on ZeroAgentsFeatured491
DeepSeek OCR 2 Demo
🚀491Try out DeepSeek-OCR-2 on your PDFs or images
- Running on ZeroAgentsFeatured282
granite-docling-258M demo
📝282Convert and query documents from images with AI
- Running on ZeroAgents42
Multimodal RAG with Granite Vision
🚀42RAG example using Granite [vision, embedding, instruct]
models 7
anyantudre/waxal-checkpoints
Automatic Speech Recognition • Updated
anyantudre/waxal-w2vbert_multi_s42
0.6B • Updated
anyantudre/waxal-whisper_small_multi_s42
Automatic Speech Recognition • 0.2B • Updated
anyantudre/chatterbox-moore-finetuned
Text-to-Speech • Updated
anyantudre/spark-tts-0.5B-lora
Text Generation • 0.5B • Updated
anyantudre/mms-mos-discriminator
83M • Updated • 9
anyantudre/Llama-3-8b-ft-unsloth
Text Generation • Updated • 9
datasets 10
anyantudre/waxal-pseudo
Viewer • Updated • 44.9k • 3
anyantudre/moore-speech-bible
Viewer • Updated • 84.5k • 6
anyantudre/MooreSpeechCorpora
Viewer • Updated • 5.54k • 10 • 3
anyantudre/moore-speech-devinettes
Viewer • Updated • 611 • 11 • 1
anyantudre/moore-speech-proverbes
Viewer • Updated • 2.32k • 8 • 1
anyantudre/moore-speech-contes
Viewer • Updated • 11.9k • 8 • 1
anyantudre/moore-speech-full-dataset
Viewer • Updated • 243k • 6
anyantudre/MooreSpeechCorporaCorrected
Viewer • Updated • 5.51k • 6
anyantudre/bf-data-raw-texts
Viewer • Updated • 1 • 20 • 1
anyantudre/chirps-morocco-daily
Viewer • Updated • 7.31M • 27 • 2