-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 279k • 205 -
microsoft/xclip-base-patch32
Video Classification • 0.2B • Updated • 97.3k • 114 -
MCG-NJU/videomae-base
Video Classification • 94.2M • Updated • 462k • 56 -
OpenGVLab/VideoMAEv2-Base
Video Classification • 86.2M • Updated • 26.9k • 19
Alban NYANTUDRE
anyantudre
AI & ML interests
ML Engineer 👨🏾💻| Deep Learning (Vision, Language, Speech)
Recent Activity
updated a model 1 day ago
anyantudre/waxal-checkpoints published a model 3 days ago
anyantudre/waxal-w2vbert_multi_s42 published a model 3 days ago
anyantudre/waxal-whisper_small_multi_s42Organizations
OCR
- Running663
MinerU Document Extraction Tools
📚663Embedded MinerU document extraction demo
- Running on ZeroAgentsFeatured492
DeepSeek OCR 2 Demo
🚀492Try out DeepSeek-OCR-2 on your PDFs or images
- Running on ZeroAgentsFeatured282
granite-docling-258M demo
📝282Convert and query documents from images with AI
- Running on ZeroAgents42
Multimodal RAG with Granite Vision
🚀42RAG example using Granite [vision, embedding, instruct]
Video-models
-
facebook/vjepa2-vitl-fpc64-256
Video Classification • 0.3B • Updated • 279k • 205 -
microsoft/xclip-base-patch32
Video Classification • 0.2B • Updated • 97.3k • 114 -
MCG-NJU/videomae-base
Video Classification • 94.2M • Updated • 462k • 56 -
OpenGVLab/VideoMAEv2-Base
Video Classification • 86.2M • Updated • 26.9k • 19
OCR
- Running663
MinerU Document Extraction Tools
📚663Embedded MinerU document extraction demo
- Running on ZeroAgentsFeatured492
DeepSeek OCR 2 Demo
🚀492Try out DeepSeek-OCR-2 on your PDFs or images
- Running on ZeroAgentsFeatured282
granite-docling-258M demo
📝282Convert and query documents from images with AI
- Running on ZeroAgents42
Multimodal RAG with Granite Vision
🚀42RAG example using Granite [vision, embedding, instruct]
models 7
anyantudre/waxal-checkpoints
Automatic Speech Recognition • Updated
anyantudre/waxal-w2vbert_multi_s42
0.6B • Updated
anyantudre/waxal-whisper_small_multi_s42
Automatic Speech Recognition • 0.2B • Updated
anyantudre/chatterbox-moore-finetuned
Text-to-Speech • Updated
anyantudre/spark-tts-0.5B-lora
Text Generation • 0.5B • Updated
anyantudre/mms-mos-discriminator
83M • Updated • 8
anyantudre/Llama-3-8b-ft-unsloth
Text Generation • Updated • 8
datasets 9
anyantudre/moore-speech-bible
Viewer • Updated • 84.5k • 9
anyantudre/MooreSpeechCorpora
Viewer • Updated • 5.54k • 11 • 3
anyantudre/moore-speech-devinettes
Viewer • Updated • 611 • 12 • 1
anyantudre/moore-speech-proverbes
Viewer • Updated • 2.32k • 9 • 1
anyantudre/moore-speech-contes
Viewer • Updated • 11.9k • 8 • 1
anyantudre/moore-speech-full-dataset
Viewer • Updated • 243k • 9
anyantudre/MooreSpeechCorporaCorrected
Viewer • Updated • 5.51k • 7
anyantudre/bf-data-raw-texts
Viewer • Updated • 1 • 17 • 1
anyantudre/chirps-morocco-daily
Viewer • Updated • 7.31M • 24 • 2