🏗️ Building on HF
Adarsh Zolekar
adarshzolekar
AI & ML interests
Exploring AI, ML, Deep Learning, models and datasets while building and contributing to the Hugging Face community.
Recent Activity
liked a model 1 day ago
google/embeddinggemma-2Organizations
Multimodal AI Models
Purpose: Models that understand text + image + audio together.
Vision Models (Image & Video)
Purpose: Text-to-image, image classification, detection, segmentation.
-
openai/clip-vit-base-patch32
Zero-Shot Image Classification • Updated • 19.2M • 1.58k -
facebook/detr-resnet-50
Object Detection • 41.6M • Updated • 173k • • 984 -
Tongyi-MAI/Z-Image-Turbo
Text-to-Image • 6B • Updated • 631k • • 5.44k -
black-forest-labs/FLUX.1-dev
Text-to-Image • 12B • Updated • 659k • • 15.5k
Embeddings & Retrieval Models (RAG)
-
sentence-transformers/all-MiniLM-L6-v2
Sentence Similarity • 22.7M • Updated • 223M • • 6.24k -
BAAI/bge-m3
Sentence Similarity • Updated • 32.9M • • 3.84k -
nomic-ai/nomic-embed-text-v1.5
Sentence Similarity • 0.1B • Updated • 12.9M • 950 -
BAAI/bge-reranker-v2-m3
Text Classification • 0.6B • Updated • 16.2M • • 1.23k
Audio & Speech Models
Purpose: Speech recognition, text-to-speech, music, audio analysis.
-
openai/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 3.74M • • 6.59k -
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.26M • • 3.43k -
hexgrad/Kokoro-82M
Text-to-Speech • Updated • 10.8M • • 7.2k -
Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice
Text-to-Speech • 2B • Updated • 2.22M • 2.1k
Text & Code Models (NLP)
Purpose: Text generation, summarization, translation, embeddings, coding.
-
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 2.07M • 3.74k -
Qwen/Qwen3-8B
Text Generation • 8B • Updated • 9.65M • • 2.09k -
deepseek-ai/DeepSeek-V4-Flash-0731
Text Generation • 304B • Updated • 4.44M • • 4.02k -
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF
Text Generation • 31B • Updated • 4.7M • 1.13k
Reasoning & Agentic Models
Embeddings & Retrieval Models (RAG)
-
sentence-transformers/all-MiniLM-L6-v2
Sentence Similarity • 22.7M • Updated • 223M • • 6.24k -
BAAI/bge-m3
Sentence Similarity • Updated • 32.9M • • 3.84k -
nomic-ai/nomic-embed-text-v1.5
Sentence Similarity • 0.1B • Updated • 12.9M • 950 -
BAAI/bge-reranker-v2-m3
Text Classification • 0.6B • Updated • 16.2M • • 1.23k
Multimodal AI Models
Purpose: Models that understand text + image + audio together.
Audio & Speech Models
Purpose: Speech recognition, text-to-speech, music, audio analysis.
-
openai/whisper-large-v3
Automatic Speech Recognition • 2B • Updated • 3.74M • • 6.59k -
openai/whisper-large-v3-turbo
Automatic Speech Recognition • 0.8B • Updated • 6.26M • • 3.43k -
hexgrad/Kokoro-82M
Text-to-Speech • Updated • 10.8M • • 7.2k -
Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice
Text-to-Speech • 2B • Updated • 2.22M • 2.1k
Vision Models (Image & Video)
Purpose: Text-to-image, image classification, detection, segmentation.
-
openai/clip-vit-base-patch32
Zero-Shot Image Classification • Updated • 19.2M • 1.58k -
facebook/detr-resnet-50
Object Detection • 41.6M • Updated • 173k • • 984 -
Tongyi-MAI/Z-Image-Turbo
Text-to-Image • 6B • Updated • 631k • • 5.44k -
black-forest-labs/FLUX.1-dev
Text-to-Image • 12B • Updated • 659k • • 15.5k
Text & Code Models (NLP)
Purpose: Text generation, summarization, translation, embeddings, coding.
-
mistralai/Mistral-7B-Instruct-v0.3
7B • Updated • 2.07M • 3.74k -
Qwen/Qwen3-8B
Text Generation • 8B • Updated • 9.65M • • 2.09k -
deepseek-ai/DeepSeek-V4-Flash-0731
Text Generation • 304B • Updated • 4.44M • • 4.02k -
unsloth/Qwen3-Coder-30B-A3B-Instruct-GGUF
Text Generation • 31B • Updated • 4.7M • 1.13k