Inference Providers
Active filters: nemotron
aufklarer/VoiceChat-11B-Perception-MLX-int5
Audio-to-Audio
• Updated • 91
• 5
pipecat-ai/NVIDIA-NemotronLabs-VoiceChat-11B-Spark
Audio-to-Audio
• Updated • 4
mlx-community/NemotronLabs-VoiceChat-11B-mlx-8bit
4B • Updated • 64
• 4
OpenMed/privacy-filter-nemotron-v2
Token Classification
• 1B • Updated • 21.3k
• • 48
OpenYourMind/OpenYourMind-NVIDIA-Nemotron-3-Ultra-550B-A55B-abliterated-uncensored-NVFP4
Text Generation
• 335B • Updated • 530
• 7
ewinregirgojr/MiniCPM5-1B-Agentic-Tooluse-Merged-FP16
Text Generation
• 1B • Updated • 1.17k
• 8
Abiray/Nemotron-3-Embed-8B-GGUF
8B • Updated • 15.8k
• 5
shadowrock-io/Nemotron-3-Embed-8B-Community-NVFP4
Sentence Similarity
• 5B • Updated • 52
• 2
mlx-community/NemotronLabs-VoiceChat-11B-mlx-bf16
11B • Updated • 103
• 2
cybermotaz/nemotron3-nano-nvfp4-w4a16
Text Generation
• 18B • Updated • 2.26k
• 14
thunkaboutit/Koleslaw-Nemotron-30B-A3B-PromptEnhance
Text Generation
• 32B • Updated • 1.07k
• 3
mudler/Nemotron-3-Nano-Omni-30B-A3B-Reasoning-APEX-GGUF
32B • Updated • 9.74k
• 12
OpenYourMind/OpenYourMind-NVIDIA-Nemotron-3-Ultra-550B-A55B-abliterated-uncensored
Text Generation
• 561B • Updated • 35
• 5
ewinregirgojr/MiniCPM5-1B-Agentic-Tooluse-GGUF
Text Generation
• 1B • Updated • 14.1k
• 73
OpenLLM-France/Luciole-23B-Instruct-1.1
Text Generation
• 23B • Updated • 1.19k
• 15
danielrmay/NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-W4A16
Text Generation
• 82B • Updated • 3.35k
• 7
shadowrock-io/Nemotron-3-Embed-8B-Community-MLX-4bit
Sentence Similarity
• 1B • Updated • 35
• 1
shadowrock-io/Nemotron-3-Embed-8B-Community-FP8
Sentence Similarity
• 8B • Updated • 21
• 1
aufklarer/VoiceChat-11B-Perception-MLX-int8
Audio-to-Audio
• Updated • 61
• 1
mlx-community/NemotronLabs-VoiceChat-11B-mlx-4bit
3B • Updated • 94
• 1
Text Generation
• 9B • Updated • 9.29k
failspy/Nemotron-4-340B-Instruct-SafeTensors
Text Generation
• 341B • Updated • 16
• 22
leafspark/Nemotron-340B-Instruct-hf
Text Generation
• Updated • 7
Text Generation
• Updated • 17.7k
• 71
Text Generation
• Updated • 5.78k
• 138
mgoin/Nemotron-4-340B-Instruct-vllm
Text Generation
• 341B • Updated • 79
mgoin/Nemotron-4-340B-Instruct-FP8-Dynamic
Text Generation
• 341B • Updated • 7
mgoin/Minitron-4B-Base-FP8
Text Generation
• 4B • Updated • 17
• 3
mgoin/Minitron-8B-Base-FP8
Text Generation
• 8B • Updated • 10
• 3
mgoin/nemotron-3-8b-chat-4k-sft-hf
Text Generation
• 9B • Updated • 80