Dell + Hugging Face

Take open AI from model to production on premises.

Discover validated models, deployable applications, and hardware-aware configurations engineered for secure AI on Dell infrastructure.

New dell-ai brings catalog discovery, configuration, and deployment to your terminal.
Model catalog

Latest optimized models

Recently added open models, ready for Dell infrastructure.

Author avatar

MiMo V2.6 Flash RL

Deployable
309B
MIT
3 Platforms

New MiMo-V2.6-Flash-RL is a 309B-parameter (15B active) omnimodal Mixture-of-Experts reasoning model with a 1M-token context window, built for coding, agentic, visual, audio, and cybersecurity workloads.

Author avatar

MiMo V2.6 Pro RL

Deployable
1T
MIT
3 Platforms

New MiMo-V2.6-Pro-RL is a 1.02T-parameter (42B active) flagship omnimodal Mixture-of-Experts reasoning model with a 1M-token context window, built for coding, agentic, visual, audio, and cybersecurity workloads.

Author avatar

Laya

Deployable
421M
Apache 2.0
7 Platforms

New Laya is a multilingual, non-autoregressive System One decision model for calibrated routing, scoring, guardrails, moderation, and other low-latency decisions.

Author avatar

NVIDIA Nemotron 3 Diarization

Deployable
100M
Other
1 Platform

100M-parameter speaker diarization model for identifying who spoke when in live and recorded audio, including overlapping speech from up to eight speakers.

Author avatar

DeepSeek V4.1 Flash

Deployable
284B
MIT
8 Platforms

New DeepSeek V4.1 Flash is a Mixture-of-Experts model built for efficient reasoning, tool use, and agentic workloads.

Author avatar

DeepSeek V4 Flash Vision Exp

Deployable
284B
MIT
7 Platforms

New DeepSeek V4 Flash Vision Exp is an experimental multimodal MoE model built for efficient visual reasoning, tool use, and agentic workloads. It matches the performance of DeepSeek V4 Flash while heavily increasing multimodal agent capabilities.

See all models →
Newsroom

Latest updates

New models, platform support, and practical deployment guidance.

Xiaomi MiMo V2.6 Pro and Flash land on Dell Enterprise Hub

Xiaomi’s MiMo V2.6 Pro RL and MiMo V2.6 Flash RL are now available on Dell Enterprise Hub, bringing frontier intelligence on prem. These native omnimodal Mixture of Experts reasoning models combine text, image, video, and audio understanding with a 1M token context window and a DFlash style speculative decoder. Deploy Flash on 4× or 8× NVIDIA H200 and B300 systems, and Pro on 8× H200 or 4× and 8× B300 systems. Deploy MiMo V2.6 Flash RL → · Deploy MiMo V2.6 Pro RL →

Laya brings low-latency decisions to Dell Enterprise Hub

Laya is now available on Dell Enterprise Hub. This multilingual, non autoregressive System One model delivers calibrated routing, scoring, guardrails, moderation, and other real time decisions through a purpose built API. Deploy it on 1× NVIDIA L40S, H100, H200, or RTX PRO 6000. Deploy Laya →

NVIDIA Nemotron 3 Diarization lands on Dell Enterprise Hub

NVIDIA Nemotron 3 Diarization identifies who spoke when in live and recorded audio, including overlapping speech from up to eight speakers. Deploy the 100M parameter model on 1× NVIDIA L40S. Deploy Nemotron 3 Diarization →

dell-ai v1.1.0 brings MCP support to Dell Enterprise Hub

dell ai v1.1.0 adds a Model Context Protocol (MCP) server, connecting AI assistants to Dell Enterprise Hub. Discover validated models and compatible Dell platforms, explore applications, and generate deployment configurations directly from your MCP client. Enable deployment tools when you want your assistant to deploy and manage workloads using the server’s local environment. Get started with MCP →

See all updates →
Application catalog

Deploy complete AI applications

Go beyond endpoints with production-ready AI experiences.

Super Analyzer

Super Analyzer is an application that identifies and evaluates common anti-patterns in C++, Java, Python, and Rust code. It uses a multi-agent architecture with three coordinated agent types—Primary, Fixer, and Chat—to deliver accurate analysis and iterative improvements. The Web UI supports natural multi-turn interactions and uses a PostgreSQL database to persist in user accounts and conversation history across sessions. Super Analyzer also provides a REST/API layer and a Python interface, all secured through user/password authentication backed by PostgreSQL. Access is protected with JWTs signed using RSA-256, ensuring a consistent and secure experience across all entry points.

nvidianemotronmulti-agent

OpenWebUI

OpenWebUI is an extensible, feature-rich, and user-friendly self-hosted AI platform designed to operate entirely offline. It provides a modern interface for interacting with various LLM runners including Ollama and OpenAI-compatible APIs. With built-in RAG (Retrieval Augmented Generation) capabilities, OpenWebUI offers a comprehensive solution for AI deployment and interaction in Kubernetes environments.

chatllmmulti-modal

AnythingLLM

AnythingLLM is an open-source, all-in-one AI chat application with built-in RAG (Retrieval Augmented Generation) capabilities. It provides a user-friendly interface for document management, LLM integration, and AI agent creation - all without complex setup. The application supports both local and cloud-based deployments, making it ideal for teams and organizations seeking a private, customizable AI solution.

chatragagents

Agentic Smart Router

In an Agentic AI application, not all prompts are the same, different prompts need to have different requirements in terms of complexity of the prompt - some need reasoning, some need access to RAG system and some need tool calling capabilities. In this custom-built Agentic Smart Router, we demonstrate that capability by leveraging NVIDIA Agent Intelligence Toolkit, NVIDIA NIMs, NVIDIA LLM Router blueprint and stitching them together to build an Agentic Smart Router application.

nvidianimsagents
See all apps →