Dear NVIDIA NIM Team,
Thank you for maintaining and continuously improving the NVIDIA NIM API platform. I would like to submit a request regarding several models that appear to be outdated or superseded by newer alternatives.
As the AI ecosystem evolves rapidly, updating the available model catalog would improve user experience, performance, and overall platform competitiveness. I would like to suggest considering the deprecation or replacement of the following models:
-
Kimi 2.6 → Upgrade to Kimi K2.7 or Kimi K3.
-
MiniMax M2.7 → Replace with an optimized MiniMax M3 deployment.
-
NVIDIA Nemotron Nano 12B V2 VL (
nvidia/nemotron-nano-12b-v2-vl). -
Qwen3 Next 80B A3B Instruct (
qwen/qwen3-next-80b-a3b-instruct) → Consider upgrading to Qwen 3.6 36B A3B or other models from the Qwen 3.6 family. -
Llama 3.1 Nemotron Nano VL 8B V1 (
llama-3.1-nemotron-nano-vl-8b-v1). -
Meta Llama 3.2 90B Vision Instruct (
meta/llama-3.2-90b-vision-instruct). -
BGE-M3.
Suggested replacement priorities:
-
Kimi K3 (or K2.7) for stronger reasoning and coding capabilities.
-
MiniMax M3 for improved efficiency and multimodal performance.
-
Qwen 3.6 family models (particularly 36B A3B) for a modern balance of quality, latency, and cost.
-
Newer Nemotron and Llama variants where applicable.
-
Updated embedding models to replace BGE-M3 if more capable alternatives are available.
Benefits of these upgrades include:
-
Better reasoning and coding performance.
-
Improved multimodal capabilities.
-
Lower latency and higher throughput.
-
Reduced maintenance overhead for legacy deployments.
-
Greater alignment with current open-model benchmarks and community adoption.
The NVIDIA NIM platform is already an excellent ecosystem for deploying AI models, and keeping the model catalog aligned with the latest generation of foundation models would further strengthen its position for developers and enterprises alike.
Thank you for considering this feedback. I appreciate the team’s efforts and look forward to future updates to the NVIDIA NIM API platform.
Best regards,
supervisorxscp