i try the nvidia / llama-3.1-nemotron-70b-instruct model with openai api key and its taking long time with no response.
tony3t3t
5
Same with me. Last night it was working but this morning waiting for a long time with no response. I use Nvidia API.
Related topics
| Topic | Replies | Views | Activity | |
|---|---|---|---|---|
| OpenAI Compatible API does not work | 6 | 1222 | August 26, 2024 | |
| Result of nvidia nims in openai SDK and API inconsistent | 0 | 152 | January 7, 2025 | |
| Llama 3.1 nemotron 70b instruct API access not working correctly | 1 | 397 | December 5, 2025 | |
| Open AI Endpoint | 0 | 386 | April 28, 2024 | |
| Models That Don't Work | 0 | 245 | July 25, 2026 | |
| Supercharging Llama 3.1 across NVIDIA Platforms | 13 | 594 | September 17, 2024 | |
| Optimizing Inference on Large Language Models with NVIDIA TensorRT-LLM, Now Publicly Available | 8 | 2194 | January 25, 2024 | |
| "404 Page Not Found" Error When api used as openai | 3 | 1928 | February 22, 2026 | |
| Turbocharging Meta Llama 3 Performance with NVIDIA TensorRT-LLM and NVIDIA Triton Inference Server | 61 | 5058 | August 28, 2024 | |
| Request to enable "Public API Endpoints" permission for my personal organization | 2 | 396 | May 8, 2026 |