$npx skillfedfor your agent

langchain-nvidia-ai-endpoints

An integration package connecting NVIDIA AI Endpoints and LangChain

Worth itPyPI Artificial IntelligenceReleased Jul 2026818.7K downloads / moMITPure Python

Decision gist · record as of 2026-08-14

pure-Python wheel — langchain_nvidia_ai_endpoints-1.4.3-py3-none-any.whl
v1.4.3 · released 2026-07-02 · Python <4.0.0,>=3.10.0 · 3 runtime deps: aiohttp, langchain-core, requests

Yes. The package has low install friction, active maintenance, no known vulnerabilities, permissive MIT licensing, and integrates seamlessly into LangChain workflows. It is worth installing if you need access to NVIDIA Foundation Models (especially Nemotron) or want the option to run models on-premises via NIM.AI-flagged interpretation of the facts on this page — verify before relying

Before you install

  • Requires NVIDIA_API_KEY environment variable set to a valid key from https://build.nvidia.com/ (format: nvapi-*), or a running NIM container endpoint if using self-hosted models.
  • Low friction installation with three runtime dependencies (aiohttp, langchain-core, requests).
  • Package is actively maintained with recent commits and no known vulnerabilities.

License · maintenance · safety

MIT (permissive) — MIT license permits free use, modification, and distribution with minimal restrictions, making it suitable for both open-source and commercial projects.

last release 2026-07-02 (43 days) · last repo commit 2026-08-11 · 210 stars

0 known vulnerabilities (OSV.dev, 2026-08-14) · 818,671 downloads/mo, #4,979 on PyPI

Verify before relying

pip install langchain-nvidia-ai-endpoints

from langchain_nvidia_ai_endpoints import ChatNVIDIA

llm = ChatNVIDIA(model="nvidia/nemotron-3-super-120b-a12b")
result = llm.invoke("Write a ballad about LangChain.")
print(result.content)
  • Whether the package supports all NVIDIA Foundation Models listed in the API Catalog or only a subset
  • Performance characteristics and latency expectations when using the NVIDIA API Catalog versus self-hosted NIM
  • Cost implications of using the NVIDIA API Catalog endpoints
Same gist for agents: .md · .json

What it is and what it does

This package bridges LangChain and NVIDIA's AI infrastructure, letting you use NVIDIA Foundation Models—particularly Nemotron and other open models—as chat and embedding providers. It works by connecting to either live endpoints on the NVIDIA API Catalog (cloud-hosted on DGX infrastructure) or to self-hosted NIM microservices running in containers on your own infrastructure.

You instantiate a ChatNVIDIA object with a model name, then use standard LangChain interfaces: invoke, stream, batch, and their async variants all work natively. The package handles authentication via an NVIDIA_API_KEY environment variable and exposes available_models to discover which models your credentials can access. It integrates with LangChain's prompt templates and output parsers, so you can build chains and agents using NVIDIA models as the LLM backbone.

Use it for

  • Build agentic AI workflows using Nemotron's reasoning and tool-calling capabilities within LangChain
  • Generate code using specialized models like meta/codellama-70b or google/codegemma-7b through LangChain chains
  • Process multimodal inputs (text and images) with models like nvidia/neva-22b for reasoning tasks
  • Stream real-time responses from NVIDIA models in LangChain applications without blocking
  • Run inference on-premises using self-hosted NIM containers for full IP and customization control

Worth the install?

AI-flagged interpretation of the facts on this page. Verify before relying on it.

Worth it

Yes.

The package has low install friction, active maintenance, no known vulnerabilities, permissive MIT licensing, and integrates seamlessly into LangChain workflows. It is worth installing if you need access to NVIDIA Foundation Models (especially Nemotron) or want the option to run models on-premises via NIM.

Install

langchain-nvidia-ai-endpoints on PyPI

Before you install

Low friction installation with three runtime dependencies (aiohttp, langchain-core, requests). Package is actively maintained with recent commits and no known vulnerabilities.

Requires NVIDIA_API_KEY environment variable set to a valid key from https://build.nvidia.com/ (format: nvapi-*), or a running NIM container endpoint if using self-hosted models.

License in practice

MIT license permits free use, modification, and distribution with minimal restrictions, making it suitable for both open-source and commercial projects.

Quickstart

pip install langchain-nvidia-ai-endpoints

from langchain_nvidia_ai_endpoints import ChatNVIDIA

llm = ChatNVIDIA(model="nvidia/nemotron-3-super-120b-a12b")
result = llm.invoke("Write a ballad about LangChain.")
print(result.content)

Verify before relying

  • Whether the package supports all NVIDIA Foundation Models listed in the API Catalog or only a subset
  • Performance characteristics and latency expectations when using the NVIDIA API Catalog versus self-hosted NIM
  • Cost implications of using the NVIDIA API Catalog endpoints

Package facts

LicenseMIT permissive
Python supportSupports the current Python release <4.0.0,>=3.10.0
Install frictionLow. Pure-Python wheel
Runtime dependencies
3 packages
aiohttplangchain-corerequests
MaintenanceActively maintained 43 days since the last release
Last repo commit
First released
Downloads818,671 / month, #4,979 on PyPI 30-day window, as of 2026-08-14
Known vulnerabilitiesNone known OSV.dev, checked 2026-08-14
Classifiers
License :: OSI Approved :: MIT LicenseProgramming Language :: Python :: 3Programming Language :: Python :: 3.10Programming Language :: Python :: 3.11Programming Language :: Python :: 3.12Programming Language :: Python :: 3.13Programming Language :: Python :: 3.14

Evidence: langchain_nvidia_ai_endpoints-1.4.3-py3-none-any.whl

Tags

Capabilities
langchain nvidia integrationnvidia foundation models apinemotron chat modelsnvidia ai endpointslangchain llm nvidianvidia nim integrationlangchain embeddings nvidia
Topics
llm-integrationnvidia-aiagentic-ai

Let your AI agent find packages like this

Example. Real query, live index.

You found this page by searching. An agent finds it by wishing: SkillFed indexes 14,416 PyPI packages by what they can do, searchable in plain language.

wish › “langchain nvidia integration”

Give your agent the search over MCP, or paste the wish link into any chat.

More Artificial Intelligence packages

litellm With conditions
PyPI · Artificial Intelligence · released Aug 2026

LiteLLM provides a unified Python interface to call 100+ LLM providers (OpenAI, Anthropic, Gemini, Bedrock, Azure, and others) using OpenAI-compatible API format, available as both a Python SDK and a self-hosted AI Gateway proxy server.

Install it if you need to work with multiple LLM providers or want to centralize LLM routing in your organization.

MITcompiled wheel
682.8Mdownloads / mo
huggingface-hub Worth it
PyPI · Artificial Intelligence · released Aug 2026

Client library and CLI tool for downloading, uploading, and managing models, datasets, and repositories on the Hugging Face Hub platform.

Install it if you work with Hugging Face Hub models or datasets.

Apache-2.0pure Python · 3.10.0+
442.4Mdownloads / mo
langchain Worth it
PyPI · Python Modules · released Aug 2026

LangChain provides a framework for building agents and LLM-powered applications by composing language models, tools, and memory through a unified API that abstracts over multiple model providers.

MITpure Python
315.4Mdownloads / mo
hf-xet With conditions
PyPI · Artificial Intelligence · released Aug 2026

hf-xet provides chunk-based deduplication and efficient file transfer for the Hugging Face Hub, enabling faster uploads and downloads of large files with local disk caching.

Apache-2.0compiled wheel · 3.8+
258.4Mdownloads / mo
tokenizers Worth it
PyPI · Artificial Intelligence · released Apr 2026

Tokenizers converts raw text into token sequences for NLP models, with support for training custom vocabularies and using pre-built tokenizers (BPE, WordPiece) optimized for speed via Rust.

Apache-2.0compiled wheel · 3.10+
222.9Mdownloads / mo
transformers Worth it
PyPI · Artificial Intelligence · released Aug 2026

Transformers provides a unified framework for loading, fine-tuning, and running state-of-the-art pretrained models across text, vision, audio, video, and multimodal tasks using PyTorch, JAX, or TensorFlow.

Install it if you need to run or train any transformer-based model for NLP, vision, audio, or multimodal tasks.

permissive licensepure Python · 3.10.0+
186.6Mdownloads / mo

See also langchain-oci · langchain-cohere · langchain-sambanova · langchain-google-genai · ngcsdk · langchain-mistralai · nvidia-nat-langchain · langchain-ibm · langchain-baseten