NVIDIA Corporation
- 28.2k followers
- 2788 San Tomas Expressway, Santa Clara, CA, 95051
- https://nvidia.com
Pinned Loading
Repositories
- Model-Optimizer Public
A unified library of SOTA model optimization techniques like quantization, distillation, pruning, neural architecture search, speculative decoding, etc. It compresses deep learning models for downstream deployment frameworks like TensorRT-LLM, TensorRT, vLLM, etc. to optimize inference speed.
- aicr Public
Tooling for optimized, validated, and reproducible GPU-accelerated AI runtime in Kubernetes
- NeMo-Relay Public
Multi-language agent runtime and library for execution scope management, lifecycle events, and middleware on tool and LLM calls.
- nvcf Public
Platform for deploying and routing GPU-accelerated inference, streaming, and batch workloads at scale.
- NeMo-Retriever Public
NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice. NeMo Retriever Library uses specialized NVIDIA NIM microservices to find, contextualize, and extract text, tables, charts and images that you can use in downstream generative applications.
Top languages
Loading…
Most used topics
Loading…