tutorials/HuggingFace at main · triton-inference-server/tutorials
github.com · 1,130 words · saved by 1 readers
This repository contains tutorials and examples for Triton Inference Server - tutorials/HuggingFace at main · triton-inference-server/tutorials
Deploying HuggingFace models Note : If you are new to the Triton Inference Server, it is recommended to review Part 1 of the Conceptual Guide . This tutorial assumes basic understanding about the Triton Inference Server. Related Pages HuggingFace model exporting guide: ONNX , TorchScript Developers often work with open source models. HuggingFace is a popular source of many open source models. The discussion in this guide will focus on how a user can deploy almost any model from HuggingFace with the Triton Inference Server. For this example, the ViT model available on HuggingFace is being used.
related reading
- tutorials/Conceptual_Guide/Part_1-model_deployment at main · triton-inference-server/tutorials · GitHubgithub.com
- Together AI | The AI Native Cloudtogether.ai
- Hugging Face – The AI community building the future.huggingface.co
- Accelerated Inference for Large Transformer Models Using NVIDIA Triton Inference Server | NVIDIA Technical Blogdeveloper.nvidia.com
- tutorials/Conceptual_Guide/Part_4-inference_acceleration/README.md at main · triton-inference-server/tutorials · GitHubgithub.com
- Replicate - Run AI with an APIreplicate.com
- How to Deploy Your Modelhtdym.sailresearch.com
- Hugging Face · GitHubgithub.com
- GitHub - mrdbourke/cs329s-ml-deployment-tutorial: Code and files to go along with CS329s machine learning model deployment tutorial.github.com
- PyTorch Inference | onnxruntimeonnxruntime.ai
- GitHub - triton-inference-server/backend: Common source, scripts and utilities for creating Triton backends. · GitHubgithub.com
- Parsed | Custom, interpretable AI systems that continuously learnparsed.com