< cd ../gallery
[DevTools]
Deploying-Llama-3.3-70B
Serve Llama 3.3 70B (with AWQ quantization) using vLLM and deploy it on BentoCloud.
@kingabzpro
maintainer
★ 32 stars
# README
Serve Llama 3.3 70B (with AWQ quantization) using vLLM and deploy it on BentoCloud. Maintained by kingabzpro on GitHub, where it has earned 32 stars from the community.
It's actively developed around bentocloud, bentoml, cloud, and is a solid reference for anyone building with these tools.
# tags
# install
npm install Deploying-Llama-3.3-70Bmore FastAPI repos