< cd ../gallery
[DevTools]

Deploying-Llama-3.3-70B

Serve Llama 3.3 70B (with AWQ quantization) using vLLM and deploy it on BentoCloud.
kingabzpro
@kingabzpro
maintainer
★ 32 stars
go to website ↗github
deploying-llama-3-3-70b — preview
Deploying-Llama-3.3-70B repository preview
# README
Serve Llama 3.3 70B (with AWQ quantization) using vLLM and deploy it on BentoCloud. Maintained by kingabzpro on GitHub, where it has earned 32 stars from the community.
It's actively developed around bentocloud, bentoml, cloud, and is a solid reference for anyone building with these tools.
# tags
# install
npm install Deploying-Llama-3.3-70B
languages
Python100%
last commit1 year ago
licenseApache-2.0
more FastAPI repos
$ made-with-fastapi
rssllmmadewithwhat

Ask MadeWithWhat

AI answers may contain mistakes — please double-check important details.