
Production Grade LLM deployment and High-Load Inferencing with vLLm, Chatbots with Memory, Local Cache of Model Weights
Production LLM Deployment: vLLM,FastAPI,Modal and AI Chatbot
InstructorPetar PetkanovDuration5h 28m
Students278
Rating4.1 (19)

Production Grade LLM deployment and High-Load Inferencing with vLLm, Chatbots with Memory, Local Cache of Model Weights
InstructorPetar Petkanov
Coupon
Coupon
Coupon
Free