Digital Library

Hands-On LLM Serving and Optimization (Chi Wang, Peiheng Hu)(Z-Library)

Chi Wang, Peiheng Hu

Hands-On LLM Serving and Optimization (Chi Wang, Peiheng Hu)(Z-Library)

Author Chi Wang, Peiheng Hu

技术

Large language models (LLMs) are rapidly becoming the backbone of AI-driven applications. Without proper optimization, however, LLMs can be expensive to run, slow to serve, and prone to performance bottlenecks. As the demand for real-time AI applications grows, along comes this comprehensive guide to the complexities of deploying and optimizing LLMs at scale. Authors Chi Wang and Peiheng Hu take a real-world approach backed by practical examples and code, and assemble essential strategies for designing infrastructures that are equal to the demands of modern AI applications. Learn the key principles for designing a model-serving system and how to build a system that meets specific business requirements.

Format EPUB
Size 10.8 MB
81
Views
0
Downloads
0.00
Total Donations

Support Author

0.00
Total Amount (¥)
0
Donation Count
Please enter an amount Minimum ¥1

You will be redirected to Alipay to complete payment, then return here.

Recommended for You

Loading recommended books...
Failed to load, please try again later
Back to List