Isbn: 9798341621497 - hands-on llm serving and optimization: hosting llms at scale (5 Ergebnisse)

- Softcover
Anbieter: PBShop.store UK, Fairford, GLOS, Vereinigtes KönigreichPBShop.store UK
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Gebraucht - Gut
EUR 50,37
EUR 5,91 VersandVersand von Vereinigtes Königreich nach USAAnzahl: 1 verfügbar
PAP. Zustand: Used - Very Good. Used - Like New Book. Shipped from UK. Established seller since 2000.

- Softcover
Anbieter: PBShop.store UK, Fairford, GLOS, Vereinigtes KönigreichPBShop.store UK
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 57,45
EUR 5,91 VersandVersand von Vereinigtes Königreich nach USAAnzahl: 3 verfügbar
PAP. Zustand: New. New Book. Shipped from UK. Established seller since 2000.

- Softcover
Anbieter: Speedyhen, Hertfordshire, Vereinigtes KönigreichSpeedyhen
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 52,03
EUR 48,24 VersandVersand von Vereinigtes Königreich nach USAAnzahl: 3 verfügbar
Zustand: NEW.

- Softcover
Anbieter: AHA-BUCH GmbH, Einbeck, DeutschlandAHA-BUCH GmbH
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 80,95
EUR 35,00 VersandVersand von Deutschland nach USAAnzahl: 2 verfügbar
Taschenbuch. Zustand: Neu. Neuware - Large language models (LLMs) are the reasoning engines of modern AI. Today, a major inflection point has arrived: as the world races to deploy AI at scale, model inference has moved to the center of the stack. Welcome to the inference era. Without proper optimization, however, LLMs can be expensive and slow to serve. Hands-On LLM Serving and Optimization is a comprehensive guide to the complexities of deploying and optimizing LLMs at scale. In this hands-on, engineering-focused book, authors Chi Wang and Peiheng Hu combine practical examples, code, and strategies for building robust, performant, and cost-efficient AI token factories. Whether you're building the LLM inference infrastructure or the applications that consume it, a deep understanding of LLM serving will make you a more effective, future-ready engineer as AI transforms how we work and build. - Learn the foundations of model serving with core concepts, design paradigms, and industry best practices - Understand the common challenges of hosting LLMs at scale - Balance latency and throughput to meet the demands of AI applications and business requirements - Host LLMs cost-effectively with practical, code-backed techniques.…

- Softcover
Anbieter: moluna, Greven, Deutschlandmoluna
Verkäufer/-in kontaktierenVerkäufer/-in mit 5 SternenZustand: Neu
EUR 76,01
EUR 48,99 VersandVersand von Deutschland nach USAAnzahl: 3 verfügbar
Zustand: New.