Inference Engineering in Practice: Serving LLMs Fast and Cheap: Batching, KV Cache, Quantization, and the Cost of Every Token - Gabriel Anhaia - cover
Inference Engineering in Practice: Serving LLMs Fast and Cheap: Batching, KV Cache, Quantization, and the Cost of Every Token - Gabriel Anhaia - cover
Dati e Statistiche
Salvato in 0 liste dei desideri
Inference Engineering in Practice: Serving LLMs Fast and Cheap: Batching, KV Cache, Quantization, and the Cost of Every Token
Disponibilità in 3 settimane
31,30 €
-5% 32,95 €
31,30 € 32,95 € -5%
Disponibilità in 3 settimane

Dettagli

Testo in English
229 x 152 mm
608 gr.
9798175785631
Informazioni e Contatti sulla Sicurezza dei Prodotti

Le schede prodotto sono aggiornate in conformità al Regolamento UE 988/2023. Laddove ci fossero taluni dati non disponibili per ragioni indipendenti da Feltrinelli, vi informiamo che stiamo compiendo ogni ragionevole sforzo per inserirli. Vi invitiamo a controllare periodicamente il sito www.lafeltrinelli.it per eventuali novità e aggiornamenti.
Per le vendite di prodotti da terze parti, ciascun venditore si assume la piena e diretta responsabilità per la commercializzazione del prodotto e per la sua conformità al Regolamento UE 988/2023, nonché alle normative nazionali ed europee vigenti.

Per informazioni sulla sicurezza dei prodotti, contattare productsafety@feltrinelli.it

​