Alle boeken

Professioneel

Apps Over Coach Inloggen Begin met lezen

Quantitative Finance · Begrippenlijst

Wat is Dynamic batching?

Definition 26.4 Machine Learning for Markets · Hoofdstuk 26 — Low-Latency Inference

Dynamic batching groups requests that arrive close together into one call of the model, trading the waiting time of the first request for a lower cost per request.

LightGBM’s prediction latency (median) against batch size through its Python interface, per call and per row. Same machine and conditions as . Data: bench_infer.py, measured_batch.csv.
Figure 26.2. LightGBM’s prediction latency (median) against batch size through its Python interface, per call and per row. Same machine and conditions as Figure 26.1. Data: bench_infer.py, measured_batch.csv.
Lees in het hoofdstuk →