Deep limit order book forecasting: a microstructural guide

Antonio Briola1, Silvia Bartolucci2, Tomaso Aste1,2

  • 1Department of Computer Science, University College London, London, WC1E 6EA, UK.

Quantitative Finance
|August 4, 2025
PubMed
Summary

Deep learning can predict stock mid-price changes using Limit Order Book data, but high accuracy doesn't guarantee profitable trading signals. New metrics are needed to assess practical forecasting in this domain.

Related Concept Videos

Orders of Magnitude01:15

Orders of Magnitude

The order of magnitude of a number is the power of 10 that most closely approximates it. Thus, the order of magnitude estimates the scale (or size) of its value. To find the order of magnitude of a number, take the base-10 logarithm of the number and round it to the nearest integer. Then the order of magnitude of the number is simply the resulting power of 10.
The order of magnitude is simply a way of rounding numbers consistently to the nearest power of 10. This makes doing rough mental math...
21.4K
Prediction Intervals01:03

Prediction Intervals

The interval estimate of any variable is known as the prediction interval. It helps decide if a point estimate is dependable.
However, the point estimate is most likely not the exact value of the population parameter, but close to it. After calculating point estimates, we construct interval estimates, called confidence intervals or prediction intervals. This prediction interval comprises a range of values unlike the point estimate and is a better predictor of the observed sample value, y. 
2.3K
Difference from Background: Limit of Detection01:05

Difference from Background: Limit of Detection

The limit of detection (LOD) is the smallest amount of analyte that can be distinguished from the background noise. The LOD value corresponds to the concentration at which the analyte signal is three times larger than the standard deviation of the blank signal. Below this value, the analyte signal cannot be differentiated from the background noise. It is calculated by dividing the calibration slope by 3 times the standard deviation of the blank signals.
The LOD indicates the presence or absence...
7.1K
End Point Prediction: Gran Plot01:07

End Point Prediction: Gran Plot

A Gran plot is used to predict the equivalence volume or endpoint of a potentiometric or acid-base titration without reaching the endpoint. Typically, titration data is collected as a function of the titrant's volume up to a point less than the equivalence volume and then transformed into a linear format. The straight line is extended to the x-axis, indicating the necessary titrant volume to achieve the equivalence point.
For potentiometric titration, the Gran plot is created by plotting...
588
Histogram01:05

Histogram

The histogram is a graphical representation in the x-y form of data distribution in a data set. The horizontal x-axis is labeled with what the data represents (for instance, distance from your home to school). The vertical y-axis is labeled either frequency or relative frequency (or percent frequency or probability).
A histogram graph consists of contiguous (adjoining) boxes. The heights of the bars correspond to frequency values. The graph will have the same shape with respective labels. The...
14.4K
Midrange01:07

Midrange

A somewhat easy to compute quantitative estimate of a data set’s central tendency is its midrange, which is defined as the mean of the minimum and maximum values of an ordered data set.
Simply put, the midrange is half of the data set’s range. Similar to the mean, the midrange is sensitive to the extreme values and hence the prospective outliers. However, unlike the mean, the midrange is not sensitive to all the values of the data set that lie in the middle. Thus, it is prone to...
3.8K