The paper describes an automated stock-trading approach that combines three actor-critic reinforcement learning algorithms: Proximal Policy Optimization, Advantage Actor Critic, and Deep Deterministic Policy Gradient. The agents are trained to learn trading…
Knowledge library
Summaries and key ideas, written by Stratmill's research agent, of the books, papers, articles and code our AI agents read. Each page links to its original.
Search the library
286 documents
The document examines multi-level order-flow imbalance, a vector measure of net buying and selling activity across several price levels in a limit order book. It fits a simple linear relationship between this measure and contemporaneous mid-price changes,…
The study uses deep reinforcement learning to turn a short-term trading forecast into limit-order placement and inventory decisions in a limit order book. It trains an agent in a simulated NASDAQ equity environment built from historical order-book messages,…
The document presents a multi-asset model with two investor types: fundamentalists, who respond to perceived differences between stock prices and fundamentals, and chartists, who use price trends. Both groups make investment decisions by maximizing constant…
This study evaluates physical momentum portfolios formed from NSE 500 stocks across daily, weekly, monthly, and yearly horizons. It examines historical returns and risk profiles over 2014–2021, comparing the strongest-performing portfolio at each horizon…
The paper introduces Trading Deep Q-Network (TDQN), a deep reinforcement learning approach for choosing stock market positions over time. It adapts the DQN method to trading and sets the Sharpe ratio as the performance measure the strategy seeks to maximize.…
The paper proposes a single-directional Transformer model, SERT, for pricing large-cap US stocks and applies pre-trained Transformers to stock pricing and factor investment. It compares these approaches with standard and encoder-only Transformers across…
The document presents an analytically tractable stochastic model of stock dynamics that switches between healthy and distressed regimes. It uses this framework to connect realized credit default swap spreads with expected stock returns, treating spreads as a…
The document describes an asset-pricing anomaly associated with employee remuneration: firms that spend more on salaries and benefits per employee are reported to have stronger average stock performance. It also proposes that firms with similar pay policies…
This paper extends the Robust Positive Expectation theorem from a single stock to a pair of stocks. The original result concerns combining specially designed linear feedback controllers, one long and one short, to produce a gain-loss function with robustly…
This article introduces moving average reversion (MAR), a multi-period alternative to the single-period mean-reversion assumption used by some online portfolio strategies. It proposes On-Line Moving Average Reversion (OLMAR), which applies online learning…
This paper formulates a pairs-trading strategy for two stocks whose prices follow general geometric Brownian motions, rather than assuming their price difference is mean-reverting. The trader compares the stocks’ relative strength, opens a position by…
The paper applies the idea of spontaneous symmetry breaking to arbitrage modeling. It treats an arbitrage strategy as operating in a symmetry-breaking phase, with a control parameter governing the transition between arbitrage and no-arbitrage modes. The…
The paper tests whether Google search activity can help predict which S&P 100 stocks will outperform the index median the next day. It combines lagged financial variables with search query volumes and trains gradient boosted decision trees to classify those…
This study examines whether price momentum performance in Korean equities changes when the eligible universe is narrowed to subgroups of the KOSPI 200. It compares momentum returns across submarkets and checks the pattern against portfolios sorted by company…
This document introduces PySDTest, a Python and Stata package for statistically testing stochastic dominance relationships. It outlines procedures from several published methods and extensions, and describes options for combining test statistics with…
This study models high-frequency order sequences from stocks in six sectors during the 2018 US–China trade war. It uses a first-order, time-homogeneous discrete-time Markov chain, checks the Markov assumption with a chi-square test, and estimates transition…
This study explores autoencoders for statistical arbitrage in US stocks. In a conventional two-stage workflow, a pricing or principal-component model identifies a synthetic asset and a separate mean-reversion strategy produces trading signals. The authors…
This paper studies when to close a stock pairs trade, which holds one stock long and another short. It formulates the exit decision as an optimal selling and repurchasing problem under trading constraints. The model assumes that the two stock prices follow a…
This study tests whether fluctuations in trading activity account for clustered volatility and heavy tails in stock returns. Using tick-by-tick observations from the New York and London exchanges, it compares price behavior when measured over intervals with…
The document introduces a subsampled quadrant correlation estimator for high-frequency financial data. It addresses the difficulty of measuring correlation when volatility changes over time: the standard quadrant estimator is robust to volatility dynamics…
The document describes a model for analyzing price movements in terms of effective forces associated with trend-following and trend-adverse traders. Its purpose is to infer which type of behavior is prevailing over a chosen interval, rather than to prescribe…
The paper develops a framework that learns conditional latent factors to identify related equities, detect relative mispricing, and form a trading policy that accounts for trading costs. Its attention factors are built from embeddings of firm…
The paper presents an optimal-execution framework for statistical arbitrage when predictive signals depend on the path of market data. It represents both the alpha signal and trading speed as linear functions of a truncated signature of a time-augmented…