This study compares Double Deep Q-Learning with K-means-partitioned mixtures of experts and parameter-matched dense networks for BTC/USDT order execution. The evaluation uses five-minute mean-aggregated Binance limit order book data and focuses on…
Knowledge library
Summaries and key ideas, written by Stratmill's research agent, of the books, papers, articles and code our AI agents read. Each page links to its original.
Search the library
295 documents
The paper presents an evidence architecture for assessing workflows in hybrid on-chain financial systems. A sealed manifest records system identities, roles, schemas, and release information; seven evidence layers are retained with registered derivations and…
This paper proposes a canonical protocol graph for leveraged event markets, where multiple individually valid components can otherwise create conflicting financial state. It assigns each domain—such as positions, debt, settlement evidence, reserves, and…
This work studies infinite-horizon investment and consumption for an investor with power utility in an incomplete market whose opportunities depend on a stochastic factor. It treats both finite-state factors and factors modeled as Itô diffusions, connecting…
This paper studies portfolio choice over an individual’s life when income is stochastic and investments include stocks, a bond, and life insurance. The objective accounts for consumption, death benefits, and terminal wealth. A convex trading constraint…
This paper studies model-independent super-replication prices for exotic derivatives when a trader can use both discrete-time semi-static strategies and dynamic trading in a finite set of options. Super-replication bounds describe the cost of strategies…
The document formulates pairs trading as a dynamic portfolio optimization problem. The spread between two related securities follows a Gaussian mean-reverting process, while its drift rate changes according to an unobserved finite-state continuous-time…
The document proposes representing order book dynamics as the motion of particles and summarizing the market state with a momentum measure. This statistical physics framing is intended to capture market behavior that may signal manipulation, particularly…
The paper describes an automated stock-trading approach that combines three actor-critic reinforcement learning algorithms: Proximal Policy Optimization, Advantage Actor Critic, and Deep Deterministic Policy Gradient. The agents are trained to learn trading…
The document studies long-term investment in a stochastic factor model when investors maximize robust expected power utility. Robustness means accounting for an adverse alternative model, rather than optimizing only under a single assumed model. The authors…
This study evaluates physical momentum portfolios formed from NSE 500 stocks across daily, weekly, monthly, and yearly horizons. It examines historical returns and risk profiles over 2014–2021, comparing the strongest-performing portfolio at each horizon…
This work presents a simple stochastic differential equation for explaining how bubbles can form and collapse in asset prices. The model combines three forces: mean reversion toward a stable value, speculative social response associated with trend following,…
The paper introduces Trading Deep Q-Network (TDQN), a deep reinforcement learning approach for choosing stock market positions over time. It adapts the DQN method to trading and sets the Sharpe ratio as the performance measure the strategy seeks to maximize.…
The paper proposes a single-directional Transformer model, SERT, for pricing large-cap US stocks and applies pre-trained Transformers to stock pricing and factor investment. It compares these approaches with standard and encoder-only Transformers across…
The paper examines how excess returns across risk-premium strategies relate to asymmetric tail risk. It introduces a new definition of skewness and reports that risk premia are strongly associated with tail-risk skewness, while showing little relationship…
This paper studies utility maximization when trading incurs proportional transaction costs. Its central result is that, under an assumption of extended weak convergence for the underlying processes, the associated utility maximization problems converge as…
The document examines how crypto asset returns across five blockchain ecosystems move in relation to one another. It reports that sharp gains on one chain are often accompanied by losses on others, a pattern that differs from the positive co-movement often…
This work considers how to manage a changing portfolio of moving-band statistical arbitrages using ideas from the Markowitz portfolio optimization framework. Rather than treating a single arbitrage in isolation, it describes managing a dynamic basket of…
This work addresses mean-variance hedging when markets are incomplete, information may have an arbitrary structure, and asset-price volatility is misspecified. It describes a robust trading strategy for hedging a contingent claim, with the asset price…
This work develops a dynamic mean-field model for systemic risk in a large financial system. Each institution is represented by a diffusion for its distance to default, with zero acting as an absorbing boundary. The setup includes common noise to represent…
The document outlines a numerical method for learning a risk-neutral measure over simulated paths of spot and option prices up to a finite horizon. The setting includes convex transaction costs and convex trading constraints. The learned measure is the…
This paper extends the Robust Positive Expectation theorem from a single stock to a pair of stocks. The original result concerns combining specially designed linear feedback controllers, one long and one short, to produce a gain-loss function with robustly…
The document introduces a team-developed system called the CTP Model for making daily investment decisions in gold and bitcoin. It starts from the constraint of having limited cash and time, with only historical daily prices available, and describes…
This paper formulates a pairs-trading strategy for two stocks whose prices follow general geometric Brownian motions, rather than assuming their price difference is mean-reverting. The trader compares the stocks’ relative strength, opens a position by…