---
title: Binary-Choice Kelly Criterion
url: https://www.emergentmind.com/topics/binary-choice-kelly-criterion
type: topic
---

# Binary-Choice Kelly Criterion

The binary-choice Kelly criterion is the log-optimal staking rule for repeated binary gambles with multiplicative wealth dynamics. In its classical form, each trial has two outcomes—win with probability $p$ and loss with probability $q=1-p$—and a bettor stakes a fraction of current wealth so as to maximize expected logarithmic growth. For net odds $b>0$, staking a fraction $f$ multiplies wealth by $1+bf$ on a win and by $1-f$ on a loss, yielding the objective
\[
g(f)=p\ln(1+bf)+q\ln(1-f).
\]
Under the usual no-shorting/no-leverage restriction $0\le f\le 1$, this problem has a unique concave optimum, the Kelly fraction, which is the canonical solution for long-run geometric growth in binary betting and the discrete prototype for broader log-optimal portfolio theory [2503.17927].

## 1. Formal model and closed-form solution

In the binary-choice setup, the one-period return $R$ takes values $b$ with probability $p$ and $-1$ with probability $q$. Wealth after $n$ rounds is
\[
W_n=W_0\prod_{i=1}^n (1+fR_i),
\]
and the per-trial log-return is
\[
X=\ln(1+bf)\quad\text{with probability }p,\qquad X=\ln(1-f)\quad\text{with probability }q.
\]
The expected log-growth is therefore
\[
g(f)=p\ln(1+bf)+q\ln(1-f).
\]

The objective is strictly concave because
\[
g''(f)=-\,p\frac{b^2}{(1+bf)^2}-q\frac{1}{(1-f)^2}<0.
\]
The first-order condition
\[
g'(f)=p\frac{b}{1+bf}-q\frac{1}{1-f}=0
\]
has the unique interior solution
\[
f^*=\frac{pb-q}{b}=\frac{p(b+1)-1}{b}.
\]
Under no shorting and no leverage, the admissible Kelly stake is the truncation of this expression to $[0,1]$:
\[
f^{\mathrm{Kelly}}=\min\{1,\max\{0,f^*\}\}.
\]
The fair-game boundary is $p=1/(b+1)$, for which $f^*=0$; if $p\le 1/(b+1)$, the optimal action under the no-shorting/no-leverage constraint is not to bet. At even net odds, $b=1$, the formula reduces to the familiar
\[
f^*=2p-1
\]
[2503.17927].

Several equivalent parameterizations recur in the literature. With decimal odds $o=b+1$,
\[
f^*=\frac{po-1}{o-1}.
\]
In prediction-market notation, if the market-implied probability is $\pi=1/(1+b)$, then
\[
f^*=\frac{p-\pi}{1-\pi}.
\]
These are algebraic restatements of the same binary Kelly rule [2107.08827].

The same closed form appears in even-money Bernoulli games written with outcome variable $\mathscr{Z}(I)\in\{-1,1\}$ and fixed stake fraction $\mathcal{F}$. There the per-trial log-utility
\[
\mathsf{U}(\mathcal{F},p)=p\log(1+\mathcal{F})+q\log(1-\mathcal{F})
\]
is globally concave on $(0,1)$, and its maximizer is
\[
\mathcal{F}_K=p-q=2p-1
\]
[2502.16859].

## 2. Optimal growth, edge, and information-theoretic structure

Substituting the optimal fraction into the growth objective gives a closed-form maximum. Using
\[
1+bf^*=p(b+1),\qquad 1-f^*=\frac{q(b+1)}{b},
\]
one obtains
\[
g(f^*)=p\ln\big(p(b+1)\big)+q\ln\!\left(\frac{q(b+1)}{b}\right)
      =\ln(b+1)+p\ln p+q\ln q-q\ln b.
\]
This expression makes the edge threshold explicit: positive long-run growth requires favorable odds in the sense $pb>q$, equivalently $p(b+1)>1$ [2503.17927].

In the even-money case, the optimal growth simplifies to an entropy identity. Evaluated at $\mathcal{F}_K=2p-1$,
\[
\mathsf{U}(\mathcal{F}_K,p)=p\log(2p)+q\log(2q)=\log 2-H(p),
\]
where
\[
H(p)=-p\log p-q\log q
\]
is the binary Shannon entropy. Thus the fair case $p=\tfrac12$ yields zero expected geometric growth, while a biased game with $p>\tfrac12$ yields $\mathsf{U}(\mathcal{F}_K,p)>0$ [2502.16859].

A related information-theoretic formulation appears in the horse-race framework. For the binary special case, the optimal expected doubling rate can be written as
\[
g(f^*)=\log_2(1+b)+p\log_2 p+q\log_2 q-q\log_2 b,
\]
and the use of side information raises the optimal doubling rate by the pragmatic information of the messages. In Weinberger’s formulation, the increase in doubling rate equals the mutual-information term associated with conditioning winning probabilities on messages, tying binary Kelly growth directly to information usage in sequential decision-making [0903.2243].

This information-theoretic viewpoint reappears in prediction markets. With Kelly traders on a binary event, the market-clearing price is a wealth-weighted average of beliefs, and wealth updates proportionally to likelihood ratios after each outcome. In that setting, the market’s cumulative log loss is within $\ln(1/w_i)$ of the best participant’s, so the binary Kelly mechanism acts as both a staking rule and an aggregation rule for probabilistic beliefs [1201.6655].

## 3. Volatility, asymptotic variance, and fractional Kelly

The standard Kelly solution maximizes asymptotic log-growth, but the associated stake is often regarded as aggressive. A recent variance-based treatment introduces the asymptotic variance of per-trial log-growth as a unified risk descriptor in the binary setting:
\[
\upsilon(f)=\mathrm{Var}[X]
          =pq\Big(\ln(1+bf)-\ln(1-f)\Big)^2.
\]
Because
\[
\frac{d}{df}\Big(\ln(1+bf)-\ln(1-f)\Big)=\frac{b}{1+bf}+\frac{1}{1-f}>0,
\]
$\upsilon(f)$ is strictly increasing on $(0,1)$. Thus larger Kelly fractions increase asymptotic volatility monotonically [2503.17927].

For i.i.d. trials, the law of large numbers and central limit theorem imply
\[
\frac{1}{n}\sum_{i=1}^n X_i\to g(f)\quad\text{a.s.},
\]
and
\[
\frac{\sum_{i=1}^n X_i-n\,g(f)}{\sqrt{n}}\Rightarrow \mathcal N(0,\upsilon(f)).
\]
Equivalently,
\[
\ln W_n\approx n\,g(f)+\sqrt{n}\,\sigma(f)\,Z,\qquad \sigma(f)=\sqrt{\upsilon(f)},\ Z\sim\mathcal N(0,1).
\]
This approximation quantifies fluctuations around mean log-growth and underlies asymptotic risk measures [2503.17927].

Two such measures have been proposed. The asymptotic Sharpe ratio is
\[
SR(f)=\frac{g(f)}{\sqrt{\upsilon(f)}}
     =\frac{p\ln(1+bf)+q\ln(1-f)}
            {\sqrt{pq}\,(\ln(1+bf)-\ln(1-f))},
\]
which measures mean log-growth per unit asymptotic volatility of log-growth. The asymptotic ridge coefficient is
\[
Ri(f,\gamma)=g(f)-\gamma\,\upsilon(f),\qquad \gamma\ge 0,
\]
which penalizes aggressive sizing via asymptotic variance. In the binary case, the corresponding first-order condition is
\[
g'(f)-\gamma\,\upsilon'(f)=0,
\]
or explicitly,
\[
p\frac{b}{1+bf}-q\frac{1}{1-f}
-2\gamma pq\big(\ln(1+bf)-\ln(1-f)\big)\left(\frac{b}{1+bf}+\frac{1}{1-f}\right)=0.
\]
Under no shorting and no leverage, this admits a unique solution in $[0,f^*]$, producing a disciplined fractional Kelly fraction that decreases as $\gamma$ increases [2503.17927].

The widely used ad hoc variant is fractional Kelly:
\[
f=\delta f^*,\qquad 0<\delta<1.
\]
Then
\[
g(\delta f^*)=p\ln(1+b\delta f^*)+q\ln(1-\delta f^*)
\]
is strictly increasing and strictly concave in $\delta$, with maximum at $\delta=1$, while
\[
\upsilon(\delta f^*)=pq\Big(\ln(1+b\delta f^*)-\ln(1-\delta f^*)\Big)^2
\]
is strictly increasing in $\delta$. The data further note that reducing $\delta$ decreases growth linearly but reduces asymptotic volatility more than linearly. This motivates variance targeting by solving $\upsilon(\delta f^*)=v_0$ or ridge sizing by solving the penalized first-order condition [2503.17927].

A simpler quadratic approximation also appears in the control-theoretic literature. Expanding $\log(1+x)\approx x-\tfrac12 x^2$ yields the approximate optimizer
\[
\hat f=\frac{\mathbb E[R]}{\mathbb E[R^2]}
      =\frac{pb-q}{pb^2+q}.
\]
At even odds, $b=1$, this approximation coincides exactly with the true Kelly fraction; for general $b$ it differs, often shrinking the stake in favorable cases with $b>1$ [2004.14048].

## 4. Finite-horizon moments, martingale regimes, and drawdown approximations

The binary-choice Kelly criterion is asymptotic in its primary objective, but several finite-horizon quantities are available in closed form. In the even-money Bernoulli model with fixed fraction $\mathcal F$,
\[
\mathscr W(N)=\mathscr W(0)\prod_{I=1}^N (1+\mathcal F\,\mathscr Z(I)),
\]
and, equivalently, if $U$ wins and $V$ losses occur with $U+V=N$,
\[
\mathscr W(N)=\mathscr W(0)\,(1+\mathcal F)^U(1-\mathcal F)^V.
\]
The expected wealth factor is
\[
\mu=1+\mathcal F(2p-1),
\]
so
\[
\mathbb E[\mathscr W(N)]=\mathscr W(0)\,\mu^N.
\]
At $\mathcal F=\mathcal F_K=2p-1$,
\[
\mathbb E[\mathscr W(N)]=\mathscr W(0)\big(1+(2p-1)^2\big)^N,
\]
which grows exponentially for $p\neq \tfrac12$ [2502.16859].

The second moment is
\[
\mathbb E[\mathscr W(N)^2]=\mathscr W(0)^2\Big(p(1+\mathcal F)^2+q(1-\mathcal F)^2\Big)^N,
\]
and the variance is
\[
\mathrm{VAR}(\mathscr W(N))
=\mathscr W(0)^2\left(\Big[p(1+\mathcal F)^2+q(1-\mathcal F)^2\Big]^N-\big[1+\mathcal F(2p-1)\big]^{2N}\right).
\]
For small $\mathcal F$ and large $N$,
\[
\mathrm{VAR}(\mathscr W(N))\approx 2\,\mathscr W(0)^2\,N\,p(1-p)\,\mathcal F^2.
\]
This provides a finite-horizon volatility proxy distinct from asymptotic log-growth variance [2502.16859].

The same paper partitions the stake domain using the sign of the per-trial log-utility. Define $\mathcal F_*$ by
\[
\mathsf U(\mathcal F_*,p)=0
\quad\Longleftrightarrow\quad
(1+\mathcal F_*)^p(1-\mathcal F_*)^q=1.
\]
Then, for $p>\tfrac12$, the wealth process is a submartingale for $\mathcal F\in[0,\mathcal F_*)$, a martingale at $\mathcal F=\mathcal F_*$, and a supermartingale for $\mathcal F\in(\mathcal F_*,1]$. For small edges, the threshold obeys the approximation
\[
\mathcal F_*\approx 2\,\mathcal F_K=2(2p-1).
\]
This regime split clarifies that positive expected log-growth and positive expected wealth drift are not identical notions, but here the sign of $\mathsf U$ determines the martingale classification stated in the paper [2502.16859].

At a fixed finite horizon $n$, the central-limit approximation for log-wealth yields an approximate shortfall probability. For a log drawdown threshold $D>0$ relative to mean,
\[
\mathbb P\!\left(\ln W_n-n\,g(f)\le -D\right)\approx
\Phi\!\left(-\frac{D}{\sqrt{n\,\upsilon(f)}}\right).
\]
The same source explicitly notes the limitations: this is asymptotic, ignores path dependence and absorbing constraints, and exact finite-$n$ behavior is a binomial mixture over two log-values rather than Gaussian [2503.17927].

## 5. Sensitivity, misspecification, and practical risk control

A central practical issue is sensitivity to estimation error. In the binary net-odds model,
\[
\frac{\partial f^*}{\partial p}=\frac{b+1}{b},
\qquad
\frac{\partial f^*}{\partial b}=-\frac{q}{b^2}
\]
in the asymptotic-variance treatment, while another source reports for the same binary formula and $a=1$ that better odds increase $f^*$ with
\[
\frac{\partial f^*}{\partial b}=\frac{q}{b^2}>0.
\]
The first derivative with respect to $p$ agrees across treatments: small errors in $p$ can produce proportionally larger errors in the stake, especially at even odds where $\partial f^*/\partial p=2$ [2503.17927; 2002.03448]. Near the fair boundary $p=1/(b+1)$, small upward errors in $p$ can move the prescribed stake from zero to materially positive values, while small downward errors truncate the optimal action to no bet under no shorting/no leverage [2503.17927].

A second-order local regret formula quantifies the cost of probability misspecification. If $p$ is estimated as $\hat p$ and $\hat f=f^*(\hat p)$, then
\[
\Delta g\approx \frac12\,\frac{(\Delta p)^2}{p(1-p)},
\qquad \Delta p=\hat p-p.
\]
In that approximation, the local growth loss depends primarily on the curvature in $p$, not directly on $b$, although the feasible stake and boundary exposure still do depend on $b$ [2508.18868].

Several practical mitigations recur across the literature. Fractional Kelly is the simplest:
\[
f_\alpha=\alpha f^*,\qquad 0<\alpha<1.
\]
It sacrifices some expected log-growth in exchange for reduced variance and greater robustness to parameter error [2002.03448]. In empirical sports-betting experiments, an adaptive variant of fractional Kelly was tuned by grid search to maximize median final wealth subject to the safety condition $Q5>0.9$ of final wealth, and this KellyFrac variant achieved zero ruin across horse racing, basketball, and football datasets in the reported study [2107.08827].

Other risk-control schemes include maximum bet caps, drawdown-constrained Kelly, and distributionally robust Kelly. The drawdown-constrained formulation in the sports-betting review uses the convex approximation
\[
\mathbb E[(O\cdot f)^{-\lambda}]\le 1,
\qquad
\lambda=\log(\beta)/\log(\alpha),
\]
to limit the probability of falling below a wealth threshold. Distributionally robust Kelly replaces the estimated probability vector by a box ambiguity set and solves a worst-case log-utility problem over that set [2107.08827]. These are not part of the classical binary-choice formula, but they arise as practical corrections when the formal assumptions behind full Kelly are considered unrealistic.

The option-based binomial extension pursues robustness differently. In a binomial stock–bond model with a European put, the paper proves that a fairly priced option does not improve log-optimal growth under correct specification, but a convex mixture of two Kelly-with-option portfolios can be robust to estimation risk in the “max-of-two” sense: asymptotically, the mixture achieves the better realized growth of the two constituent option-augmented Kelly portfolios under the true environment [2508.18868]. This suggests that robustness can be engineered structurally rather than only by scalar stake shrinkage.

## 6. Generalizations and adjacent formulations

The binary-choice Kelly criterion is the base case for several generalizations. If payoffs on wins are random rather than constant, with nonnegative payoff variable $B$, the expected log-growth becomes
\[
G(f)=p\,\mathbb E[\ln(1+fB)]+q\,\ln(1-f),
\]
and the optimal fraction solves the fundamental integral equation
\[
p\,\mathbb E\!\left[\frac{B}{1+f^*B}\right]-\frac{q}{1-f^*}=0.
\]
The resulting optimal fraction is smaller than the classical constant-payoff Kelly fraction computed using the average payoff $\mathbb E[B]$, with equality only in the degenerate constant-payoff case [1411.3615]. A plausible implication is that payoff uncertainty itself acts as an endogenous conservatism mechanism through the concavity of $\ln(1+fB)$.

Another extension replaces log utility by a general concave utility function and allows extraneous wealth. In a one-step binary gamble, the interior first-order condition becomes
\[
p\,b\,U'(W_{\mathrm{ext}}+W(1+fb))
=
q\,U'(W_{\mathrm{ext}}+W(1-f)).
\]
For log utility this yields
\[
f^*=\frac{W_{\mathrm{ext}}+W}{bW}(bp-q),
\]
while CRRA and CARA utilities produce different closed-form stakes. The same work shows that, in an IID binomial tree with concave utility, the optimal local action at a node depends only on the node’s state and primitives, enabling an $O(n^2)$ dynamic program rather than exponential path enumeration [1611.09130].

The continuous-time analogue arises in diffusion models with return process $R_t=\mu t+\sigma B_t$. There the log-growth and asymptotic variance are
\[
g_R(f)=f\mu-\tfrac12 f^2\sigma^2,
\qquad
\upsilon_R(f)=f^2\sigma^2,
\]
the continuous-time Kelly fraction is $f^*=\mu/\sigma^2$ (or $(\mu-r)/\sigma^2$ with risk-free rate $r$), and the ridge-optimal fraction becomes
\[
f^{Ri}=\frac{f^*}{1+2\gamma}.
\]
The discrete binary case and diffusion case share the same qualitative structure—concave growth objective, volatility increasing with stake, and fractionalization under variance penalization—although the binary model retains nonlinear wealth multipliers and hard boundaries under no shorting/no leverage [2503.17927; 2002.03448].

Finally, the single-bet binary formula is no longer sufficient when multiple binary bets are placed simultaneously. In the multivariate case with $N$ simultaneous wagers, the Kelly objective involves the joint outcome space of size $2^N$. Recent work replaces explicit enumeration by an integral-transform formulation for independent bets, reducing objective evaluation from $O(2^N)$ to $O(N)$ per quadrature node, and studies lower and upper bounds via decomposition subproblems [2604.24723]. This establishes the single binary criterion as the tractable primitive from which higher-dimensional Kelly optimization departs.

## 7. Prediction markets, forecast evaluation, and interpretation

In binary prediction markets with Kelly bettors, the criterion determines both individual trade size and market aggregation. If agent $i$ has belief $p_i$ and normalized wealth $w_i$, the competitive equilibrium price is
\[
p_m=\sum_i w_i p_i.
\]
After a realized outcome, wealth updates proportionally to likelihood ratios:
\[
w_i'=\frac{p_i}{p_m}w_i\quad\text{if }y=1,
\qquad
w_i'=\frac{1-p_i}{1-p_m}w_i\quad\text{if }y=0.
\]
Thus wealth reweighting is exactly Bayesian, and the next-period market price is the posterior wealth-weighted average of beliefs. With fractional Kelly, the equilibrium price becomes a wealth-and-confidence-weighted average,
\[
p_m=\frac{\sum_i \lambda_i w_i p_i}{\sum_i \lambda_i w_i},
\]
and each trader behaves as if holding the effective belief
\[
p_i'=\lambda_i p_i+(1-\lambda_i)p_m
\]
[1201.6655].

A newer forecast-evaluation framework treats each probabilistic model as a canonical Kelly bettor whose bankroll evolves in real time as probabilities are updated before a single binary outcome resolves. For a model that already holds win shares $w$ and cash bankroll $B$, the generalized Kelly fraction is
\[
f=p-\frac{1-p}{b}\left(1+\frac{w}{B}\right).
\]
Cash updates as
\[
B_t=B_{t-1}(1-f_t),
\]
win shares update as
\[
w_t=w_{t-1}+(1+b_t)f_tB_{t-1},
\]
and the market-clearing consensus probability in the binary case is
\[
m_t=\frac{\sum_i p_{i,t}B_{i,t}}{1-\sum_i p_{i,t}w_{i,t}}.
\]
The framework interprets
\[
\text{credibility}_{i,t}=B_{i,t}+m_t w_{i,t}
\]
as real-time model credibility and emphasizes a Bayesian analogue in which bankroll is a proxy for posterior credibility [2602.09982].

Simulation results reported in that study show that Kelly-bankroll evaluation can distinguish correct from incorrect time-updating binary models more accurately than average log-loss or Brier score in several settings, including faulty recency bias and random-walk miscalibration [2602.09982]. This suggests that the binary Kelly criterion functions not only as a staking prescription but also as an evaluation functional for sequential probabilistic forecasts, especially when the timing and confidence of updates matter.

Across these formulations, the central structure remains unchanged: a binary event, multiplicative wealth, a log-based objective, and an optimal fraction determined by the discrepancy between subjective probability and implied odds. The classical formula
\[
f^*=\frac{pb-q}{b}
\]
is therefore both a specific betting rule and the seed from which broader theories of risk-sensitive growth, Bayesian aggregation, and probabilistic forecast evaluation are constructed.

Source: https://www.emergentmind.com/topics/binary-choice-kelly-criterion