mirror of
https://github.com/theodore-song/polymarket-analyst.git
synced 2026-08-24 04:58:08 +00:00
Replace rejected sports pilot with forward shadow learner
This commit is contained in:
@@ -15,12 +15,12 @@ https://polymarket-site-eta.vercel.app/personal.html
|
||||
|
||||
The site fetches live Polymarket markets, generates agent suggestions, lets you
|
||||
run frequent paper cycles, and syncs the shared arena state through Neon or
|
||||
Vercel Blob. Build 72 also installs an offline app shell and caches timestamped
|
||||
Vercel Blob. Build 73 also installs an offline app shell and caches timestamped
|
||||
market snapshots. During an outage, cycles continue locally; cached entries are
|
||||
allowed for 90 minutes, older snapshots become mark-only, and all cached data
|
||||
expires after 24 hours.
|
||||
|
||||
Build 72 distinguishes a temporary order-book pause from settlement. Exact
|
||||
Build 73 distinguishes a temporary order-book pause from settlement. Exact
|
||||
market refreshes still mark paused positions to the latest published price, but
|
||||
the engine cannot simulate a stop, policy exit, or settlement while
|
||||
`acceptingOrders` is false and `closed` is still false. Only a closed market
|
||||
@@ -32,7 +32,7 @@ team names, Over/Under, or another pair are rejected instead of being silently
|
||||
reinterpreted as Yes/No. The same semantic check applies to complete event
|
||||
bundles and the offline evaluators.
|
||||
|
||||
Build 72 ranks the competition by each agent's return since Strategy 57 began.
|
||||
Build 73 ranks the competition by each agent's return since Strategy 58 began.
|
||||
Historical replay equity remains visible for context, but it no longer makes an
|
||||
agent look like the current leader when the live adaptive strategy is losing.
|
||||
|
||||
@@ -43,7 +43,7 @@ The learner shrinks small samples toward neutral, caps sizing changes to
|
||||
15% of candidates for deterministic exploration so a stale regime cannot become
|
||||
permanent.
|
||||
|
||||
Strategy 57 treats each binary stake as capable of falling to zero even when the
|
||||
Strategy 58 treats each binary stake as capable of falling to zero even when the
|
||||
18% stop cannot fill. New core positions are capped at 2.5%-4% of equity and
|
||||
aggressive positions at 3%-5%, with lower limits for near-term, extreme-price,
|
||||
reversal, and fast-moving setups. Oversized positions inherited from older
|
||||
@@ -66,8 +66,8 @@ cohort can demote it.
|
||||
The Build 69 re-audit loaded all 500 requested histories with no failures. The
|
||||
six-hour family lost 1.27% net on average across 106 independent events, with
|
||||
its full 90% interval below zero; no tested rule was robustly positive at 6,
|
||||
12, 24, or 72 hours. Build 72 therefore uses six hours only to stop bad regimes
|
||||
sooner. It also closes every stale pre-Strategy-57 directional holding at the
|
||||
12, 24, or 72 hours. Build 73 therefore uses six hours only to stop bad regimes
|
||||
sooner. It also closes every stale pre-Strategy-58 directional holding at the
|
||||
next fresh mark, including legacy records missing a signal label, while leaving
|
||||
complete arbitrage bundles and paired maker inventory under their own accounting.
|
||||
|
||||
@@ -98,15 +98,16 @@ evidence proves an edge.
|
||||
Run `npm run evaluate:sports-favorites` for the separate pregame favorite audit.
|
||||
It anchors decisions to the published game start, rejects stale prices, charges a
|
||||
modeled five-cent cost, takes only the highest-priced eligible favorite per event,
|
||||
and uses a chronological 60/20/20 split. The clean 3,000-market run produced nine
|
||||
train-pass rules but zero strict validation selections. The exact 24-hour,
|
||||
60%-85% rule nevertheless had positive point estimates in all three partitions:
|
||||
2.97% train, 3.10% validation, and 19.83% untouched holdout, while its train and
|
||||
validation confidence bounds still crossed zero. Strategy 57 therefore permits
|
||||
only a labeled paper-capital pilot in Favorite Backer: 1.25% of equity per event,
|
||||
6% total, with a five-cent cost booked immediately. It disables after 12 completed
|
||||
events with materially negative evidence or a 2.5% account loss, and scales to
|
||||
2.5% only after 30 events produce a positive 90% lower confidence bound.
|
||||
and uses a chronological 60/20/20 split. The refreshed August 20 run loaded all
|
||||
3,000 histories with no failures and found zero train-pass rules. The Strategy 57
|
||||
24-hour, 60%-85% capital rule averaged -11.98% across 168 independent training
|
||||
events and -4.04% across 100 untouched holdout events after the five-cent cost,
|
||||
so Strategy 58 retires it at the next fresh executable mark. The narrower 12-hour,
|
||||
60%-75% rule had positive point estimates in train, validation, and holdout but
|
||||
did not establish a positive lower confidence bound in train or holdout. Favorite
|
||||
Backer therefore records it with zero capital, one market per event, and persists
|
||||
the pending and completed forward ledger offline. Only 30 new independent closed
|
||||
events with a positive 90% lower confidence bound can promote 1.25% positions.
|
||||
|
||||
Run `npm run evaluate:settlement-calibration` for the stricter settlement-bias
|
||||
search across up to 5,000 resolved markets. It uses a 60/20/20 chronological
|
||||
@@ -115,10 +116,10 @@ stability windows, a 24-hour market-age minimum, and a recent non-flat price
|
||||
history requirement. Before those activity and overlap controls, four sports
|
||||
rules appeared to pass holdout because correlated props shared one event and
|
||||
some histories contained inactive default prices. After correction, 11 of
|
||||
1,400 rules passed training and zero passed validation. Strategy 57 therefore
|
||||
1,400 rules passed training and zero passed validation. Strategy 58 therefore
|
||||
does not install any other static side, category, price-band, or settlement-horizon bet.
|
||||
|
||||
Strategy 57 also removes the last emotion-driven sizing path. Agent mood and
|
||||
Strategy 58 also removes the last emotion-driven sizing path. Agent mood and
|
||||
leaderboard urgency remain visible in reports, but neither can increase capital.
|
||||
A positive peer signal receives at most a 5% sizing lift, and only when both the
|
||||
agent's realized-trade cohort and the independent walk-forward market cohort are
|
||||
@@ -139,7 +140,7 @@ and 24-hour horizons. Zero rules passed training, validation, or untouched
|
||||
holdout. At three hours, the broad 0.5-cent quote-gap rule still lost 0.63% per
|
||||
event in holdout; only 0.61% of observations completed both legs while 23.33%
|
||||
produced adverse one-leg inventory. Wider quotes traded less but remained
|
||||
negative. Strategy 57 therefore does not risk paper capital on an unproven
|
||||
negative. Strategy 58 therefore does not risk paper capital on an unproven
|
||||
maker rule.
|
||||
|
||||
Run `npm run evaluate:reward-maker` to stress current reward-qualified books
|
||||
|
||||
Reference in New Issue
Block a user