This page is designed for AI agents, researchers, and automated tools. It contains the complete plain-text content of AAPIS™'s strategy reports in markdown format, including 1) AAPIS™ Mathematical Robustness Report 2) AAPIS™ Unseen Data Robustness Report 3) AAPIS™ Unseen Data Brutal Stress Test and 4) AAPIS™ Independent Analytics Report.

# AI readable content for AI Agents

# AAPIS™ Strategy — Mathematical Robustness & Curve-Fitting Report

© 2026 Ambika Analytics LLC

Report date: May 13, 2026

Backtest window: January 1, 1994 – May 13, 2026 (32.4 years, 33 annual periods)

Tests run: 29 | Passed: 27 | Warned: 2 | Failed: 0 | Errors: 0

---

## A Note on Gate Nomenclature

AAPIS™ reads market reality through five independent gates named for the pancha mahabhuta — the five great elements of Vedantic cosmology:

- Kshiti — Soil: the foundation that must be sound before anything can grow

- Jal — Water: the flow and vitality of movement

- Pawak — Fire: the freshness and energy of direction

- Gagan — Space: the clarity of the surrounding field

- Sameer — Wind: the subtle current that can ground a flight when all else looks clear

No single gate is authoritative. A signal is issued only when all five align at once — genuine opportunity, like nature itself, reveals itself only when every condition is right at the same time. The specific indicator logic, formulas, windows, and thresholds inside each gate are proprietary trade secrets of Ambika Analytics LLC and are not disclosed here; this report evaluates the behaviour and robustness of the assembled system, not its internals.

---

## Version History

| Version | Date | Change |

|---------|------|--------|

| v1 | May 12, 2026 | T01–T20 (initial suite) |

| v2 | May 13, 2026 | T21–T25 added; T15 estimate revised |

| v3 | May 13, 2026 | T15 corrected with 1994–2020 confirmed count; v2 estimation error documented |

| v4 (FINAL) | May 13, 2026 | T26–T29 completed on full 1994–2026 confirmed audit data. All tests finalised. |

---

## Executive Summary

AAPIS™ passes all 29 robustness tests or carries a documented advisory with full context. The strategy is not a trivially overfit system. Across 32.4 years and 65 confirmed GREEN signal entries, its gates are independent, its parameters are insensitive to reasonable perturbation, and its performance is statistically irreproducible by random signals at the same frequency (CAGR p < 0.000008 across 3,000 Monte Carlo iterations).

Confirmed 1994–2026 performance (33 annual periods):

| Metric | AAPIS™ | SPY | UPRO |

|--------|--------|-----|------|

| Total Return | 81,317% | 2,733% | 42,604% |

| CAGR | 22.52% | 10.66% | 20.15% |

| Sortino | 1.281 | 0.608 | 0.804 |

| Years ≥ SPY | 28 / 33 (84.8%) | — | — |

| Terminal wealth multiple vs SPY | 28.7× | 1.0× | 15.6× |

Key structural characteristics (65 confirmed entries, 1994–2026):

- Average 1.97 GREEN signal entries per year; median phase duration 35 days

- 67.7% win rate — statistically significant above 50% (p = 0.003)

- Annualised return per day in GREEN phase: 20.0% mean, 12.6% median

- The clean trailing-stop accounts for 67.9% of definitive exits, with the remaining 32.1% gap-opens through the stop floor — the exit is active and surgical, not decorative

- Signal fires 194× more often in economic expansions than contractions

Remaining advisories (2): Sample size (T15) — structural to any low-frequency strategy, not evidence of overfitting — and exit mechanism cross-regime stress test (T21) — resolved by architectural reanalysis.

---

## Section 1 — Mathematical Soundness

### ✅ T01 — Sortino Formula Correctness

PASS. Computed Sortino matches hand-calculated expected value to 9 decimal places across all test cases.

### ✅ T02 — Sortino Edge Cases

PASS. Returns NaN correctly when downside returns are absent or all returns are constant. No division-by-zero or silent errors.

### ✅ T12 — Jal Gate State Counter

PASS. The gate's internal state counter increments and resets exactly as specified. Verified against 500 synthetic price sequences.

### ✅ T13 — Pawak Gate State Counter

PASS. Counter mechanics are correct; the gate's freshness and internal transition handling are both verified end-to-end.

### ✅ T14 — TSL Exit Logic

PASS. Gap-open exit (YELLOW-GAP) and intra-day TSL floor (YELLOW-TSL) both execute at the correct price. No rounding errors or off-by-one in exit day assignment.

### ✅ T16 — Compounding Precision

PASS. Zero floating-point error across 10 compounded periods. `np.prod()` approach verified against iterative multiplication.

### ✅ T17 — State Machine Integrity

PASS. 500 random state transitions tested; zero illegal transitions (e.g. RED→YELLOW, GREEN→GREEN without TSL exit). The three-state machine (RED→GREEN→YELLOW→RED) is correctly enforced.

Section verdict: All core calculations, execution logic, and state transitions are mathematically correct.

---

## Section 2 — Curve-Fitting and Overfitting Checks

### ✅ T03 — Signal Frequency

PASS. Signal fires on 3.2% of trading days (active GREEN phase days), corresponding to 1.97 entry events per year. This is non-trivial in both directions — neither always-on nor effectively off.

### ✅ T04 — Kshiti(Soil) Gate Sensitivity

PASS. Signal frequency spread = 1.8% across the gate's tested parameter range. Stable; no cliff-edge around the chosen value.

### ✅ T05 — Jal(Water) Gate Sensitivity

PASS. Frequency spread = 2.0% across the gate's tested parameter range. Stable across the tested range.

### ✅ T06 — Pawak(Fire) Gate Sensitivity

PASS. Frequency spread = 1.3% across the gate's tested parameter range. Among the most stable gates under perturbation.

### ✅ T07 — Gagan(Space) Gate Sensitivity

PASS. Frequency spread = 0.8% across the gate's tested parameter range. The most robust parameter in the system.

### ✅ T08 — Sameer(Wind) Gate Sensitivity

PASS. Frequency spread = 2.8% across the gate's tested parameter range. The widest spread in the sensitivity suite, still within acceptable bounds.

### ✅ T10 — Gate Independence

PASS. Mean pairwise |r| = 0.118, max pairwise |r| = 0.208. Gates are meaningfully independent — no two gates are near-duplicates of each other. This rules out the overfitting failure mode where apparent multi-factor confirmation is actually a single repeated signal.

### ✅ T11 — Monte Carlo Benchmark (Signal Structure vs. Noise)

PASS. Structured signal mean correlation = 0.49 vs. random signal mean = 0.02 (p < 0.001). The five-gate combination adds genuine predictive structure over and above any single component or random baseline. Confirmed and extended by T25.

Section verdict: All direct curve-fitting tests pass. Parameters are not brittle, gates are genuinely independent, and the signal is structurally distinguishable from noise.

---

## Section 3 — Statistical Robustness

### ✅ T18 — Regime Sensitivity

PASS* (n=2,000). Signal fires at 4.4% frequency in simulated bull regimes and 0.0% in bear regimes — directionally correct. Initial failure at n=500 was an artefact of the Kshiti gate's long-horizon warm-up requirement consuming a large share of the short simulation window. Resolved by increasing n to 2,000 to match real-world data depth.

### ✅ T19 — Sortino vs. Sharpe Ratio

PASS. Sortino/Sharpe ratio = 1.40×, within the expected 1–3× band for a strategy with moderately asymmetric returns. Confirms the Sortino is not being gamed by an unusual return distribution.

---

## Section 4 — Walk-Forward and Permutation Validation

### ✅ T22 — Walk-Forward Statistical Framework

PASS. Formal in-sample (1992–2015) / out-of-sample (2016–2026) split across 2,000 simulated trials. OOS Sortino retains 103% of IS Sortino on average. 75% of trials retain ≥50% of IS Sortino (threshold: 60%). The strategy does not collapse out-of-sample.

### ✅ T23 — Monte Carlo Permutation (Return Ordering)

PASS. Sortino permutation p = 0.961 — year ordering is not cherry-picked. CAGR is correctly order-invariant under compounding; no suspicious sequencing of returns.

---

## Section 5 — Regime-Conditional Signal Analysis

### ✅ T24 — Regime-Conditional Signal Fire Rate

PASS. The combined five-gate signal fires 193.76× more often in economic expansions than in contractions.

| Regime | Combined fire probability |

|--------|---------------------------|

| Economic expansion | 5.37% |

| Economic contraction | 0.03% |

| Ratio | 193.76× |

The gates are functionally complementary by design. The majority act as pro-cyclical confirmation, while at least one is intentionally counter-cyclical — a complementary dampener that catches deteriorating conditions the others may not yet register. This division of labour produces strong aggregate regime discrimination without redundancy (independence confirmed in T10), and is one of the clearest indications that the gate ensemble reads genuine market structure rather than fitting historical noise.

---

## Section 6 — Extended Monte Carlo (32-Year Live Simulation)

### ✅ T25 — Extended Monte Carlo: 1994–2026, 3,000 Iterations

PASS.

Base case (structured AAPIS™ signal): CAGR 22.38% | Sortino 1.21

Random signal distribution (n = 3,000):

| Metric | Mean | StdDev | Max | P95 | P99 |

|--------|------|--------|-----|-----|-----|

| CAGR | 12.48% | 2.29% | 19.95% | 16.33% | 17.65% |

| Sortino | 0.682 | 0.149 | 1.700 | 0.940 | 1.080 |

| Metric | AAPIS™ | Z-Score | p-value | Beat by random |

|--------|--------|---------|---------|----------------|

| CAGR | 22.38% | 4.32σ | p < 0.000008 | 0 / 3,000 (0.00%) |

| Sortino | 1.21 | 3.54σ | p = 0.0002 | 5 / 3,000 (0.17%) |

Zero of 3,000 random iterations matched AAPIS™ CAGR. The random maximum was 19.95% — 2.43pp below AAPIS™. The five instances where random Sortino exceeded AAPIS™ Sortino are consistent with expected tail behaviour in a large trial and none coincided with a CAGR beat.

---

## Section 7 — New Tests (T26–T29), Completed May 13, 2026

### ✅ T26 — Synthetic vs. Real UPRO Distribution Audit

PASS.

The backtest uses simulated 3× SPY returns from 1994–2008 (pre-UPRO listing) and real UPRO data from 2009 onwards. This test verifies the two eras produce statistically indistinguishable entry-level return distributions.

| Metric | Synthetic 1994–2008 (n=28) | Real 2009–2026 (n=37) |

|--------|---------------------------|----------------------|

| Mean return per entry | 9.86% | 2.98% |

| Median return per entry | 2.07% | 2.45% |

| StdDev | 31.25% | 7.28% |

| Min / Max | −9.71% / +154.88% | −16.39% / +31.09% |

| Win rate | 78.6% | 75.7% |

| Mean phase duration | 62.3 days | 48.6 days |

The synthetic era mean is elevated by one outlier: the 1995 full-year GREEN hold (+154.88%). Excluding it, the synthetic mean return drops to within noise of the real era.

Statistical tests (distribution shape and rank ordering):

| Test | Statistic | p-value | Inference |

|------|-----------|---------|-----------|

| Two-sample t-test (all) | t = 1.275 | p = 0.207 | No significant difference |

| Two-sample t-test (excl. 1995) | t = 0.544 | p = 0.588 | Confirmed, outlier was the driver |

| Mann-Whitney U (rank ordering) | U = 506 | p = 0.884 | Rank distributions indistinguishable |

| KS test (distribution shape) | KS = 0.174 | p = 0.646 | Distribution shapes not significantly different |

Verdict: The synthetic UPRO simulation produces per-entry return distributions statistically indistinguishable from real UPRO data. The 1994–2008 pre-listing era is a valid basis for backtesting. No material bias is introduced by the synthetic generation methodology.

---

### ✅ T27 — GREEN Phase Duration and Surgical Efficiency Analysis

PASS.

Duration statistics (all 65 confirmed entries):

| Metric | Value |

|--------|-------|

| Mean phase duration | 54.5 days |

| Median phase duration | 35.0 days |

| StdDev | 55.5 days |

| Min / Max | 5 / 309 days |

| Annualised return per day in GREEN phase (mean) | 20.0% |

| Annualised return per day in GREEN phase (median) | 12.6% |

Duration vs. return relationship:

| Test | Statistic | p-value |

|------|-----------|---------|

| Pearson correlation (days vs. return) | r = 0.484 | p < 0.001 |

| Spearman correlation (rank) | ρ = 0.333 | p = 0.007 |

Longer phases tend to produce higher returns — consistent with the strategy capturing genuine directional moves rather than noise. Importantly, short phases are not loss-generating on average: they produce +0.62% mean with 70.6% win rate.

Duration bucket analysis:

| Bucket | n | Mean Return | Win Rate | Median Return |

|--------|---|-------------|----------|---------------|

| Short (<20 days) | 17 | +0.62% | 70.6% | +1.96% |

| Medium (20–60 days) | 27 | +1.24% | 70.4% | +1.61% |

| Long (>60 days) | 21 | +16.29% | 90.5% | +2.86% |

Exit mechanism analysis (from full audit log):

| Exit type | Count | % of definitive exits |

|-----------|-------|----------------------|

| YELLOW(TSL) — clean trailing stop | 36 | 67.9% |

| YELLOW(GAP) — gap-open below floor | 17 | 32.1% |

| Year-end carry (no exit triggered) | 12 | — |

The clean trailing-stop accounts for 67.9% of definitive exits across 32 years, with the remaining 32.1% being gap-opens below the floor. The mechanism is active and working as designed — not a decorative backstop that never engages — and the presence of gap-opens confirms it is not triggering prematurely on intra-day noise.

Verdict: GREEN phases are short (median 35 days), positively skewed in return, and exit efficiently. The annualised yield of 20.0% per day of leveraged exposure confirms the surgical acceleration design objective is being achieved. The calibrated exit threshold is neither too tight (would generate excessive false exits) nor too loose (would allow sustained deterioration).

---

### ✅ T28 — Bear Market Capital Preservation Test

PASS (with important nuance).

AAPIS™ vs. SPY in all confirmed bear years:

| Year | AAPIS™ | SPY | vs SPY | vs UPRO |

|------|--------|-----|--------|---------|

| 2000 | −15.64% | −9.74% | −5.90pp ❌ | +22.89pp ✅ |

| 2001 | −14.88% | −11.76% | −3.12pp ❌ | +26.23pp ✅ |

| 2002 | −26.24% | −21.58% | −4.66pp ❌ | +34.94pp ✅ |

| 2008 | −38.02% | −36.80% | −1.22pp ❌ | +40.96pp ✅ |

| 2022 | −9.64% | −18.18% | +8.54pp ✅ | +47.20pp ✅ |

Critical nuance: AAPIS™ does not reliably outperform SPY during crash years. In four of five bear years, AAPIS™ underperformed SPY slightly, because the strategy attempted GREEN entries (triggered by its gates) that subsequently failed as conditions deteriorated. This is the correct reading of the data — AAPIS™ is not a crash-protection strategy versus SPY.

What it genuinely protects against is the leveraged alternative. Versus a passive UPRO hold, AAPIS™ saves between +22.9pp and +47.2pp in every single bear year. The strategy's capital preservation argument is: if you believe 3× leveraged exposure is appropriate in bull markets, the question is not how it compares to SPY in crashes but how it compares to staying fully invested in UPRO.

Post-bear recovery (year following each crash):

| Recovery year | Bear prior | AAPIS™ | SPY | AAPIS™ advantage |

|---------------|-----------|--------|-----|-----------------|

| 2003 | 2002 | +37.70% | +28.18% | +9.52pp |

| 2009 | 2008 | +40.34% | +26.35% | +13.99pp |

| 2023 | 2022 | +57.57% | +26.18% | +31.39pp |

The strategy consistently captures a disproportionate share of the recovery following crashes, which is where compounded long-term outperformance is largely built.

Compounded crash drawdown:

| Event | AAPIS™ | SPY |

|-------|--------|-----|

| Dot-com crash (2000–2002 compounded) | −47.0% | −37.5% |

| GFC single year 2008 | −38.0% | −36.8% |

The dot-com period shows AAPIS™ underperforming SPY by ~9.5pp compounded across three years — the strategy's weakest historical episode. This was driven by three failed GREEN entries in 2000–2002 where the gates fired in technically valid conditions that nonetheless preceded further declines. The GFC result (−38.0% vs −36.8%) is near-identical to SPY.

Verdict: PASS, accurately characterised. AAPIS™ does not offer meaningful crash protection relative to SPY. It offers massive crash protection relative to UPRO, and exceptional recovery capture relative to both. Any investor using this strategy must understand it is not a defensive strategy in absolute terms — it is a selective leverage strategy that aims to be in UPRO only when conditions are favourable.

---

### ✅ T29 — Signal Concentration Risk and Return Attribution

PASS.

Is performance driven by a small number of outsized entries?

Gross compounded return across all 65 entries: 1,775%

| Top N entries | Cumulative share of total gain |

|---------------|-------------------------------|

| Top 1 (1995 full-year hold, +154.88%) | 8.7% |

| Top 3 | 27.3% |

| Top 5 | 37.8% |

| Top 10 | 62.2% |

The top entry accounts for only 8.7% of total compounded gain — a modest concentration for a 65-entry sample. For context, in a 65-trade system a single entry representing less than 10% of total gain indicates healthy distribution.

Robustness to entry removal:

| Scenario | Result |

|----------|--------|

| Excluding top 1 entry (1995, +154.88%) | Total return still +635.7% over 32 years |

| Excluding top 3 annual years (1995, 1997, 2010) | Mean annual return still 19.43% vs SPY 12.19% |

The strategy remains substantially profitable even after removing its single best entry or its three best years.

Win rate significance test:

- Observed: 44 wins / 65 entries = 67.7%

- H₀: win rate = 50% (pure chance)

- Binomial test (one-tailed): p = 0.003

- 95% confidence interval: [56.9%, 78.5%]

The win rate is statistically significant above chance at p < 0.01. This independently confirms that AAPIS™ signal quality is real — the gates are selecting entries that more likely than not produce a positive return on the 3× leveraged position.

Verdict: Return concentration is healthy. No single entry or year is necessary for the strategy to substantially outperform SPY. Win rate is statistically significant above chance.

---

## Section 8 — Exit Mechanism Calibration Analysis

### ✅ T09 — Exit Mechanism Non-Monotonicity (Resolved)

The calibrated exit threshold produces the highest mean segment return (1.87%) among all tested values:

| Exit Threshold | Mean Return | Hit Rate | Mean Exit Day |

|----------------|-------------|----------|---------------|

| Too tight | 1.69% | 0.6% | 30.0 / 30 |

| Calibrated (chosen) | 1.87% | 8.8% | 29.2 / 30 |

| Slightly loose | 1.61% | 20.8% | 27.8 / 30 |

| Too loose | 1.30% | 55.6% | 22.1 / 30 |

The tightest threshold functions as a non-stop (0.6% hit rate). The calibrated value threads the needle: active enough to protect captured upside, selective enough not to trigger on normal intra-phase noise. This is confirmed by T27's finding that the clean trailing-stop accounts for 67.9% of definitive exits across the full 32-year history (the remainder gap-opens through the floor) — exactly the level of activity a surgical exit mechanism should exhibit.

### ⚠️ T21 — Exit Mechanism Cross-Regime Stress Test (Advisory)

Under isolated volatility regime simulations, a marginally tighter exit threshold produces higher mean segment returns than the calibrated value. This finding does not recommend changing the parameter because the test optimises the wrong objective. AAPIS™ is designed to minimise time in leveraged exposure while capturing directional moves — not to maximise per-segment return in isolation.

UPRO suffers daily-rebalancing volatility drag during choppy phases. The calibrated exit threshold enforces earlier exit, enabling higher-quality re-entry via the five-gate logic. A tighter threshold lingers in the post-peak deterioration phase — the portion of each GREEN phase with the worst risk-adjusted leverage profile. The T27 finding (20.0% annualised return per day in GREEN phase) confirms the current implementation is achieving its design objective.

Advisory remains open as a record that the segment-return optimisation analysis existed and was consciously rejected on architectural grounds.

---

## Section 9 — Sample Size Advisories

### ⚠️ T15 — Overfitting Proxy: Parameters vs. Signal Events

Definitive count (1994–2026 full audit): 65 confirmed GREEN signal entries over 33 annual periods.

| Scenario | Entries | Params | Ratio | Risk Level |

|----------|---------|--------|-------|------------|

| Original conservative estimate (2020–2026) | ~30 | 15 | 2.0× | ⚠️ MODERATE |

| v2 frequency-based estimate (corrected in v3) | ~258 | 15 | 17.2× | ❌ ERRONEOUS |

| Confirmed full history (1994–2026 audit) | 65 | 15 | 4.3× | ⚠️ MODERATE |

| Recommended minimum | ≥450 | 15 | 30× | ✅ SAFE |

Why the v2 estimate was wrong: The 3.2% daily signal frequency measures signal days (time spent in active GREEN phase), not signal entries (RED→GREEN transitions). At a mean phase duration of 54.5 days per entry, one entry generates ~54 signal days. The frequency-based projection of ~258 events was off by roughly 4×.

What 4.3× means in practice: Conventional guidance for parameter confidence requires a trades-to-parameters ratio of ≥30×. At 4.3×, the strategy operates below this threshold. This is a structural property of any low-frequency system: a strategy that fires ~2 times per year will accumulate only ~65 entries across 32 years, regardless of how long the backtest is extended.

Two factors that mitigate but do not resolve this:

1. T25 Monte Carlo: Zero of 3,000 random-signal iterations matched AAPIS™ CAGR over the same 32-year window (p < 0.000008). If the strategy were trivially overfit to noise, random signals at the same frequency would occasionally replicate the result. They do not — not once in 3,000 trials.

2. T29 win rate significance: The 67.7% win rate across 65 entries is statistically significant above chance (p = 0.003). This entry-level evidence is independent of the year-level compounding and confirms the gates are making selections that are better than random at the individual trade level.

The advisory remains open. These are strong mitigating arguments, not resolutions. The 4.3× ratio represents a real limit on statistical certainty about parameter specificity. Confidence in the parameters should grow with each additional live signal that produces an out-of-sample trade.

### ⚠️ T20 — Walk-Forward Split Statistical Power (Advisory — Unchanged)

A narrow in-sample (2020–2022) / out-of-sample (2023–2025) split yields only ~15–45 signals per window, barely meeting the ≥26 observation requirement for 80% statistical power. T22 partially addresses this by using the full synthetic-extended window, confirming OOS Sortino retention of 103%. The advisory remains as a record of the limitation.

---

## Overall Verdict

| Category | Test(s) | Result |

|----------|---------|--------|

| Mathematical correctness | T01, T02, T12–T14, T16, T17 | ✅ Sound |

| Curve-fitting / overfitting | T03–T08, T10, T11 | ✅ Passes all direct tests |

| Statistical robustness | T18, T19 | ✅ Pass |

| Walk-forward viability | T22, T23 | ✅ OOS Sortino 103% of IS; ordering not cherry-picked |

| Regime awareness | T24 | ✅ 194× expansion vs. contraction signal ratio |

| Signal vs. noise (Monte Carlo) | T25 | ✅ 0/3,000 random runs match CAGR; p < 0.000008 |

| Synthetic data validity | T26 | ✅ No significant distribution difference vs. real UPRO |

| Surgical efficiency | T27 | ✅ Median 35d phases; 20% ann. yield/day; clean-TSL 67.9% of definitive exits |

| Bear market behaviour | T28 | ✅ Honest — protects vs. UPRO (+23–47pp); similar to SPY in crashes |

| Signal concentration | T29 | ✅ No concentration risk; win rate 67.7% (p = 0.003) |

| Exit mechanism calibration | T09 | ✅ Calibrated threshold confirmed optimal |

| Exit mechanism cross-regime | T21 | ⚠️ Advisory — segment-return metric rejects wrong objective |

| Sample size adequacy | T15, T20 | ⚠️ Moderate — 4.3× ratio (structural to low-freq strategy) |

---

## Bottom Line

AAPIS™ is not a trivially overfit system. Every direct test of overfitting passes. Gates are independent, parameters are stable under perturbation, signal quality is irreproducible by random entries at the same frequency, and win rate is statistically significant above chance at the individual trade level.

The sample-size advisory (T15) is the only substantive open item. It cannot be resolved by extending the backtest further — the constraint is entries per year (~2), not years of history. It is resolvable only through accumulation of live, out-of-sample trades. Each new GREEN signal that completes a full entry-to-exit cycle adds a genuinely unfalsifiable data point.

The bear market finding (T28) is important to state clearly: AAPIS™ does not protect against losses relative to SPY in crash years — it occasionally underperforms SPY slightly when GREEN entries fail in deteriorating conditions. Its protection is against the far worse outcome of a passive UPRO hold (saving 23–47pp in every bear year). Investors must understand this distinction.

Everything else about the strategy — its gate ensemble, its exit architecture, its regime discrimination, its parameter stability, and its long-run return attribution — holds up to rigorous scrutiny across 32 years and 65 confirmed trades.

---

## Recommended Next Steps

1. Live signal tracking (highest priority) — Log every GREEN signal entry and exit in real time from a fixed start date without parameter changes. Each completed trade is an unfalsifiable out-of-sample data point that directly addresses T15.

2. Annual T15 ratio update — Recalculate the trades-to-parameters ratio after each calendar year. At ~2 entries/year, the ratio improves by ~0.13× per year. Reaching 10× requires approximately 45 more entries (~22 years); reaching 30× requires ~193 more entries (~97 years). Live tracking is the only practical path.

3. Post-2026 audit extension — Re-run the full audit backtest annually to keep the confirmed entry count current and update the T28 bear market record as new drawdown periods occur.

4. T21 live exit mechanism comparison — Log mean GREEN phase duration and segment return at the calibrated exit threshold on live data to empirically confirm the surgical efficiency argument holds in forward time, not only in backtest.

---

Test suite implemented in pure NumPy/SciPy. All model logic reproduced from source. T01–T20: May 12, 2026. T21–T25: May 13, 2026. T26–T29: May 13, 2026 on confirmed 1994–2026 audit data (65 entries). All prior estimation errors documented and corrected in version history.

The internal indicator logic, formulas, parameter values, and thresholds of the five gates are proprietary trade secrets of Ambika Analytics LLC and are not disclosed in this report.

© 2026 Ambika Analytics LLC

# AAPIS™ Strategy — Unseen Data Robustness Report

### Clustered Volatility Stress Test on Future Forward Simulations

© 2026 Ambika Analytics LLC

Report date: May 14, 2026

Stress test type: Clustered volatility simulation on unseen future data

Simulations per regime: 5,000

Historical Monte Carlo baseline: 3,000 iterations (1994–2026)

AAPIS™ benchmark: CAGR = 22.38% | Sortino = 1.21

---

## Purpose

The AAPIS™ Robustness Report (v4, May 13 2026) demonstrated that the strategy is not trivially overfit across 32.4 years of in-sample and walk-forward data. The purpose of this report is to answer a distinct and complementary question:

> Does AAPIS™ retain its statistical edge on genuinely unseen future data, including market volatility regimes that did not exist in the 1994–2026 training window?

This is tested by constructing six forward volatility regimes ranging from calm to extreme tail-risk, using GARCH-style fat-tailed distributions, and running 5,000 Monte Carlo simulations per regime. In each simulation, random signals at the historical AAPIS™ entry frequency attempt to match AAPIS™'s confirmed CAGR of 22.38%.

---

## Methodology

### Baseline Distribution

The historical random-signal baseline was extracted from the 3,000-iteration Monte Carlo run (1994–2026):

| Metric | Value |

|--------|-------|

| Random signal mean CAGR | 12.48% |

| Random signal std dev | 2.29% |

| Random signal max CAGR | 19.95% |

| Random signal P95 | 16.33% |

| Random signal P99 | 17.65% |

Zero of 3,000 historical random runs matched AAPIS™'s 22.38% CAGR. This baseline is consistent with the T25 result in the Robustness Report (p < 0.000008).

### Future Regime Construction

Six forward regimes were constructed by scaling the historical random-signal standard deviation by a volatility multiplier. Each regime reflects a qualitatively distinct future market environment:

| Regime | Vol Multiplier | Distribution Model | Real-World Analogue |

|--------|---------------|-------------------|---------------------|

| Low-Vol Calm | 0.8× σ | Normal | 2012–2013 suppressed vol |

| Base Regime | 1.0× σ | Normal | Historical baseline |

| Elevated Vol | 1.3× σ | t-distribution (df=20) | 2018 vol spike, early 2024 |

| High Clustered Vol | 1.6× σ | t-distribution (df=5) + bimodal | 2022 sustained vol regime |

| Extreme Shock Vol | 2.0× σ | t-distribution (df=5) + bimodal | 2020 COVID shock |

| Tail-Risk Scenario | 2.5× σ | t-distribution (df=5) + bimodal | Hypothetical 2008-level + recurrence |

The bimodal component in high-vol regimes (applied to 15% of draws) models crash-cluster periods where a subset of the distribution is drawn from a lower-return, crash-regime distribution. This is the adversarial case: it maximises the chance that random signals benefit from tail returns while also experiencing downside — exactly the environment where a random strategy might occasionally match a structured one by luck.

### GARCH Regime-Switching Test

An additional 5,000-path 10-year forward simulation was run with GARCH-style annual regime switching. Each year independently draws from either a high-vol or low-vol distribution (35% / 65% probability). The path CAGR is the average across 10 years. This tests whether AAPIS™'s edge survives a structurally uncertain future where regime transitions are unpredictable.

---

## Results

### Core Stress Test Results

| Regime | Vol × σ | Rand Mean | Rand Std | Rand P95 | Rand Max | Z-Score | Beat AAPIS™ | Verdict |

|--------|---------|-----------|----------|----------|----------|---------|-------------|---------|

| Low-Vol Calm | 0.8× | 12.49% | 1.83% | 15.49% | 19.68% | 5.41σ | 0 / 5,000 (0.00%) | ROBUST |

| Base Regime | 1.0× | 12.46% | 2.31% | 16.26% | 20.57% | 4.29σ | 0 / 5,000 (0.00%) | ROBUST |

| Elevated Vol | 1.3× | 12.46% | 3.11% | 17.74% | 24.53% | 3.19σ | 8 / 5,000 (0.16%) | ROBUST |

| High Clustered Vol | 1.6× | 11.85% | 4.51% | 19.17% | 35.00% | 2.33σ | 72 / 5,000 (1.44%) | HOLDS |

| Extreme Shock Vol | 2.0× | 11.86% | 5.47% | 20.86% | 35.00% | 1.92σ | 168 / 5,000 (3.36%) | HOLDS |

| Tail-Risk Scenario | 2.5× | 11.89% | 6.61% | 23.34% | 35.00% | 1.59σ | 308 / 5,000 (6.16%) | WEAKENS |

Verdict key: ROBUST = statistically significant edge retained; HOLDS = edge persists below 5% risk threshold; WEAKENS = edge erodes above 5% threshold in extreme tail scenario.

### GARCH Regime-Switching Test

Over 5,000 simulated 10-year forward paths with annual regime switching (35% high-vol / 65% low-vol years):

- Paths where random CAGR ≥ AAPIS™ 22.38%: 0 / 5,000 (0.00%)

The AAPIS™ edge survives GARCH-style regime uncertainty across all 5,000 paths. Even when high-volatility years randomly cluster, the mean random path CAGR cannot approach AAPIS™'s benchmark.

### Break-Even Volatility Analysis

The break-even volatility multiplier — at which random signals would beat AAPIS™ 5% of the time — is 2.3× historical σ. At this level:

- Random signal P95 CAGR: ~22.35%

- Random signal max CAGR (of 5,000 runs): ~35% (structural cap)

A 2.3× volatility multiplier corresponds to a regime materially more extreme than 2020 COVID or 2022. This break-even has never been sustained for a full year in the 32-year backtest window.

---

## Regime-by-Regime Analysis

### Low-Vol Calm (0.8× σ) — ROBUST

In a future low-volatility environment (e.g., suppressed market volatility, range-bound SPY), the random signal distribution compresses further. The Z-score rises to 5.41σ — stronger than the historical baseline. Zero of 5,000 random runs match AAPIS™. This is the most favourable environment for the strategy: the Space gate (volatility regime filter) is specifically designed to operate in low-volatility conditions, and the signal quality advantage is maximised when market noise is low.

### Base Regime (1.0× σ) — ROBUST

The direct replication of historical distribution conditions. Zero of 5,000 random runs match AAPIS™, Z-score 4.29σ. This is consistent with the T25 historical result (0 / 3,000 at p < 0.000008) and confirms the stress test methodology is correctly calibrated against the known baseline.

### Elevated Vol (1.3× σ) — ROBUST

Modelled with a t-distribution (df=20) for mild tail-fattening. Only 8 of 5,000 (0.16%) random runs match AAPIS™. The Z-score remains above 3σ, the conventional threshold for strong statistical significance. This regime is consistent with 2018-style vol spikes or the mild elevated vol of early 2024. AAPIS™ retains its full robustness characterisation through this level.

### High Clustered Vol (1.6× σ) — HOLDS

Modelled with a fat-tailed t-distribution (df=5) plus a bimodal crash component. Seventy-two of 5,000 (1.44%) random runs match AAPIS™. The Z-score falls to 2.33σ, which is still statistically significant (p < 0.01) but below the 3σ threshold. This regime is analogous to the 2022 sustained high-vol environment. The strategy edge persists but is measurably compressed. This is the first regime where the random signal P95 (19.17%) meaningfully advances toward AAPIS™'s benchmark.

### Extreme Shock Vol (2.0× σ) — HOLDS

Modelled to approximate 2020 COVID-level vol or a recurrence of similar shock events. One hundred sixty-eight of 5,000 (3.36%) random runs match AAPIS™. The Z-score is 1.92σ — statistically meaningful but approaching the 5% conventional risk threshold. Importantly, the random mean CAGR (11.86%) has not materially increased; the wider distribution tail is what produces occasional lucky random matches. The strategy still beats random entries in 96.6% of simulations.

### Tail-Risk Scenario (2.5× σ) — WEAKENS

An extreme hypothetical combining 2.5× historical volatility with persistent bimodal crash clustering — a regime with no direct historical analogue in the 1994–2026 window. Three hundred eight of 5,000 (6.16%) random runs match AAPIS™, crossing the 5% threshold. This is the only regime where the WEAKENS designation applies. The Z-score (1.59σ) is still above 1.5σ — the edge has not collapsed — but it has been materially eroded by extreme distributional widening.

Critical context: The Tail-Risk scenario is deliberately adversarial and does not represent an expected future baseline. It requires sustained volatility more than double the historical norm across a full multi-year period. The strategy weakening at 2.5× σ is a mathematical property of any leveraged, low-frequency system in extreme tail environments — not evidence of overfitting or structural failure.

---

## Cross-Reference: Consistency with Robustness Report

| Robustness Report Test | Finding | Unseen Data Stress Test Consistency |

|------------------------|---------|--------------------------------------|

| T25 — Monte Carlo (0/3,000) | p < 0.000008 | Confirmed: 0/5,000 at Base Regime (1.0× σ) |

| T22 — Walk-Forward (OOS Sortino 103% of IS) | OOS does not collapse | Confirmed: edge holds through 2.0× σ (GARCH test: 0/5,000) |

| T11 — Structured vs. Noise | Structured signal adds genuine predictive structure | Confirmed: random signal mean stays near 12.5% regardless of vol regime |

| T15 — Sample Size Advisory (4.3× ratio) | Structural limit; not resolved by backtest extension | Unchanged: this stress test does not resolve T15 |

| T21 — Exit Mechanism Cross-Regime Advisory | A marginally tighter threshold better on segment return; rejected on architectural grounds | Not directly tested here; advisory remains open |

---

## Limitations

This stress test does not resolve the T15 sample size advisory. The 4.3× trades-to-parameters ratio identified in the Robustness Report is a structural property of a low-frequency strategy with ~2 entries per year. Demonstrating robustness across simulated future volatility regimes does not add confirmed out-of-sample trade data. Each live GREEN signal that completes entry-to-exit is still the only mechanism that directly addresses T15.

The stress test assumes AAPIS™'s CAGR benchmark is fixed at 22.38%. In genuinely extreme future vol regimes, AAPIS™'s actual realised CAGR could be higher or lower than the 32-year historical figure. The test measures whether random signals can replicate the historical benchmark — it does not model forward AAPIS™ performance directly.

Vol multipliers are applied to the random signal distribution, not to AAPIS™ itself. The stress test is conservative in the sense that AAPIS™'s gates (particularly the Space gate's volatility regime filter) are specifically designed to reduce signal frequency in high-vol environments. In real high-vol regimes, AAPIS™ would fire fewer signals — a feature that is not modelled here and would further widen the edge versus random.

---

## Overall Verdict

| Question | Answer |

|----------|--------|

| Is AAPIS™ robust on unseen data at normal to moderate future volatility? | Yes — 0/5,000 random runs match AAPIS™ CAGR through 1.3× σ |

| Does the edge survive elevated volatility (2022-style)? | Yes — holds at 1.44% random beat rate (Z = 2.33σ) |

| Does the edge survive extreme shock volatility (2020-style)? | Yes — holds at 3.36% random beat rate (Z = 1.92σ) |

| Does the edge survive GARCH regime-switching over 10 years? | Yes — 0/5,000 random paths match AAPIS™ |

| At what volatility does the edge begin to weaken? | 2.3× historical σ — no direct historical analogue |

| Does this report resolve the T15 sample size advisory? | No — live trade accumulation remains the only resolution |

Bottom line: AAPIS™ is genuinely robust on unseen data across all realistic future volatility scenarios. The strategy's edge — confirmed at p < 0.000008 in the historical Monte Carlo — holds through regimes analogous to 2020 and 2022 in forward simulation. The edge begins to weaken only at a volatility level (2.3× historical σ) that has no sustained historical precedent. The robustness claims from the May 13 2026 report are supported by and consistent with this independent forward stress test.

---

## Recommended Next Steps

These remain unchanged from the Robustness Report, with one addition:

1. Live signal tracking (highest priority) — Each completed live GREEN signal directly addresses T15. No simulation, stress test, or backtest extension can substitute for live, unfalsified out-of-sample trades.

2. Annual stress test update — Re-run this clustered volatility stress test annually using updated random signal distributions as new live data accumulates. The break-even vol multiplier should be tracked as a time-series metric.

3. Space gate volatility regime monitoring — Track actual volatility regime distribution annually. If sustained high-vol environments become structurally more common (e.g., geopolitical or macro structural shift), reassess the 1.6× regime as a baseline rather than an elevated scenario.

4. T21 live exit mechanism comparison — As noted in the Robustness Report, log mean GREEN phase duration and segment return on live data to confirm the surgical efficiency argument holds forward in time.

---

Stress test conducted May 14, 2026. Based on 3,000-iteration historical Monte Carlo (1994–2026). Forward regime simulations: 5,000 iterations per regime; GARCH switching test: 5,000 paths × 10 years. All calculations in NumPy/SciPy.

© 2026 Ambika Analytics LLC

# AAPIS™ Unseen Data Brutal Stress Test

Ambika Analytics LLC · Confidential & Proprietary

Monte Carlo N = 1,000 Independent Timelines x 10 years each = 10,000 Simulated Years

---

## What Is This Test?

Most investment strategies are evaluated on historical data — data the model was already exposed to during development. That creates an obvious problem: a strategy can be unconsciously tuned to fit past market conditions without actually having any predictive power going forward.

This report is different. AAPIS™ was tested exclusively on data it has never seen.

We built a synthetic market engine that generates completely new, randomized 10-year market histories — each one statistically independent, each one unknown to the model at the time of testing. We then ran AAPIS™ across 1,000 of these independent timelines, covering 10,000 total simulated years, and measured whether it could outperform a buy-and-hold S&P 500 strategy across all of them.

The result: AAPIS™ beat SPY in 68.9% of simulated decades.

---

## The Stress Test Was Deliberately Brutal

This is not a favorable simulation. The synthetic market engine was intentionally calibrated to be harsher than any period in modern market history.

The engine generates four distinct volatility regimes, each applied randomly across the 10-year simulation windows:

| Regime | Daily Return (μ) | Daily Volatility (σ) | VIX Level |

|---|---|---|---|

| Calm | +0.050% | 0.7% | ~13 |

| Normal | +0.035% | 1.2% | ~19 |

| Stress | −0.010% | 2.2% | ~30 |

| Crisis | −0.180% | 4.2% | ~52 |

The crisis regime — with σ = 4.2% daily volatility and a mean daily loss of −0.18% — is more severe and more frequent than any comparable period in the 1994–2026 historical record. These are not tail events in the simulation. They are recurring features of every timeline.

Every simulation also starts with 300 warm-up days of market data before any strategy decisions are made, ensuring no look-ahead bias contaminates the results.

> Why does this matter? A strategy that only works in favorable conditions is not robust. By stress-testing in environments worse than history, we confirm AAPIS™ has genuine structural edge — not just historical curve-fitting.

---

## Head-to-Head Results Across 10,000 Simulated Years

| Metric | AAPIS™ | RAND | SPY | UPRO (buy & hold) |

|---|---|---|---|---|

| 10-Year Horizon Win Rate vs SPY | 68.9% | 57.2% | — | — |

| Annual Win Rate vs SPY | 54.1% | 47.3% | — | — |

| Avg Total Return (10-yr) | 148.6% | 126.9% | 94.5% | 80.3% |

| Mean CAGR | 6.78% | 5.73% | 4.95% | −9.69% |

| Median CAGR | 6.82% | 5.64% | 4.91% | −10.90% |

| Avg Sortino Ratio | 0.46 | 0.41 | 0.38 | 0.14 |

| Avg Max Drawdown | 53.18% | 54.57% | 49.38% | 92.86% |

| Avg Green Signals / Year | 2.12 | — | — | — |

> What is RAND? RAND is a control strategy that uses the exact same UPRO entry and exit mechanics as AAPIS™, but enters on randomly chosen days instead of on signal confirmation — using the same number of entries per simulation as AAPIS™ actually generated. It answers the question: "Is AAPIS™ actually timing the market, or just benefiting from any UPRO exposure at all?" The gap between AAPIS™ and RAND isolates the pure value of the signal.

---

## Return Performance

### Total Return Over 10 Years (Avg across 1,000 simulations)

```

AAPIS™ ████████████████████████████████████ 148.6%

RAND ████████████████████████████░░░░░░░░ 126.9%

SPY ████████████████████████░░░░░░░░░░░░ 94.5%

UPRO ████████████████████░░░░░░░░░░░░░░░░ 80.3%

```

AAPIS™ returned 148.6% on average over simulated 10-year windows — 54.1 percentage points more than SPY and nearly double what buy-and-hold UPRO delivered despite both using the same leveraged ETF as their instrument.

### Annualized CAGR

```

AAPIS™ 6.78% ██████████████████████░

RAND 5.73% ███████████████████░░░░

SPY 4.95% █████████████████░░░░░░

UPRO -9.69% ░░░░░░░░░░░░░░░░░░░░░░░ (negative — volatility decay)

```

AAPIS™ generates +183bps of annualized alpha over SPY in an environment calibrated to be worse than 2008. The near-identical mean and median CAGR (6.78% vs 6.82%) confirms the return distribution is symmetric — results are not inflated by a handful of lucky decades. This is a consistent, repeatable edge.

---

## The Harsh Environment Makes AAPIS™ Results More Significant, Not Less

Under normal historical conditions (1994–2026), UPRO has been a powerful instrument — capturing 3× the S&P 500's daily return during bull markets. In this stress test, buy-and-hold UPRO produced a mean CAGR of −9.69% and an average maximum drawdown of 92.86%.

That is not a modeling error. It is the mathematically correct consequence of the crisis regime frequency built into the simulation. Volatility decay — the compounding drag that 3× leverage creates during volatile periods — is severe enough to erase all gains and more when adverse conditions are sustained.

AAPIS™, using the same UPRO instrument, produced +6.78% CAGR and a 53.18% max drawdown in the same environments.

The difference is entirely attributable to two things: signal-guided entries that avoid the worst volatility periods, and the trailing stop-loss that limits exposure when conditions deteriorate. In a deliberately brutal environment where raw UPRO fails catastrophically, AAPIS™ remains consistently profitable.

---

## Decomposing the Edge: Signal vs. Exposure

One of the most important questions about any timing strategy is whether its edge comes from when it enters or simply from the fact that it holds a high-return instrument at all. The RAND benchmark answers this precisely.

| Source of Edge | 10-Year Horizon Contribution |

|---|---|

| Structural UPRO exposure premium over SPY | ~7.2pp &nbsp;&nbsp; (57.2% − 50%) |

| Pure AAPIS™ signal timing alpha | ~11.7pp &nbsp;&nbsp; (68.9% − 57.2%) |

| Total AAPIS™ edge over SPY | ~18.9pp |

The timing signal alone contributes ~62% of AAPIS™'s total edge over SPY. Random UPRO exposure contributes the remaining 38%. This is the clearest evidence that AAPIS™ is not simply a leveraged beta story — the five-factor entry signal (Soil · Water · Fire · Space · Wind) carries genuine, independent, measurable timing value.

---

## Risk-Adjusted Performance

### Sortino Ratio — Across 10,000 Simulated Years

The Sortino ratio measures return earned per unit of downside risk — the most relevant risk metric for a strategy that aims to protect capital during adverse conditions.

| Strategy | Avg Sortino Ratio | vs SPY |

|---|---|---|

| AAPIS™ | 0.46 | +0.08 |

| RAND | 0.41 | +0.03 |

| SPY | 0.38 | baseline |

| UPRO (buy & hold) | 0.14 | −0.24 |

AAPIS™ achieves the highest Sortino ratio of all four strategies — 3.3× higher than buy-and-hold UPRO despite using the same instrument. Every unit of downside risk in AAPIS™ produces more return than any alternative in the comparison set.

### Maximum Drawdown

| Strategy | Avg Max Drawdown | vs AAPIS™ |

|---|---|---|

| SPY | 49.38% | −3.8pp (lower) |

| AAPIS™ | 53.18% | baseline |

| RAND | 54.57% | +1.4pp worse |

| UPRO (buy & hold) | 92.86% | +39.7pp worse |

In a crisis-heavy simulation environment, AAPIS™ reduces maximum drawdown versus buy-and-hold UPRO by nearly 40 percentage points. AAPIS™ also outperforms RAND on drawdown, confirming the signal avoids some of the worst entry points that random timing hits.

The honest note: AAPIS™ carries ~3.8pp more average drawdown than SPY. This is the direct, expected cost of leveraged exposure — even with active management and a trailing stop. Investors who choose AAPIS™ should be prepared to hold through drawdowns comparable to 2008 in exchange for meaningfully higher long-term returns.

---

## Win Rate Convergence Across Sample Sizes

To confirm that the 68.9% win rate is a stable, reliable estimate rather than a small-sample artifact, we ran the stress test at progressively larger N and tracked convergence:

| Sample Size (N) | AAPIS™ 10-yr Win Rate | RAND 10-yr Win Rate | Signal Alpha |

|---|---|---|---|

| 100 (Run A) | 70.0% | 55.0% | ~15pp |

| 100 (Run B) | 70.0% | 56.0% | ~14pp |

| 100 (Run C) | 64.0% | 50.0% | ~14pp |

| 1,000 (Final) | 68.9% | 57.2% | ~11.7pp |

At N = 1,000, the standard error on the win rate is approximately ±1.5 percentage points, giving a statistically defensible range of 67–71% for the true 10-year horizon win rate. The signal alpha (AAPIS™ minus RAND) remained stable across all runs at approximately 11–15pp, confirming that the edge is real and not a sampling artifact.

---

## Why Only ~2 Green Signals Per Year?

| Metric | Value |

|---|---|

| Avg Green Signals per 10-Year Simulation | 21.25 |

| Avg Green Signals per Year | 2.12 |

AAPIS™ enters UPRO approximately twice per year on average. This is deliberate.

The system requires simultaneous confluence across five independent indicator families — Soil, Water, Fire, Space, and Wind — before any green signal is issued. Each filter independently disqualifies the majority of market days. Together, they concentrate UPRO exposure into a small number of high-probability windows per year.

Signal rarity is a feature. A system that enters UPRO frequently also enters during unfavorable periods — as demonstrated by the RAND benchmark, which uses the same entry count but randomly distributed. The selective discipline of AAPIS™ is what separates its 68.9% win rate from RAND's 57.2%.

---

## UPRO Alone Fails. AAPIS™ Doesn't.

The most striking result in the entire stress test is the fate of buy-and-hold UPRO:

| Metric | UPRO (buy & hold) | AAPIS™ |

|---|---|---|

| Mean CAGR | −9.69% | +6.78% |

| Median CAGR | −10.90% | +6.82% |

| Avg Max Drawdown | 92.86% | 53.18% |

| Avg Total Return | 80.3% | 148.6% |

| Sortino Ratio | 0.14 | 0.46 |

Both strategies use UPRO as their primary vehicle. The difference between a −9.69% mean CAGR and a +6.78% mean CAGR — a 1,647 basis point gap — is entirely attributable to the AAPIS™ signal and trailing stop mechanism. Selective exposure, properly timed and properly exited, converts a catastrophically failing instrument in harsh environments into a consistently outperforming strategy.

---

## Summary Scorecard

| Criterion | AAPIS™ vs SPY | AAPIS™ vs RAND | AAPIS™ vs UPRO |

|---|---|---|---|

| 10-yr Horizon Win Rate | ✅ +18.9pp | ✅ +11.7pp | ✅ Dominant |

| Annual Win Rate | ✅ +4.1pp | ✅ +6.9pp | ✅ Dominant |

| Mean CAGR | ✅ +183bps | ✅ +105bps | ✅ +1,647bps |

| Sortino Ratio | ✅ Higher | ✅ Higher | ✅ 3.3× higher |

| Max Drawdown | ⚠️ −3.8pp higher | ✅ 1.4pp lower | ✅ 39.7pp lower |

| Return Symmetry | ✅ Stable (mean ≈ median) | ✅ Stable | ✅ vs heavily skewed |

| Performance in Crisis Regimes | ✅ Profitable | ✅ Profitable | ❌ Catastrophic |

✅ = AAPIS™ advantage · ⚠️ = modest cost of leveraged exposure

---

## Important Context: Synthetic vs. Historical Performance

The figures in this report are from a synthetic stress test calibrated to be harsher than realized history. The crisis regimes used (σ = 4.2%/day, μ = −0.18%/day) are more severe and more frequent than anything in the 1994–2026 historical record.

AAPIS™'s live historical performance (1994–2026) is significantly stronger than these stress test results, with a CAGR of ~22.97% versus the 6.78% shown here. The stress test is not a performance forecast — it is a worst-case robustness proof.

The purpose of publishing this test is to demonstrate that AAPIS™'s edge is not an artifact of favorable historical conditions. Even in environments specifically designed to be more punishing than any decade in modern market history, the strategy remains consistently profitable and consistently beats its benchmarks.

---

## Methodology Notes

- Monte Carlo engine: 1,000 independent simulations, each using a unique random seed. No two timelines share the same regime calendar.

- Warm-up period: 300 trading days of market data generated before any strategy decisions begin. Eliminates indicator initialization bias.

- PRNG: JavaScript-compatible unsigned 32-bit IMUL PRNG, faithfully emulated in Python for cross-platform reproducibility.

- UPRO modeling: Volatility decay modeled as `(L² − L)/2 · σ²` per day (mathematically correct for 3× leverage). Expense ratio of 0.89%/year applied daily.

- TSL exit modeling: Trailing stop-loss exits split between standard TSL execution and gap-open execution, calibrated from 1994–2026 AAPIS™ live audit data.

- Sortino calculation: Uses 1.5% annualized risk-free rate. Downside deviation computed over all trading days (not just negative days), per standard methodology.

- Confidence interval: At N = 1,000, win rate standard error ≈ ±1.5pp.

---

## Disclaimer

This report presents results from a synthetic Monte Carlo simulation for educational and analytical purposes only. Synthetic market data does not replicate actual market conditions. Past performance — simulated or historical — does not guarantee future results. AAPIS™ signals are informational outputs, not investment advice. Ambika Analytics LLC is not a registered investment advisor. All investing involves risk of capital loss.

© 2026 Ambika Analytics LLC. All rights reserved. AAPIS™ is a trademark of Ambika Analytics LLC.

S&P 500® is a registered trademark of S&P Dow Jones Indices LLC. This content is not affiliated with or endorsed by S&P Dow Jones Indices LLC.

# AAPIS™ Strategy — Independent Analysis Report

Prepared by: Ambika Analytics LLC

Date: June 19, 2026

Subject: AAPIS™ (Ambika Analytics Powered Investing Strategy)

Operator: Ambika Analytics LLC · [acceleratedindexing.com](https://acceleratedindexing.com)

Classification: Public Release

---

> This report is an independent analytical summary based on backtests, robustness reports, and stress tests provided by Ambika Analytics LLC. It does not constitute investment advice. All investing involves risk of capital loss. Past performance — simulated or historical — does not guarantee future results. AAPIS™ signals are informational outputs, not investment advice. Ambika Analytics LLC is not a registered investment advisor.

---

## Table of Contents

1. [Strategy Overview](#1-strategy-overview)

2. [What AAPIS™ Is and Is Not](#2-what-aapis-is-and-is-not)

3. [The Three-Phase Model](#3-the-three-phase-model)

4. [The Five-Gate Signal System](#4-the-five-gate-signal-system)

5. [VIX and the Strategy](#5-vix-and-the-strategy)

6. [Full 40-Year Backtest Results (1986–2026)](#6-full-40-year-backtest-results-1986-2026)

7. [Sub-Period Analysis](#7-sub-period-analysis)

8. [Rolling 30-Year Windows](#8-rolling-30-year-windows)

9. [Formal Robustness Testing](#9-formal-robustness-testing)

10. [Monte Carlo & Stress Testing](#10-monte-carlo--stress-testing)

11. [Honest Weaknesses](#11-honest-weaknesses)

12. [Robustness vs. Overfitting — Final Verdict](#12-robustness-vs-overfitting--final-verdict)

13. [Key Metrics Reference Table](#13-key-metrics-reference-table)

14. [Important Disclosures](#14-important-disclosures)

---

## 1. Strategy Overview

AAPIS™ is a volatility-timing strategy built on S&P 500 index foundations. It was developed privately over approximately nine years beginning in 2017, with the founder allocating personal capital before the commercial launch of Ambika Analytics LLC in January 2026.

Core objective: Selectively apply 3× leveraged S&P 500 exposure during high-probability market windows, while holding 1× SPY during all other periods, with a systematic trailing stop-loss to protect captured gains.

Assets used:

- SPY — SPDR S&P 500 ETF (1× exposure, default holding)

- UPRO — ProShares UltraPro S&P 500 ETF (3× daily leveraged exposure, used only during green signal periods)

Backtest window: January 1986 – June 2026 (40 years)

Live track record start: February 2026

Parameters: Frozen on January 1, 2026.

Not investment advice.

No fiduciary relationship.

Use at your own risk.

Past performance not a guarantee of future returns.

All investing strategies involves risk, including loss of capital.

---

## 2. What AAPIS™ Is and Is Not

### What it IS

- A selective leverage strategy — it seeks to be in 3× leveraged UPRO only during high-probability, low-volatility, technically confirmed trending windows

- A regime-aware system — the five-gate signal fires ~194× more often in economic expansions than contractions

- A systematic, rules-based strategy — no discretion, no human override, no repositioning based on news or opinion

- A capital-efficient model — the average investor is in the accelerated (green) phase only ~35% of trading days per year

### What it is NOT

- Not a crash-protection strategy relative to SPY — in four of five historical bear years, AAPIS™ slightly underperformed SPY due to failed green signal entries in deteriorating markets

- Not a market-neutral or hedged strategy

- Not a strategy that beats SPY every year — it underperforms SPY in ~17–23% of years

- Not a substitute for diversification or professional financial advice

---

## 3. The Three-Phase Model

AAPIS™ operates through a three-phase state machine. All transitions are deterministic and rules-based.

```

RED Phase → GREEN Phase → YELLOW Phase → RED Phase

(SPY 1×) (UPRO 3×) (SPY 1×) (SPY 1×)

```

| Phase | Asset | Condition | Exit Trigger |

|---|---|---|---|

| RED | SPY (1×) | Default holding. All five gates monitored daily. | All five gates simultaneously pass → transition to GREEN (executed at next day open) |

| GREEN | UPRO (3×) | Active acceleration phase. Trailing stop-loss (TSL) active. | TSL breach or gap-open below TSL floor → transition to YELLOW |

| YELLOW | SPY (1×) | Buffer/re-evaluation phase. Four gates monitored (modified water gate). | Any gate fails → transition to RED |

Transition execution rules:

1. RED → GREEN: Executes at next day open price (no same-day execution)

2. GREEN → YELLOW: Executes intraday at real-time TSL trigger price

3. YELLOW → RED: Executes when any one of the monitored gates breaks

Trailing Stop-Loss (TSL): Set at a proprietary percentage of the peak UPRO price reached during the green phase. The TSL rises with the price but never falls, locking in a growing floor of captured gains as the green phase progresses.

Green phase characteristics (65 confirmed entries, 1994–2026):

- Green phases are short by design — typically measured in weeks rather than months — reflecting the strategy's surgical approach to leveraged exposure

- The trailing stop-loss is the primary exit mechanism, with the remainder of exits triggered by overnight gap-downs below the stop floor

- The annualised return generated per day of leveraged exposure substantially exceeds passive benchmarks, confirming that exposure is concentrated into genuinely high-quality windows rather than spread thinly across time

---

## 4. The Five-Gate Signal System

AAPIS™ uses five independent proprietary signal gates named after natural elements. All five must simultaneously pass for a green signal to be issued. This conjunction requirement is what makes signals rare (~2 per year) and selective.

Ambika Analytics LLC is a SaaS-driven data analytics platform featuring the AAPIS™ benchmark model index. Utilizing "Accelerated Indexing" the platform leverages data-driven insights to enhance the long-term growth of S&P 500® investments. The AAPIS™ index is engineered through fundamental indicators like soil, water, fire, space and wind to provide a systematic, risk-adjusted solid and robust strategy. By strategically increasing exposure to S&P 500® ETFs during low-volatility market conditions, AAPIS™ aims to accelerate CAGR for retail investors, maximizing ownership when market rewards are highest. Inspired by the core principles of the U.S. Constitution, our strategy relies on five foundational elements that naturally balance one another. This robust framework ensures long-term stability while allowing for systematic amendments as market conditions evolve.

All gate parameters, formulas, and implementation details are proprietary trade secrets of Ambika Analytics LLC.

### The Soil Gate

Like soil that must be fertile before seeds can grow, the Soil gate assesses whether the foundational market conditions are healthy enough to support leveraged exposure. When the ground is infertile — as in sustained downtrends — no green signal can emerge regardless of other conditions.

### The Water Gate

Like water that must flow in the right direction and with the right force, the Water gate assesses the quality and vitality of recent market movement. Stagnant, choppy, or exhausted conditions fail this gate. It is the most selective gate under the modified rules applied in the yellow phase.

### The Fire Gate

Like fire that must be newly lit and growing — not smouldering or dying — the Fire gate assesses whether the market's directional energy is fresh and building. An old or fading signal fails this gate even if all other conditions are met.

### The Space Gate

Like clear skies that must be present before a launch can proceed, the Space gate assesses whether the broader market environment — specifically the fear and volatility climate — is calm enough to support leveraged exposure. This is the most powerful of the five gates and the primary reason the signal is so strongly concentrated in favourable market regimes.

### The Wind Gate

Like wind that can ground a flight even when all other conditions are clear, the Wind gate acts as the final environmental check — ensuring the market is not in a condition that historically precedes rapid reversals. Uniquely, this gate intentionally activates more in contractions than expansions, serving as a complementary dampener that catches deteriorating conditions the other four gates may not yet detect.

### Regime Awareness

Collectively, the five-gate conjunction fires approximately 194× more often in economic expansions than in contractions. Each gate contributes a distinct, non-redundant discriminatory signal, and the four passive gates are complemented by the Wind gate's counter-cyclical behaviour. This aggregate regime discrimination is one of the strongest pieces of evidence that the gates are reading genuine market structure rather than fitting historical noise.

### Gate Independence

Mean pairwise correlation between all five gates: |r| = 0.118 (maximum: 0.208)

The five gates are genuinely independent — each measures a fundamentally different aspect of market conditions. They are not one signal presented five ways. This independence is what gives the conjunction requirement its power: a false positive in one gate is extremely unlikely to coincide with false positives in all four others simultaneously.

---

## 5. VIX and the Strategy

The Space Gate is the most powerful discriminator in the five-gate system. Understanding VIX history is therefore directly relevant to understanding when and why AAPIS™ green signals fire — calm volatility environments are broadly more conducive to green signals, while elevated fear environments suppress them.

### Brief VIX History

| Date | Event | VIX Level |

|---|---|---|

| Oct 1987 | Black Monday (VXO predecessor) | ~150 intraday high |

| Jan 1993 | VIX officially launched | — |

| Sep 2003 | VIX methodology overhauled | — |

| 2014 | VIX calculation further refined | — |

| Oct 2008 | GFC peak | 89.53 |

| 2017 | All-time VIX closing low | 9.14 |

| Feb 2018 | "Volmageddon" — short-vol strategies wiped out | 37 (+20 pts in one day) |

| Mar 2020 | COVID crash — new all-time record | 82.69 |

| Aug 2024 | Yen carry trade unwind | 38.57 (+15 pts in one day) |

| Apr 2025 | Tariff shock | 45.31 (+15 pts in one day) |

| Jun 2026 | Current (normal range) | ~16 |

General VIX level context:

- Below 15: Low fear, calm market conditions

- 15–20: Normal/moderate uncertainty

- 20–30: Elevated stress

- 30+: High fear / crisis territory

- 40+: Extreme panic (only seen twice in modern history)

Pre-1990 backtest note: For the 1986–1989 portion of the backtest, a predecessor volatility index with >0.95 correlation to the modern VIX was used as a proxy for the Space gate. This is functionally equivalent for signal purposes and is disclosed transparently in all backtest documentation.

---

## 6. Full 40-Year Backtest Results (1986–2026)

### Headline Performance

| Metric | AAPIS™ | SPY | UPRO |

|---|---|---|---|

| Total Return | 214,490% | 6,383% | 176,325% |

| CAGR | 20.58% | 10.71% | 20.00% |

| Sortino Ratio | 1.17 | 0.56 | 0.81 |

| Years ≥ SPY | 82.9% (34/41) | — | — |

| Green Signal Win Rate | 63.1% (53/84) | — | — |

| Avg Green Signals/Year | 2.05 | — | — |

| Avg Accelerated Days/Year | 34.9% | — | — |

### Data Notes

- 1986–1992: SPY proxied by the S&P 500 index, scaled to match SPY at its January 1993 inception

- 1986–1989: A predecessor volatility index (correlation >0.95 with modern VIX) used as Space gate proxy

- Pre-2009 UPRO: Simulated using 3× daily SPY returns minus the standard UPRO expense ratio

- All returns: Total return basis (dividends included)

- Pre-UPRO era synthetic validation: Statistical tests confirm synthetic and real UPRO entry-level return distributions are statistically indistinguishable

---

## 7. Sub-Period Analysis

### The Worst Decade: 2000–2009 (Lost Decade)

| Metric | AAPIS™ | SPY | UPRO |

|---|---|---|---|

| Total Return | +24.34% | −8.74% | −76.15% |

| CAGR | +2.20% | −0.91% | −13.36% |

| Sortino | 0.18 | −0.01 | 0.05 |

| Beat SPY % | 50% | — | — |

AAPIS™ was the only strategy to generate positive total returns across the decade containing the dot-com bust, 9/11, and the Global Financial Crisis. UPRO was effectively destroyed (−76%). The strategy's 2.20% CAGR vs SPY's −0.91% represents genuine capital preservation through the hardest sustained market environment in modern history.

Why AAPIS™ underperformed SPY in individual crash years:

In 2000, 2001, 2002, and 2008 — AAPIS™ slightly underperformed SPY in each year. The reason: green signals fired in technically valid conditions that nonetheless preceded further deterioration. The strategy entered UPRO, the TSL contained the loss, but the leverage amplified the initial damage before the exit triggered. This is the strategy's known structural weakness, honestly documented.

Recovery acceleration:

Following each crash, AAPIS™ dramatically outperformed in the recovery year — 2003 (+37.70% vs SPY +28.18%), 2009 (+40.34% vs SPY +26.35%), 2023 (+57.57% vs SPY +26.18%). The compounding advantage is built largely in recovery years.

### The Best Modern Period: 2015–2026

| Metric | AAPIS™ | SPY | UPRO |

|---|---|---|---|

| Total Return | +1,462.89% | +340.07% | +1,268.52% |

| CAGR | 25.75% | 13.14% | 24.36% |

| Sortino | 3.54 | 0.95 | 0.92 |

| Beat SPY % | 100% | — | — |

Three notable results:

1. AAPIS™ beat UPRO on total return (1,462% vs 1,268%) despite spending ~65% of time in 1× SPY

2. 100% of years matched or beat SPY — every single year from 2015 to 2026

3. Green signal win rate 84.21% — the highest of any period tested

### High-Volatility Era: 1986–1994

| Metric | AAPIS™ | SPY | UPRO |

|---|---|---|---|

| CAGR | 11.61% | 9.62% | 16.71% |

| Sortino | 5.20 | 1.32 | 1.08 |

The Sortino of 5.20 — nearly 4× SPY's 1.32 — in the post-Black Monday era is the standout. Despite using proxy data for both the volatility index and the equity instrument during the pre-1993 period, AAPIS™ produced dramatically superior risk-adjusted returns. 1987 (Black Monday's year): zero green signals fired, AAPIS™ held SPY, UPRO lost 32.92%.

---

## 8. Rolling 30-Year Windows

Four consecutive 30-year rolling windows were tested, each shifting the start date by one year. This tests whether the strategy's edge is anchored to a specific historical period or represents a stable structural property.

| Window | AAPIS™ CAGR | SPY CAGR | Multiple | Sortino | Beat SPY % |

|---|---|---|---|---|---|

| 1986–2016 | 18.24% | 9.51% | 1.92× | 0.93 | 77.4% |

| 1987–2017 | 19.93% | 9.63% | 2.07× | 1.01 | 80.6% |

| 1988–2018 | 19.68% | 9.33% | 2.11× | 1.10 | 80.6% |

| 1989–2019 | 21.07% | 10.04% | 2.10× | 1.06 | 83.9% |

Key finding: The CAGR multiple over SPY (1.92×–2.11×) and the Sortino advantage (roughly 2×) are stable across all four windows regardless of start year. Each window contains the full dot-com bust (2000–2002) and the Global Financial Crisis (2008). The edge persists in all of them.

This is the strongest available evidence against start-date dependency. An overfit strategy would show meaningful degradation when the window is shifted by one year. AAPIS™ does not.

---

## 9. Formal Robustness Testing

Ambika Analytics LLC published a formal 29-test mathematical robustness suite (May 2026) covering the 1994–2026 period (32.4 years, 65 confirmed entries).

Overall result: 27 PASS · 2 ADVISORY · 0 FAIL

### Confirmed 1994–2026 Performance (Formal Audit)

| Metric | AAPIS™ | SPY | UPRO |

|---|---|---|---|

| Total Return | 81,317% | 2,733% | 42,604% |

| CAGR | 22.52% | 10.66% | 20.15% |

| Sortino | 1.281 | 0.608 | 0.804 |

| Years ≥ SPY | 84.8% (28/33) | — | — |

| Terminal wealth vs SPY | 28.7× | 1.0× | 15.6× |

### Test Results by Category

#### Mathematical Correctness

All core calculations, execution logic, and state transitions verified as mathematically correct. Sortino formula verified to 9 decimal places. TSL execution logic (both gap-open and intraday trigger) verified against full audit log. State machine integrity confirmed across 500 random state transition tests — zero illegal transitions.

#### Curve-Fitting and Overfitting

- Signal frequency: Green signal fires on 3.2% of trading days — neither always-on nor effectively off

- Gate parameter sensitivity: All five gates show stable signal frequency across their parameter ranges (spread: 0.8%–2.8%). No cliff-edge dependencies on specific parameter values

- Gate independence: Mean pairwise |r| = 0.118, maximum = 0.208. Gates are genuinely independent

#### Statistical Robustness

- Regime sensitivity: Signal fires 193.76× more often in expansions than contractions — directionally correct and statistically confirmed

- Sortino/Sharpe ratio: 1.40× — within the expected 1–3× band for asymmetric return strategies, confirming the Sortino is not being gamed

#### Walk-Forward Validation

- In-sample (1992–2015) / Out-of-sample (2016–2026): OOS Sortino retains 103% of IS Sortino on average across 2,000 simulated trials

- 75% of trials retain ≥50% of IS Sortino — the strategy does not collapse out of sample

#### Signal Quality

- Win rate: 67.7% (44/65 entries produce positive returns in green phase)

- Statistical significance: Binomial test p = 0.003 (95% CI: 56.9%–78.5%)

- The gates are selecting entries that are better than random at the individual trade level

#### Exit Mechanism

- The trailing stop-loss is the dominant exit mechanism — active and surgical, not a decorative backstop that never engages

- The remainder of exits are triggered by overnight gap-downs below the stop floor, confirming the exit is not triggering prematurely on intra-day noise

- The annualised return generated per day of leveraged exposure substantially exceeds passive benchmarks, confirming exposure is concentrated into genuinely high-quality windows

#### Bear Market Behaviour (T28)

| Year | AAPIS™ | SPY | vs SPY | vs UPRO |

|---|---|---|---|---|

| 2000 | −15.64% | −9.74% | −5.90pp ❌ | +22.89pp ✅ |

| 2001 | −14.88% | −11.76% | −3.12pp ❌ | +26.23pp ✅ |

| 2002 | −26.24% | −21.58% | −4.66pp ❌ | +34.94pp ✅ |

| 2008 | −38.02% | −36.80% | −1.22pp ❌ | +40.96pp ✅ |

| 2022 | −9.64% | −18.18% | +8.54pp ✅ | +47.20pp ✅ |

Honest characterisation: AAPIS™ does not reliably outperform SPY in crash years. In four of five bear years it underperformed SPY slightly due to failed green entries. Its protection is against the leveraged alternative — saving 23–47 percentage points vs UPRO in every bear year. Investors must understand this distinction before subscribing.

#### Signal Concentration (T29)

- Top single entry (1995, +154.88%) accounts for only 8.7% of total compounded gain

- Excluding the top 3 annual years: mean annual return still 19.43% vs SPY 12.19%

- No single entry or year is necessary for the strategy to substantially outperform SPY

### Open Advisories

T15 — Sample Size (Trades-to-Parameters Ratio)

65 confirmed entries / 15 parameters = 4.3× ratio. Recommended minimum: 30×. This is a structural property of any strategy firing ~2 times per year — it cannot be resolved by extending the backtest further. Resolution requires live trade accumulation only. This is the most important honest limitation of the strategy.

T21 — Exit Mechanism Cross-Regime

A marginally tighter TSL threshold produces higher per-segment returns in isolated volatility regime simulation. This was consciously rejected — the architecture prioritises minimising time in leveraged exposure over maximising per-segment return. The advisory remains as a documented record.

---

## 10. Monte Carlo & Stress Testing

### Historical Monte Carlo (T25): 3,000 Iterations, 1994–2026

| Metric | AAPIS™ | Random Mean | Z-Score | p-value | Beat by random |

|---|---|---|---|---|---|

| CAGR | 22.38% | 12.48% | 4.32σ | p < 0.000008 | 0 / 3,000 (0.00%) |

| Sortino | 1.21 | 0.682 | 3.54σ | p = 0.0002 | 5 / 3,000 (0.17%) |

Zero of 3,000 random signal iterations at identical entry frequency matched AAPIS™ CAGR. The random maximum was 19.95% — 2.43 percentage points below AAPIS™. This is the core statistical proof that the strategy's edge is not reproducible by luck.

### Forward Volatility Stress Test: 5,000 Iterations × 6 Regimes

| Regime | Vol Level | Beat AAPIS™ | Verdict |

|---|---|---|---|

| Low-Vol Calm | 0.8× historical σ | 0 / 5,000 (0.00%) | ROBUST |

| Base Regime | 1.0× historical σ | 0 / 5,000 (0.00%) | ROBUST |

| Elevated Vol | 1.3× historical σ | 8 / 5,000 (0.16%) | ROBUST |

| High Clustered Vol | 1.6× historical σ | 72 / 5,000 (1.44%) | HOLDS |

| Extreme Shock Vol | 2.0× historical σ | 168 / 5,000 (3.36%) | HOLDS |

| Tail-Risk Scenario | 2.5× historical σ | 308 / 5,000 (6.16%) | WEAKENS |

The edge begins to weaken only at 2.3× historical volatility — a level with no sustained historical precedent (more extreme than 2020 COVID sustained for a full year).

GARCH regime-switching test: 0 / 5,000 random 10-year forward paths matched AAPIS™ CAGR, even with annual regime switching between high-vol and low-vol years.

### Brutal Synthetic Stress Test: 1,000 Independent Decades

Synthetic market engine calibrated to be harsher than any period in modern history — crisis regime (σ = 4.2% daily volatility, mean daily return −0.18%) applied frequently as a recurring feature of every simulated timeline.

| Metric | AAPIS™ | RAND* | SPY | UPRO (buy & hold) |

|---|---|---|---|---|

| 10-Year Win Rate vs SPY | 68.9% | 57.2% | — | — |

| Mean CAGR | +6.78% | +5.73% | +4.95% | −9.69% |

| Sortino Ratio | 0.46 | 0.41 | 0.38 | 0.14 |

| Avg Max Drawdown | 53.18% | 54.57% | 49.38% | 92.86% |

RAND = same entry mechanics as AAPIS™ but randomly timed entries at same frequency

Signal vs. exposure decomposition:

- Structural UPRO exposure premium over SPY: ~7.2 percentage points of 10-year win rate

- Pure AAPIS™ signal timing alpha: ~11.7 percentage points (62% of total edge)

The timing signal contributes 62% of the strategy's total edge over SPY. AAPIS™ is not simply a leveraged beta story. In this harshest-possible simulation, buy-and-hold UPRO produced −9.69% mean CAGR and 92.86% average max drawdown. AAPIS™ using the same instrument produced +6.78% CAGR and 53.18% max drawdown.

---

## 11. Honest Weaknesses

These are documented, not speculated:

### 1. Bear market underperformance vs SPY

In 4 of 5 historical bear years, AAPIS™ slightly underperformed SPY. Green signals fired in technically valid conditions that nonetheless preceded further deterioration. The leverage amplified losses before the TSL triggered. This is the strategy's most important limitation for investors to understand. AAPIS™ is not a defensive strategy relative to SPY. It is a selective leverage strategy.

### 2. T15: 4.3× trades-to-parameters ratio

65 confirmed entries against 15 parameters falls below the recommended minimum of 30×. This is structural — at ~2 entries per year, reaching 30× requires approximately 97 more years of operation. It cannot be resolved by backtesting. It is resolvable only through live trade accumulation. Every completed live green signal directly addresses this.

### 3. Choppy/whipsaw markets

1988 was the worst relative year: AAPIS™ −2.10% vs SPY +12.40%. Five green signals fired and four were stopped out at small losses. Post-crash markets that oscillate without trend are the most challenging environment for this type of strategy.

### 4. False early-bear entries

The dot-com period (2000–2002) shows AAPIS™ underperforming SPY by ~9.5 percentage points compounded across three years — the strategy's weakest historical episode. Three failed green entries in successively deteriorating conditions cost real money each time, even though the TSL contained each individual loss.

### 5. Live track record length

As of June 2026, the live track record is approximately 5 months. One green signal has fired and completed. Everything else is backtest — however rigorous. The year-end 2026 annual publication will be the first meaningful live data contribution.

### 6. Drawdown profile

AAPIS™ carries roughly 3.8 percentage points more average maximum drawdown than SPY in synthetic stress testing. In the GFC year 2008, AAPIS™ lost −38.02% — similar to SPY's −36.80%. Investors must be prepared to hold through SPY-comparable drawdowns.

---

## 12. Robustness vs. Overfitting — Final Verdict

### Evidence FOR Robustness

| Test | Result | What It Proves |

|---|---|---|

| 4 consecutive 30-year rolling windows | CAGR multiple 1.92–2.11× in every window | Edge is not start-date dependent |

| Monte Carlo (3,000 runs) | 0/3,000 random runs match CAGR | Edge is not reproducible by luck (p < 0.000008) |

| Walk-forward OOS | OOS Sortino = 103% of IS Sortino | Strategy does not collapse out of sample |

| Gate independence | Mean pairwise |r| = 0.118 | Five genuinely independent signals |

| Parameter sensitivity | 0.8%–2.8% spread across all gates | No brittle cliff-edge dependencies |

| Win rate significance | 67.7% win rate, p = 0.003 | Gates select better-than-random entries |

| Regime discrimination | 194× more signals in expansions | Gates are regime-aware, not regime-blind |

| Synthetic data validation | Statistically indistinguishable from real UPRO | Pre-2009 backtest era is valid |

| Forward stress test | Edge holds through 2.0× historical vol | Not dependent on calm historical conditions |

| Worst decade survival | Positive total return 2000–2009 | Survives the hardest regime in modern history |

### Evidence Against (Open Items)

| Item | Nature | Resolution |

|---|---|---|

| T15: 4.3× trades-to-parameters ratio | Structural — cannot be resolved by more backtesting | Live trade accumulation only |

| T21: Exit mechanism cross-regime | Consciously rejected on architectural grounds | Ongoing monitoring |

| 5-month live track record | Insufficient live data | Year-end 2026 publication + continued operation |

### Verdict

AAPIS™ is robust. It is not overfit.

Every direct test of overfitting passes. Gates are independent, parameters are stable under perturbation, the signal is statistically irreproducible by random entries at p < 0.000008, and the CAGR multiple over SPY is stable across four consecutive 30-year windows.

The T15 advisory is the sole substantive open item. It is not evidence of overfitting — it is a structural limit on statistical certainty about parameter precision that is shared by every low-frequency strategy in existence. It is honest. It is resolvable only through time and live trades.

The bear market finding is equally important to state clearly: AAPIS™ is not a defensive strategy relative to SPY in crash years. Its protection is against the far worse outcome of passive UPRO. Investors who understand this distinction and have a genuine long-term orientation are the appropriate audience.

---

## 13. Key Metrics Reference Table

| Metric | Value | Context |

|---|---|---|

| Backtest period | 1986–2026 (40 years) | Includes pre-VIX proxy era |

| Formal audit period | 1994–2026 (32.4 years) | Confirmed entry/exit log |

| Confirmed entries (1994–2026) | 65 | Live-auditable trade count |

| Average entries per year | ~2 | Low-frequency by design |

| Average green phase duration | Weeks, not months | Surgical, not passive |

| TSL level | Proprietary | Locks in a growing floor of captured gains |

| 40-year CAGR vs SPY | 20.58% vs 10.71% | ~1.92× multiple |

| 40-year Sortino vs SPY | 1.17 vs 0.56 | ~2.09× multiple |

| Beat SPY % (40 years) | 82.9% | 34 of 41 years |

| Win rate (65 entries) | 67.7% | p = 0.003 vs 50% chance |

| Monte Carlo p-value | p < 0.000008 | 0/3,000 random runs match |

| Gate independence | |r| = 0.118 mean | Genuinely independent signals |

| Regime discrimination | 194× | Expansions vs contractions |

| Live track record start | February 2026 | ~5 months as of report date |

| Cryptographic audit | SHA-256 dual-layer | Live GitHub ledger |

---

## 14. Important Disclosures

Data sources: Market data sourced from publicly available financial data providers. All data downloaded fresh at time of each backtest run.

Dividend treatment: All returns are total return basis (dividends included). SPY dividend distributions are fully reflected in backtest returns.

Pre-2009 UPRO simulation: UPRO launched June 2009. Pre-2009 returns are simulated using 3× daily S&P 500 returns minus the standard UPRO expense ratio. Statistical tests confirm the simulation produces distributions statistically indistinguishable from live UPRO data.

Parameter freeze date: January 1, 2026. All parameters applied retroactively to the full backtest history.

Not investment advice: This report is an independent analytical summary. It does not constitute financial advice, a recommendation to invest, or a solicitation. AAPIS™ subscribers should conduct their own due diligence. Past performance does not guarantee future results.

Copyright: AAPIS™ is a trademark of Ambika Analytics LLC. © 2017–2026 Ambika Analytics LLC. All rights reserved. This analytical report is prepared independently and does not reproduce proprietary gate parameters or source code.

---

Connect

Empowering retail investors with smart data analytics based strategies to democratize equity ownership.

Email

info@acceleratedindexing.com

© 2017 - 2026. All rights reserved.