Kinetic Alpha Research
AI in Investment & Wealth Management · Part I of a series
Series Launch · AI in Wealth Management · Part I

From 60/40 to Autonomous Agents

Seventy years of portfolio construction compressed into three regimes: the policy portfolio, the robo-advisor, and now the AI agent. This piece maps the evolution, surveys what today's AI offerings actually do — versus what their marketing says — and frames the question the rest of this series will attack: how do you prove a model is behaving inside your risk tolerance?

−16%
Globally diversified 60/40 return in 2022 — the first year in the modern era stocks and bonds both fell double digits
Vanguard
$2.5T→$5.9T
Robo-advised assets, 2022 actual to 2027 projected
PwC Global AWM Survey
98%
Morgan Stanley advisor teams using the firm's GPT-4 research assistant
Morgan Stanley / OpenAI
62%
US retail investors who have used AI tools to inform investment decisions
Investing.com survey, Apr 2026 (n=938)
52%
Finance firms already operating agentic AI systems
Cambridge CCAF, cited by Bank of England, Jun 2026
81%
Failure rate of GPT-4-Turbo with retrieval on FinanceBench SEC-filing questions — why validation architecture matters
Patronus AI, arXiv 2311.11944

01Three regimes of portfolio construction

The story of modern wealth management is usually told as continuous progress. It reads better as three distinct regimes, each defined by what the industry chose to automate — and each ending when its central assumption broke.

The policy portfolio (1952–2008)

Harry Markowitz's 1952 "Portfolio Selection" gave the industry its operating system: diversification as a computable trade-off between expected return and variance. ERISA (1974) institutionalized it, and by the 1980s the 60/40 stock/bond mix had become the default policy portfolio for pensions and, through balanced funds, for households. The track record justified the default. A US 60/40 blend delivered roughly 8.3% annualized over the trailing 30 years with single-digit volatility, and — measured at calendar year-ends since 1928 — never posted a negative 10-year return. The human advisor sat on top of this machinery, charging roughly 1% of assets largely for allocation, discipline, and hand-holding.

The regime's central assumption was the stock-bond hedge. 2022 broke it: with inflation forcing rates up, a globally diversified 60/40 fell about 16% — by some measures the strategy's worst year since 1937, and the first modern year in which both stocks and bonds fell by double digits simultaneously. The portfolio wasn't dead (it returned roughly +30% cumulatively over the following two years, and Vanguard still defends the construct while tilting its 2026 model toward bonds), but the episode ended 60/40's claim to being a complete answer. It also created the market opening every AI pitch now walks through: if the static allocation can fail, perhaps the allocation should think.

The automation regime (2008–2022)

Betterment (founded 2008, launched 2010) and Wealthfront (launched as a robo in 2011) automated the policy portfolio rather than rethinking it: risk questionnaire in, MPT-optimized ETF portfolio out, with threshold rebalancing and — from 2012 — automated tax-loss harvesting. The innovation was economic, not intellectual. At ~25 basis points versus ~100 for a human advisor, robo-advice quadrupled the addressable market and forced fee deflation across the industry.

Two things are worth being precise about, because they shape how to read today's AI claims. First, robo-advisors were never AI — they were deterministic rules engines executing 1950s portfolio theory. No learning, no language, no inference. Second, the pure-play economics mostly failed. Goldman's Marcus Invest was sold to Betterment (2024); JPMorgan, UBS, and US Bank all shut their robos (2024–2025); BlackRock wound down FutureAdvisor (2023). Wealthfront survived to a December 2025 IPO — but its S-1 revealed that roughly three-quarters of revenue now comes from cash-management spread, not advisory fees. The survivors are those attached to a giant (Vanguard Digital Advisor at ~$312B including hybrid programs, Schwab, Empower) or those that found a different monetization. Automation compressed fees faster than it created margin.

The intelligence regime (2022–)

ChatGPT's November 2022 launch — the same month the 60/40 was completing its worst modern year — marks the boundary. What changed was not portfolio math but the interface and the scope: models that read documents, hold conversations, personalize at scale, and increasingly act. The adoption curve since has been steep on the sell side (98% of Morgan Stanley advisor teams on the firm's GPT-4 assistant; 95% of wealth and asset managers with generative AI scaled into multiple use cases per EY's 2025 survey) and surprisingly steep on the retail side (62% of surveyed US retail investors have used AI to inform decisions). The frontier moved again in 2025–2026 with agentic execution: Robinhood's sandboxed agentic-trading accounts (100,000+ funded since May 2026) and broker APIs like Alpaca's MCP server, where LLM agents — not humans — now drive order flow.

Each regime automated the previous one's scarce resource. The policy portfolio automated judgment into math; the robo automated the math into product; AI is now automating the conversation, the research, and — at the frontier — the decision itself. Which is precisely why the control question, not the capability question, is where this series will spend most of its time.

Seventy years, three regimes
Milestones in the evolution from static allocation to autonomous agents. Hover any row for detail.
Policy portfolio · 1952–2008
Automation · 2008–2022
Intelligence · 2022–
    Era assignment reflects the dominant production technology of the period, not the disappearance of prior regimes — 60/40 allocation and robo rebalancing both remain in force inside AI-era products.
    The automation regime, measured
    Reported assets under management, $B — the two flagship independent robo-advisors
    Wealthfront Betterment
    Company-reported AUM, year-end unless noted (Wealthfront Jun 2026; Betterment May 2026). Wealthfront's post-2022 acceleration is driven substantially by cash-management assets — its S-1 attributes ~74% of FY2026 revenue to cash products, not advisory fees. Sources: company disclosures compiled by investingintheweb.com; Wealthfront S-1.

    02One technology, three very different buyers

    This series will ultimately treat institutional allocators, advisory practitioners, and retail investors as three separate discussions — the requirements barely overlap. As an opening frame, here is what each constituency is actually buying today, what it must demand, and where its specific risk sits. Later pieces expand each column into its own evaluation.

    Institutional / Allocator

    PMs, CIOs, risk officers, consultants evaluating AI-driven strategies and tooling
    What AI does today

    Research synthesis at scale (AlphaSense, Hebbia, Rogo), portfolio-analytics copilots (Aladdin Copilot, MSCI AI Portfolio Insights, Bloomberg Document Insights), and genuine ML alpha generation inside quant and multi-strategy funds (Bridgewater's ~$2B AIA fund, Point72's Turion).

    Requirements

    Model inventory and validation lineage; explainability sufficient for an investment committee; diligence on foundation-model dependence (a vendor's model-version bump is a model change); data licensing and audit trails.

    The specific risk

    Backtest credibility and crowding. A 2026 survey of LLM-trading studies found only 2 of 19 had clean time-consistent data splits; regulators' emerging worry is correlated agent behavior — herding at machine speed.

    Advisor / Practitioner

    RIAs, wirehouse FAs, private banks deciding what to adopt and how to supervise it
    What AI does today

    The productivity layer: meeting capture and CRM automation (Jump, Zocks, Morgan Stanley Debrief, Merrill's Meeting Journey), plan generation (Conquest), proposals (Powder), next-best-action (TIFIN AG), and household-level agents (Savvy Intelligence).

    Requirements

    FINRA 3110 supervision extended to AI output; books-and-records for prompts and drafts; client consent for recording; human review before anything reaches a client; Marketing-Rule discipline on what you claim your AI does.

    The specific risk

    Paying for AI without a data foundation — 64% of wealth firms lack a unified data layer, and ROI remains "elusive" per F2 Strategy's 2026 survey. And fiduciary duty cannot be delegated to a model: the advisor owns every recommendation the machine drafts.

    Retail Investor

    Individuals choosing between robos, AI advisors, and agentic trading products
    What AI does today

    Portfolio digests and screening (Robinhood Cortex), holistic advisory recommendations (PortfolioPilot), AI-built custom indexes (Public's Generated Assets), and — newest — sandboxed agents that execute trades (Robinhood Agentic Trading).

    Requirements

    Know the autonomy level you're granting; know whether "assets on platform" means managed or merely linked; know who executes and who is accountable; verify AI claims — the SEC's first AI-washing cases hit consumer-facing advisers.

    The specific risk

    Confidently wrong output and unaudited performance marketing. Studies put LLM error rates on personal-finance questions near 35%; "70% win rate" claims from AI stock-picker vendors are unaudited; and the "self-directed" legal framing shifts responsibility for agent behavior onto the user.

    A definitional line this series will hold: automation is not intelligence. Threshold rebalancing and tax-loss harvesting are automation — valuable, deterministic, testable. "AI" should mean systems that infer, generalize, or generate. Much current marketing launders the former as the latter; the SEC's AI-washing docket exists precisely because the distinction has economic value.

    03The autonomy ladder

    The most useful single axis for evaluating any "AI investing" product is not model quality — it is how much decision authority the system holds. Nearly everything shipped at scale today sits on the bottom two rungs. The interesting engineering, and all of the interesting risk, lives in the climb.

    L0Static
    Rules-based automation

    Deterministic MPT allocation, threshold rebalancing, tax-loss harvesting. The robo-advisor stack. No inference.

    Control model: conventional software QA + disclosure accuracy (the Schwab cash-drag case shows where that fails).
    L1Informational
    AI that explains

    Digests, summaries, research Q&A: Schwab Portfolio Insights, Bloomberg Document Insights, Aladdin Copilot, meeting notetakers. Human decides everything.

    Control model: retrieval grounding to curated corpora, citation requirements, hallucination evals, human review before client contact.
    L2Advisory
    AI that recommends

    Personalized, actionable recommendations the user executes: PortfolioPilot, Arta AI, Cortex trade ideas, Conquest's plan engine. This is where fiduciary and suitability questions become live.

    Control model: suitability mapping, recommendation logging, conflict analysis, output classifiers on advice language.
    L3Supervised execution
    AI that acts inside a fence

    Agents execute within hard deterministic constraints: Robinhood's sandboxed pre-funded agentic accounts, LLM agents on Alpaca's MCP rails, advisor-supervised rebalancing (Vise). The frontier of what ships to retail in 2026.

    Control model: pre-trade compliance engines, position/loss limits, rate limits, kill switches outside the agent's control plane — the 15c3-5 pattern.
    L4Discretionary
    AI that manages

    Model holds the mandate. Exists today only inside quant/multi-strat funds with institutional risk infrastructure (Bridgewater AIA, Point72 Turion, Voleon). Not currently permissible as a retail advisory product — "RIAs aren't allowed to hire an AI agent to manage money… it doesn't suit the rules currently" (TradePMR's Robb Baldwin).

    Control model: full model-risk lifecycle — independent validation, champion-challenger, drift monitoring, capital-at-risk limits, human accountability chain.

    04The current landscape: what's actually on offer

    Below is a working map of the offerings that define the space as of August 2026 — filterable by market segment and by rung on the autonomy ladder. Two honest caveats baked into the data: platform-reported "assets" for retail AI tools generally mean linked or analyzed assets rather than discretionary AUM, and several vendors listed have unaudited performance claims or enforcement history, flagged inline.

    Segment
    Autonomy
    Offering Segment Autonomy What it actually does Scale

    Scale figures are the most recent company-reported or press-reported numbers as of Aug 2026; bases differ (AUM vs linked assets vs users vs adoption) and are labeled per row. Hover a row for sourcing and caveats.

    What the map says

    Three patterns stand out. First, the shipped product is overwhelmingly L1–L2. The scaled deployments — Morgan Stanley's assistant, JPMorgan's LLM Suite, Merrill's Meeting Journey, Schwab's Portfolio Insights — are drafting and explanation layers with a human firmly in the loop. Discretion remains confined to quant funds and sandboxes. Second, the guardrail pattern is converging across every serious deployment: consent-gated inputs, retrieval restricted to curated corpora, human review before client contact, sandboxed and pre-funded execution, "self-directed" legal framing. Firms discovered the same containment architecture independently because the regulatory physics is the same. Third, the economics rhyme with the robo era. Budgets are surging while measured ROI stays thin (F2 Strategy, 2026), notetaker startups are raising dueling Series Bs while platforms absorb their feature set, and at least one pure-play AI robo (Q.ai) has already shut down. The capability is real; the business models are still sorting.

    05The control problem: proving the model stays inside your tolerance

    "Does it generate alpha?" is the question everyone asks. "How do I know it's behaving inside my risk tolerance — and how would I prove it?" is the question that determines whether any of this is investable. Two facts frame the answer in 2026, and both are uncomfortable.

    Fact one: the regulatory scaffolding just got thinner, not thicker. The SEC withdrew its only AI-specific rule proposal (the predictive-data-analytics conflicts rule) in June 2025. In April 2026, the Fed, OCC, and FDIC replaced SR 11-7 — the model-risk bible since 2011 — with shorter, principles-based guidance that explicitly excludes generative and agentic AI from scope, promising a future request for information instead. FINRA's position is technology-neutral ("the rule is the rule, no matter how the method changes"), the EU AI Act does not classify investment advice as high-risk, and its high-risk deadlines are themselves slipping under the Digital Omnibus. Meanwhile every SEC AI enforcement action to date — Delphia, Global Predictions, Rimar — punishes lying about AI, not AI misbehaving. The burden of making AI safe inside a mandate has been left, almost entirely, to the firms deploying it.

    Fact two: the raw models are not trustworthy enough to skip the engineering. On FinanceBench — real questions against real SEC filings — GPT-4-Turbo paired with a retrieval system failed 81% of the time; with idealized retrieval, accuracy jumped to ~89%. That ~70-point swing is the single most important number in AI wealth management: it says the control point is the retrieval and validation architecture around the model, not the model choice. The same lesson generalizes to execution: a 2026 survey of LLM-trading research found essentially none of the published performance claims auditable (2 of 19 studies with clean data splits, 1 of 19 with transaction costs).

    The architecture that answers the risk-tolerance question is old, proven, and borrowed from electronic trading: the stochastic model proposes; a deterministic layer disposes. Knight Capital's $460M, 45-minute self-destruction in 2012 — and the Market Access Rule (15c3-5) enforcement that followed — established the template: hard pre-trade checks, independent kill paths, deployment controls, all sitting outside the thing being controlled. Applied to an AI manager, the stack looks like this:

    Stochastic coreLLM / ML model generating research, recommendations, or orders — assumed fallible by design
    Layer 1 · Deterministic mandate enforcementMachine-readable IPS: position & concentration caps, asset-class ranges, restricted lists, leverage limits, tracking-error and VaR budgets, drawdown de-risking triggers. Pre-trade engines hard-block violations regardless of why the order was generated.
    Layer 2 · Validation regimeBacktest hygiene (deflated Sharpe, combinatorial purged CV), champion–challenger promotion, shadow/paper trading periods, groundedness and hallucination evals run continuously — not once at launch.
    Layer 3 · Runtime containmentOrder-rate and budget throttles, anomaly detection on order flow vs. historical distribution, human-escalation tripwires, kill switches outside the agent's control plane.
    Layer 4 · GovernanceThree lines of defense; AI use-case inventory (shadow AI is a top exam finding); model-version change management — a foundation-model bump is a model change; immutable audit trail of prompt, retrieval, and output; NIST AI RMF / ISO 42001 filling the vacuum SR 11-7's successor left.

    Notice what this stack implies for the buyer's diligence question. "Is your AI good?" is unanswerable. "Show me the deterministic constraint set my mandate compiles into, your eval suite and its failure rates, your last champion-challenger promotion decision, and the kill path that doesn't depend on the agent's own cooperation" — that is answerable, and today almost no retail-facing product will answer it. Closing that gap is where this series goes next. The Bank of England is already asking the systemic version of the same question: with half of finance firms running agentic AI, Deputy Governor Sarah Breeden noted in June that "our frameworks were not built to contemplate autonomous agents" — and that human-in-the-loop for every agent action is "unlikely to be realistic." The controls have to be structural.

    06Where this series goes from here

    Part I set the map. The proposed continuation — each piece standalone, each building the evaluation framework the finale needs:

    PART II

    Functional teardown: retail AI advisors

    Hands-on evaluation of PortfolioPilot, Cortex, Public's Generated Assets, Magnifi and peers: identical prompts and test portfolios, scored on recommendation quality, consistency across sessions, risk-profile adherence, and what happens when you push against the guardrails.

    PART III

    The advisor stack, audience by audience

    From notetaker to next-best-action: where the ROI actually shows up, the data-layer prerequisite, supervision and books-and-records obligations, and a build-vs-buy framework for RIAs — expanding the practitioner lens from §02.

    PART IV

    Agentic execution architecture

    The MCP-broker rails (Robinhood, Alpaca), sandbox design, and the core engineering question: how an IPS becomes a machine-readable constraint set — encoding risk tolerance as tracking-error budgets, VaR caps, concentration limits, and drawdown triggers the agent cannot argue with.

    PART V

    The validation playbook

    Proving behavior, not asserting it: eval design and hallucination benchmarks, deflated-Sharpe and CPCV backtest hygiene, champion–challenger promotion, shadow trading, drift monitoring, and model-version change management — the answer to "how do I know it's true."

    PART VI

    Governance & regulatory tracker

    The post-SR-11-7 vacuum, SEC/FINRA posture, EU AI Act drift, state regimes, and the BoE kill-switch debate — maintained as a living reference page, updated as the 2026–2027 rulemaking cycle resolves.

    PART VII

    Future state: the AI-native wealth platform

    Proposed functionality enhancements: what a platform designed around the control stack — rather than retrofitted with it — should look like, from continuous suitability to portfolio-level agent orchestration, and which incumbents are closest.

    Companion deliverables per piece: an interactive dashboard for kineticalpha.com and a LinkedIn distribution post, consistent with prior Kinetic Alpha research releases.

    07Selected sources

    1. Vanguard — "The global 60/40 portfolio: Steady as it goes" (data through Sept 2024)
    2. CFA Institute Research & Policy Center / Monash — "The Performance of the 60/40 Portfolio" (Feb 2025)
    3. Ben Carlson, A Wealth of Common Sense — 60/40 historical-drawdown series (2022–2023)
    4. CNBC — "Vanguard likes a 40/60 portfolio for 2026" (Jan 2026)
    5. PwC — Global Asset & Wealth Management Survey (Jul 2023)
    6. Wealthfront S-1 / Forbes IPO coverage (Sept–Dec 2025); investingintheweb.com robo AUM compilations (2026)
    7. The Daily Upside — "Why US Bank, UBS, JPMorgan All Shut Down Their Robo Advisors" (Nov 2025)
    8. EY — Generative AI in Wealth & Asset Management Survey (2025); Advisor360° GenAI survey (Feb 2025)
    9. F2 Strategy via InvestmentNews — "AI in wealth management: budgets surge but ROI remains elusive" (Jul 2026)
    10. Investing.com — retail AI usage survey, n=938 (Apr 2026)
    11. Morgan Stanley press releases: AI @ MS Assistant (2023), Debrief (Jun 2024), AskResearchGPT (Oct 2024); OpenAI case study
    12. Robinhood Newsroom — Cortex/Strategies launch (Mar 2025); Axios — agentic trading & prediction markets (Dec 2025); InvestmentNews — Cortex for Advisors (Jun 2026)
    13. Axios — Public.com "Generated Assets" (Nov 2025)
    14. SEC Press Release 2024-36 — first AI-washing actions: Delphia & Global Predictions (Mar 2024); PR 2024-167 — Rimar Capital (Oct 2024)
    15. SEC — withdrawal of predictive-data-analytics proposal, S7-12-23 (Jun 2025)
    16. OCC Bulletin 2026-13 — revised interagency model risk management guidance (Apr 2026); Sullivan & Cromwell and Orrick client memos
    17. FINRA — Regulatory Notice 24-09 (Jun 2024); 2026 Annual Regulatory Oversight Report, Gen-AI section (Dec 2025)
    18. K&L Gates — EU AI Act / Digital Omnibus status (Jan 2026); Covington — UK regulators' AI approach (Apr 2026)
    19. Bank of England — Deputy Governor Breeden remarks on agentic AI and kill switches, Sintra (Jun 2026); Cambridge CCAF survey
    20. Foster–Sherman letter to SEC Chair Atkins on agentic AI trading (Jun 2026)
    21. Patronus AI — FinanceBench (arXiv 2311.11944, Nov 2023); Presenc — retrieval-vs-model accuracy analysis (2026)
    22. "Agentic Trading: When LLM Agents Meet Financial Markets" — survey of 77 studies (arXiv, 2026)
    23. Lopez-Lira & Tang — "Can ChatGPT Forecast Stock Price Movements?" (arXiv 2304.07619); look-ahead-bias critique (arXiv 2309.17322)
    24. Bailey & López de Prado — "The Deflated Sharpe Ratio" (2014)
    25. SEC PR 2013-222 — Knight Capital Market Access Rule action (Oct 2013); Phil Venables — "High Frequency Trading and Lessons for Agentic AI"
    26. SEC admin proceeding 34-95087 — Schwab robo-adviser settlement, $187M (Jun 2022)
    27. BlackRock — Aladdin Copilot; Bloomberg LP — Document Insights; Hedgeweek — Bridgewater AIA fund (Jul 2024); Business Insider — hedge-fund AI adoption (Nov 2025)
    28. Alpaca — AI-agent trading; Dealroom — $435M raise, agent-driven volume 4× QoQ (2026)
    29. Funding rounds: Jump $80M (2025); Conquest $80M (Jun 2025); Farther $150M Series D (May 2026); Range $60M Series C (Nov 2025); AlphaSense $350M at $7.5B (Jun 2026)
    30. CNBC — "Don't rely on AI for personal finance advice" (Jul 2026)