Rolling Sharpe
30-day Sharpe tracked against the validation baseline. Watch and critical bands when performance drifts.
Independent validation desk
A descriptive research report. No black-box score. No investment advice.
A backtest that looks good isn’t one you can trust yet. Send us your hypothesis — we run it through an institutional, multi-gate validation gauntlet built to expose overfitting, then hand back an independent report. Typically within 24-72 business hours for Tier A scope.
What you can validate
Backtesting is how a trading idea earns trust: replaying a strategy over decades of historical market data to measure how it would have behaved — before any capital is at risk. Validraft is built for rigorous, reproducible quantitative research. You don’t need a fully coded algo: discretionary, systematic, and hybrid hypotheses all start the same way — a brief in plain English. We translate your rules into a testable specification, then run it through the same institutional engine with walk-forward and out-of-sample validation by default.
We group strategies by the type of edge they claim: price behavior first, then order flow, news/NLP, relative value, fundamentals, event catalysts, macro regimes, and alternative data. Pick an edge group to see concrete example backtests, or describe your own idea and we’ll scope it.
Don’t see your exact approach? Discretionary playbooks and partial rules are exactly what the brief is for — we scope what can be tested honestly before any compute runs.
How it works
Every brief travels the same eight-stage pipeline — from immutable, point-in-time data to a signed-off report with full lineage. Step through it, or let it walk you through itself.
STAGE 01 / 08
Brief & scope
You describe the asset, timeframe, rules, and what success looks like — in plain English. We translate it into a precise, testable specification and confirm data coverage and feasibility, so the engagement is scoped to what the data lake can genuinely support.
Validation framework
Each engagement ends with a structured validation report: what was tested, which controls ran, where the edge held up, where it failed, and what the result does not prove. No black-box score. No recommendation language.
Behind every report sits a scoped validation profile drawn from 66+ controls: data integrity, statistical validation, stress testing, risk measurement, benchmarking, and reproducibility.
Phase 4 · statistical validation
Control 4.1
A frozen candidate is scored on data never used in selection, across an explicit train/test boundary.
Post-validation monitoring
A strategy that passed validation yesterday can drift tomorrow. For ongoing engagements, we keep watching rolling Sharpe, drawdown, signal decay, and regime behaviour — benchmarked against the same gates that produced your report.
Alerts and history live in your client workspace, with a direct line back to the desk when something needs a human read.
Client workspace · preview
Opening-range breakout — NQ
Illustrative metrics — your live workspace reflects each delivered engagement.
30-day Sharpe tracked against the validation baseline. Watch and critical bands when performance drifts.
Graded alerts as peak-to-trough loss approaches stress-test envelopes from your report.
Second-half vs first-half performance comparison — early warning when the edge starts to fade.
Conditional correlation and behaviour across market states — flagged when regime context changes.
Data
Every brief is tested against the datasets your hypothesis actually needs — price, fundamentals, macro, sentiment and alternative signals, on the same immutable, point-in-time lake our research desk runs. We are actively ingesting new feeds, cleaning and normalizing them, and adding coverage once each dataset passes internal checks.
CME futures
CME futures data across continuous rolls: minute bars, ticks, L2/L3 order book, and contract metadata.

US equities
US equity cross-section with trades, minute aggregates, adjusted bars, reference, corporate actions, and processed daily bars.
Fundamentals & multi-asset markets
The broadest static snapshot in the lake: fundamentals, events, ownership, index membership, ETF holdings, FX, crypto, and COT.
Macroeconomic indicators
Vintage-aware macro data: roughly 231 FRED/ECB series with point-in-time layouts for regime and macro overlays.
News & sentiment
FMP news streams across equities, press releases, forex, macro, and crypto, then locally classified with FinBERT sentiment scoring.
Congressional trading
Congressional trading exports preserved as snapshots and normalized into event ledgers for alternative-data tests.
Coming soon
These feeds are not part of the active lake yet. They are the next provider candidates we plan to evaluate, license where required, normalize, and document before they can be used in validation.

Market news & events
Planned market-moving news, analyst, calendar, and event feeds for sentiment and event-driven validation.

Options analytics
Planned options chains, implied-volatility history, greeks, and surface features for options-aware validation.
FX & rates
Planned foreign-exchange rates and currency reference data for macro, carry, and cross-asset overlays.
Crypto market data
Planned exchange-normalized crypto trades, candles, reference, and multi-venue market data.
Polymarket & Kalshi
Validraft-built prediction-market dataset planned from market metadata, odds, liquidity, volume, and resolution history.
Tell us what the hypothesis needs. If the feed is available and licensable, we source it, ingest it into the research lake, and document the data assumptions before validation starts.
Provider names and marks are shown for identification only. Validraft is independent and not affiliated with or endorsed by the listed providers. Coverage shown reflects the current research lake; some feeds refresh on demand for scoped engagements.
Privacy & confidentiality
A trading hypothesis can be the most sensitive thing you send us. Validraft separates two ideas clearly: personal data is handled under the privacy notice, and strategy materials are treated as confidential engagement materials by default.
Data surfaces we design around
Briefs, parameters, files, reports, and monitoring notes stay tied to the engagement they came from.
Client hypotheses are not added to a public template library, marketing example, or cross-client benchmark.
Reports and workspace updates are delivered through authenticated access or agreed private channels.
Standard briefs stay lightweight; sensitive business, Tier C, or institutional scopes can move under NDA before details are shared.
Operational practices
Positioning
Validraft is built for the moment when a backtest looks promising, but you need an independent desk to test whether the evidence survives proper validation.
Validraft position
Human judgment where it matters, standardized validation where it must be repeatable.
You operate everything
Flexible if you already own the pipeline, but every data bug, look-ahead issue, and overfit parameter choice is yours to catch.
Independent validation desk
Human-scoped validation, institutional gates, and a descriptive report without turning the engagement into open-ended consulting.
Custom research project
Useful for bespoke build-outs, but often heavier, slower, and harder to compare because methodology changes engagement by engagement.
Best fit
Builders who want full control
Founders, traders, and managers who need an external validation read
Bespoke research or platform build-outs
Client effort
High
Brief-driven
High-touch
Validation gates
Whatever you implement
Standardized control stack
Variable by engagement
Delivery model
Self-service tool
Scoped report
Open-ended project
Research notes
Plain-English essays on overfitting, validation gates, data discipline, and the difference between an attractive simulation and a durable hypothesis.
First article
01 / Overfitting
The curve can be beautiful and still be wrong.
Backtest overfitting
A practical checklist for spotting curve-fitting before a beautiful equity curve becomes an expensive mistake.
Send a brief. We confirm scope, run the validation, and deliver the report. Research and simulation only.