Skip to content

About Mainline

Product transparency

Mainline is a personalized chess training program that adapts as you play. Every recommendation carries an explicit evidence grade.

The thesis

No training activity has been proven to cause a measured chess rating gain. Mainline helps you train smarter on the best available evidence. It never promises you a rating.

Every chess-specific study is observational or correlational: we know what strong players do differently, but we cannot prove that copying those activities will raise your rating. Mainline is built around that fact, not around hiding it.

The vision

The training-program layer, not another tool.

Mainline is not another puzzle trainer, game analysis tool, or spaced-repetition deck. This app sits one layer up: it is the orchestration layer that decides what you should work on, with which resources, and why, then revises that plan continuously as you train and play.

Engagement and progress

Built to support consistency, not to extract attention.

Progress in Mainline means training signals: whether you are showing up, completing the planned work, keeping reviews healthy, and building skill estimates with uncertainty. Rating is noisy, and no activity here is treated as proven to cause rating gain.

The engagement layer exists because consistency is part of training. It uses forgiving reminders, capped streak cycles, and competence feedback to make practice easier to resume. It does not use ads, leaderboards, shame, unbreakable streaks, or paywalled training quality.

The boundaries

What this app deliberately is not.

Every exclusion is a design decision with a reason.

  • No LLM/AI at runtimeGenerative AI plays chess poorly and introduces opacity. Mainline uses deterministic algorithms and local Stockfish. You can verify every decision.
  • No competing game platformLichess and Chess.com provide great play servers. Mainline connects to external platforms rather than replacing them.
  • No hosted copyrighted contentBooks and courses are recommended and logged, never hosted. Mainline points you to resources you own.
  • No social or multiplayerNo leaderboards, chat, or comparative rankings. Social comparison harms long-term practice habits.
  • No self-reported skill diagnosisSelf-assessment in chess is prone to error. Mainline measures play behaviorally from your games and calibration puzzles.
  • No infinite streaksUnbreakable streaks create loss aversion and burnout. Mainline caps streak cycles and forgives missed days.
  • No global leaderboardsComparative leaderboards harm motivation for most learners. Mainline focuses strictly on personal training consistency.
  • No puzzle volume chasingCorrelation between raw puzzle volume and rating is near zero. Spaced repetition and deliberate calculation matter more than quantity.
  • No opening memorization for beginnersBeginner and intermediate games are decided by tactical blunders. Time is better spent on pattern recognition and calculation.

The evidence framework

Borrowed from the board: every claim is annotated.

Every recommendation, methodology value, and claim on this page carries a grade: an unverified assumption can never pose as established fact.

Grade A

Strong, replicated

Used for robust, replicated findings such as retrieval-practice and spacing effects.

Grade B

Suggestive, limited

Suggestive studies with limited sample size, context, or generalizability.

Grade C

Theory / best-guess

Logical inference, provisional rule, or calibration estimate. Treated as a starting point, never a proven prescription.

Grade D

Contradicted myth

Popular chess advice that Mainline avoids because evidence contradicts it.

Grade answers how strong the science is. A second axis, confidence, answers a different question: how much of your own data backs this specific call to you. The same Grade-A finding can land with low confidence, as when we know spaced repetition works but you've only imported three games. Or it can land with high confidence, well-backed by your own play. The distinction keeps a band prior from masquerading as a personalised verdict.

Insufficient

Not enough data yet

We do not have enough of your games or reviews to make this call. Mainline displays uncertainty plainly instead of guessing.

Low

Population baseline

The recommendation rests on what players at your level tend to need, not on your personal data yet. It sharpens as you train.

Medium

Partial data

Partially grounded in your games or reviews. A working hypothesis under active refinement.

High

Strong personal data

Drawn from enough of your own play to represent your specific strengths and leaks.

Where the science is now

Honest about the current state.

Active methodology · research-1.4.0

This active research release encodes the approved methodology values and copy, while retaining every documented best guess and deliberate stub as evidence-labeled data.

The release is reproducible: historic programs keep the version and rationale snapshot they were generated with. The central caveat remains unchanged: no training activity has been proven to cause a measured rating gain.

Still deliberately unresolved

  • · Training-fit feedback is not evidence of chess skill and no adherence or rating effect is claimed.

Aggregate basis

None. P9 enables controlled observational export, but no Mainline aggregate has informed this release.

Methodology release history

Every change keeps its limits and rollback path.

research-1.4.0 · 2026-07-15

Adds the sparse P8 training-fit prompt policy and restricts subjective fit to a positive tie-break inside the evidence-led daily mix.

Aggregate basis: None. P9 enables controlled observational export, but no Mainline aggregate has informed this release.

Evidence changes:

  • · Added Grade C training-fit policy values and retained the Grade A boundary separating self-report from behavior.

Limitations:

  • · Prompt timing and positive tie-break effects are unvalidated product best guesses and do not establish adherence or rating effects.

Rollback: Restore research-1.3.0 if prompts or fit ordering cause operational harm; preserve 1.4.0 artifacts for replay.

research-1.3.0 · 2026-07-12

Keeps the three-item calibration and revises weekly-focus alternative copy so optional user choice is clear without weakening its evidence caveat.

Aggregate basis: None. No Mainline observational aggregate informed this release.

Evidence changes:

  • · Clarified user-facing rationale without changing its Grade C, Tier 2 evidence.

Limitations:

  • · Alternative-choice framing has not been validated for chess training adherence.

Rollback: Restore research-1.2.0 if the revised optional-choice copy is misleading.

research-1.2.0 · 2026-07-12

Shortens new-user calibration to one three-item tactical track while preserving historic assessment behavior by methodology version.

Aggregate basis: None. No Mainline observational aggregate informed this release.

Evidence changes:

  • · No evidence grade changed; the shorter calibration is explicitly Grade C.

Limitations:

  • · Three calibration items may trade completion against measurement quality and require observational review.

Rollback: Restore research-1.1.0 for new assessments; never rescore historic assessments silently.

research-1.1.0 · 2026-07-12

Adds stable weekly focus selection, confidence-gated revision, and bounded process-goal alternatives.

Aggregate basis: None. No Mainline observational aggregate informed this release.

Evidence changes:

  • · Added Grade C weekly-focus policy values without upgrading the underlying evidence.

Limitations:

  • · Focus stability and bounded alternatives are product best guesses, not demonstrated causes of adherence or rating gain.

Rollback: Restore research-1.0.0 if weekly-focus selection is operationally unsafe; historic programs retain 1.1.0.

research-1.0.0 · 2026-07-10

First research-channel methodology release for the current Phase 1 seams.

Aggregate basis: None. No Mainline observational aggregate informed this release.

Evidence changes:

  • · Published the reviewed research synthesis with its existing grades and citations; no evidence was upgraded.

Limitations:

  • · Chess-specific activity effects remain observational or extrapolated and cannot establish rating causation.

Rollback: Restore stub-0.1.0 only as an owner-reviewed operational rollback while preserving historic program versions.

stub-0.1.0 · 2026-06-21

Pre-release initial configuration retained for historic programs.

Aggregate basis: None. No Mainline observational aggregate informed this release.

Evidence changes:

  • · None. This was a pre-release initial configuration.

Limitations:

  • · Not an evidence-complete methodology release.

Rollback: Historic only. Do not activate without owner review.