Methodology

How Sportslyx Analytics Works

Where the data comes from, how a market line becomes a probability, what the language model is and is not allowed to decide, how every read is graded, and where the whole approach breaks down. Everything on this page describes the product as it runs today.

Inputs

Data Sources

Every number on Sportslyx traces back to a public feed. There is no proprietary data and no hidden source.

The primary feed for all seven sports is the ESPN public scoreboard. It supplies schedules, start times, live and final scores, base market lines and box scores without an API key, which is what makes it a dependable backbone: it does not expire, it does not meter, and it covers NFL, NBA, MLB, NHL, soccer, UFC and golf from one shape of data. Optional enrichment providers — The Odds API, API-Sports and football-data — add lines and competition detail where they are configured, and a free odds feed supplies lines for some markets. MLB player prop lines come from the ESPN core endpoint.

From those feeds Sportslyx builds two things: team profiles — name, abbreviation and crest, refreshed daily and holding nothing else — and a per-game package that carries statistics, recent form, head-to-head history, market lines, injuries and availability, and probable starters or goaltenders where the sport has them. Nothing game-specific is stored on a profile, and there is no player profile; the package is the only thing the language model reads. The full list of feeds, what each one provides, the refresh cadence and the known gaps are on the data sources page.

Models

Statistical Models

Four kinds of arithmetic run under the product. None of them is a neural network guessing winners; each is a specific, inspectable calculation.

Market de-vig to fair probability. A moneyline pair carries the bookmaker’s margin: the two implied probabilities add up to more than 100%. Sportslyx removes that margin proportionally across the two sides so the pair sums to exactly 100%, and the favoured side’s share is the win probability shown on the game. For soccer the read is priced on the 1X2 line: the two side prices from the three-way market are de-vigged against each other, the draw is left out of the probability, and a drawn match is graded as a push. UFC and golf are individual events; there is no home side and no venue adjustment, only the market for each fighter or the outright market for each player.

Team and player statistics. The per-game package aggregates the public statistics that matter for the sport — efficiency, pace and rest in basketball; probable starting pitchers, run scoring and pitching in baseball; goaltending, special teams and shot metrics in hockey; form, goals and shots per match and fixture congestion in soccer; striking and grappling profiles in MMA; strokes gained, recent finishes and made-cut rate in golf. A baseball analyst also thinks about the bullpen and the park, and a soccer analyst about expected goals; those are not feeds Sportslyx ingests, so they appear in a breakdown only as context, never as data. These numbers describe the matchup; they do not override the market probability. They are shown on the statistics pages for every team the feeds cover.

Studio custom weights. Studio lets a subscriber choose the statistics they believe in, assign weights and run a live game through that model. The projection is the subscriber’s own; it is saved alongside the market number and the AI read, and it is graded after the game like anything else. Studio does not add hidden factors — the weights you set are the whole model.

Combos from one score distribution. A same-game combo is priced from a single model of how the score is likely to unfold, so correlated legs — a favourite’s moneyline and an under, a run line and a team total — are priced together rather than multiplied as if they were independent. The product page on Combos explains the pricing in more detail.

Language model

AI Match Analysis

The written breakdown is produced by a large language model accessed through OpenRouter, which keeps the product provider-agnostic. The model writes; it does not price.

The model receives structured inputs, not a free-form question: the per-game package, the current line and the market-derived probability, plus a data-quality tier — rich, partial or weak — that reflects how much of the expected data actually arrived. The tier is part of the prompt and affects calibration: a breakdown written from a weak package is instructed to hedge and is written more cautiously. The output follows a fixed structure so that every game reads the same way and nothing important is skipped.

Breakdowns are pre-generated for the upcoming slate by scheduled jobs, validated before they are published, and cached so every reader sees the same analysis. A response that fails validation is held for a short period and regenerated rather than shown as the final read. A validated breakdown is written once: the market-derived percentage on the card follows the line with each board refresh, but the text is not rewritten when the line moves, so a late scratch or a goaltender change that lands after generation is not in the read. The AI sports analysis page walks through the pipeline stage by stage and describes each sport’s inputs.

The percentage

The Win Probability

The number next to a game is one number, defined one way, everywhere in the product.

What it is. The percentage shown on the overview, on the matches board, on a game card and inside the opened breakdown is the market-derived win probability described above — the de-vigged share of the favoured side. All four surfaces read the same source, so a game cannot show one number on the board and a different one in its analysis.

What it is not. It is not the language model’s self-reported certainty, it is not a proprietary rating, and it is not an estimate of how likely the analysis is to be right. Language models are poorly calibrated about their own certainty; Sportslyx therefore does not surface that figure at all, and asks the model to write around the market number instead. What the model output does carry is a data-quality tier — rich, partial or weak — which says how complete the inputs were.

Calibration. A calibrated 60% should win about six times in ten over a large enough sample. Because the number comes from an efficient market, it inherits the market’s calibration, which is good, and the market’s errors, which exist. The monthly track record is the calibration check: if reads at a given probability win noticeably more or less often than the probability implies, that is visible in the record rather than hidden behind a rating.

Grading

Prediction Tracking

Every read — the company board’s and your own — is graded by the same rules, from official results, automatically.

Settlement jobs run after games finish and grade each read from the official final score or, for MLB player props, from the official box score. A tracked pick closes 15 minutes after the scheduled start: after that the analysis stays open to read, but a pick can no longer be added or changed, so no record contains a read made with the result partly known. The rules below are applied without exception, and they are the same rules the subscriber-facing product uses.

Grading rules
RuleHow it is applied
Result sourceOfficial final score from the schedule feed, including overtime and shootouts where the market settles the game that way.
Player props (MLB only)Player props are covered for MLB only and graded from the official box score against the posted line. A player who did not play is void — the read is removed, not counted as a loss. Prop lines are half-point lines, so a prop never pushes.
Soccer drawA drawn match is graded as a push: neither a win nor a loss. The read is priced on the 1X2 line, and this rule is stated openly because it inflates ROI relative to a true draw-no-bet price.
VoidsCancelled fights, no-contests, postponed or abandoned games are voided and removed from the record entirely.
GolfGolf previews are published for information only — the company board places no golf reads, so golf does not appear in the graded record on the Pick Tracker.
Cut-offPicks close 15 minutes after the scheduled start. The breakdown remains readable; the pick window does not.
PeriodOne calendar month in UTC, product-wide. The month a read belongs to is the month it was settled in.
Hit rateWins divided by wins plus losses. Pushes and voids are excluded from the denominator.
ROINet result divided by total stake; ROI can be negative. The company record is graded at a flat stake, one per read. Your own tracked picks are graded at the stake you entered, so your personal record shows a net result in money rather than a count of flat stakes.

Your own tracked picks are graded by these same jobs, so a personal record and the company record are directly comparable on the Pick Tracker. The company board publishes one read per game on the slate in the six graded sports — the market favourite on the moneyline, and MLB player props where available — and each is settled after the final result. Golf previews are published for information only and never enter the record. Rare larger-stake reads on the Super Board are tracked separately so a single large position cannot distort the flat-stake record.

Where it breaks

Model Limitations

A methodology is only honest if it names the ways it fails. These are the ones we know about.

  • Delays. Public feeds lag the real world. A late scratch, a goaltender change or a pitching change can arrive after the breakdown was written, and the text is not rewritten when it does: the percentage on the card follows the line, the written read does not. Breakdowns are pre-generated for the upcoming slate and cached, and nothing regenerates them when the market moves. Check the availability panel and the line before you rely on the text.
  • Incomplete feeds. Some competitions carry no shot data, some cards no full fighter records, some golf fields no strokes-gained history for a debutant. The quality tier records this; it does not fix it.
  • Coverage edges. Player props are covered for MLB only. UFC is priced on two markets — the fight moneyline and the round total — so the method of victory is discussed in the prose but never priced. Golf has one market, the outright winner, and its previews are published for information only: golf never enters the graded record.
  • Model error. Language models can misread an input or assert something that is not in the package. Structured inputs, a fixed brief and output validation reduce this; nothing eliminates it.
  • Small samples. A month of reads in one sport can be a few dozen games. Hit rates over samples that small swing widely for reasons that have nothing to do with skill, which is why the record is shown by month and by sport rather than as a single headline number.
  • Market efficiency. Because the probability is derived from the line, it cannot systematically outperform the market it comes from. The value of the product is in the reading, the consistency and the record, not in a claimed edge over the price.
  • The draw-push rule. Soccer reads are priced on the 1X2 line but a draw is graded as a push. A true draw-no-bet price would be shorter, so the published soccer ROI is higher than a two-way market would produce. This is a deliberate, disclosed convention, not an oversight, and it is repeated on every soccer page.

Track record

How Results Are Evaluated

The public record is the product’s only claim about itself. Here is how to read it.

The Pick Tracker publishes, for each calendar month, the number of settled reads, wins, losses and pushes, the hit rate and the ROI at a flat stake — across all sports and for each sport separately. Each sport’s predictions page shows the same figures for that sport, so you can see, for example, whether a strong overall month was carried by one sport’s small sample.

What “good” looks like has to be judged against market efficiency. A board that publishes the market favourite on the moneyline will, by construction, post a high hit rate — favourites win most games — while the ROI at a flat stake hovers near the bookmaker’s margin below zero over a long run, because that is what the margin is. A positive ROI over a month is therefore ordinary variance until it persists; a hit rate is meaningful only next to the average price of the reads behind it. The honest use of the record is to check calibration and consistency over many months, not to celebrate one.

Two conventions make the record easier to read and are worth repeating: the month a read belongs to is the month it settled in, and the figures on every page are the figures the subscriber sees inside the product. There is no marketing version of the numbers.

Frequently Asked Questions

Is the percentage the language model’s own estimate?

No. The percentage shown on a game is a market-derived win probability: the moneyline pair is de-vigged into a fair two-way probability. The language model is told that number and writes around it. What the model’s output does carry is a data-quality tier — rich, partial or weak — which describes how complete the inputs were, not how sure the model is.

How is a prediction graded?

From the official final score, including overtime and shootouts where the market settles that way. MLB player props — the only sport with props — are graded from the box score; a player who did not play is void, not a loss. A soccer draw is a push. Cancelled fights and postponed games are voided and removed. Golf previews are published for information only and are never graded.

What period does the track record use?

One calendar month in UTC, across the whole product. The Pick Tracker, sport hubs and prediction pages all read the same monthly figures, so a number on one page is the same number on every other page.

Why does the draw-as-push rule matter?

Because it removes a losing outcome from the soccer record while keeping the 1X2 price. The published soccer ROI is therefore higher than it would be at a true draw-no-bet price. Sportslyx states this openly rather than presenting the figure as if it were priced on a two-way market.

Can I audit the methodology against the results?

Yes. Every settled read is counted in the public record by month and by sport, and the definitions of hit rate and ROI on this page are the definitions the product uses. If a figure looks wrong, write to support and we will check the settlement.

Check the Record Before You Decide

The 7-day Pro trial lets you test this methodology on live games instead of taking it on trust: every breakdown, the daily board and your own picks graded by exactly the rules on this page.

7-day trial · $0 today · card required · cancel anytime before the trial ends

The definitions on this page are the definitions the product uses. If a published figure looks wrong, see the contact page and we will check the settlement.

Sportslyx produces statistical estimates and analytical breakdowns for information and entertainment. Nothing on this page is betting advice, and no outcome is certain. 18+ only. Please play responsibly — see our Responsible Play page.