SplitAtlas

Product scoring methodology - SplitAtlas

Principle: scores are computed from normalized structured product data. No hand-tuned per-product numbers and no sponsorship influence. Published pages show the methodology link, eligible category bands, excluded categories, and a data-completeness indicator.

Scale

SplitAtlas scores use an absolute 6.5-10.0 display scale. Every catalogued system is certified equipment, so visible scores run from solid to exceptional rather than from bad to good. Category calculations start on a 0-100 normalized scale, then the overall weighted result is mapped into the public 6.5-10.0 score.

Current methodology version: scoring-methodology.v2.

Changelog

  • 2026-08-27 - scoring-methodology.v2: Category scores changed from fixed min/max-style anchors to percentile ranks among scored published peers. For categories where capacity changes the fair market baseline, the comparison is made within standard capacity buckets; otherwise it is made across the scored catalogue. Scores moved because the comparison changed. The public 6.5-10.0 overall display scale and category weights did not change.

Category bands on model pages

Model pages show scored categories as comparison bands, not as bare 0-100 category scores. A band is derived from the stored category percentile score. Because efficiency, cold-weather behavior, noise, and price norms vary by size, those categories normally compare a model against systems in the same standard capacity bucket: <=9k, 12k, 18k, 24k, 36k, or >36k. When a category/capacity bucket has fewer than 30 scored peers, the category falls back to the scored published catalogue and the page labels that fallback.

The band labels are bottom quartile, below average, above average, top quartile, and top 10%. The model page also shows the underlying specification figure used for the category, such as SEER2 for cooling efficiency or HSPF2 for heating efficiency.

Category weights

Cooling efficiency 15%. Heating efficiency 15%. Cold-weather performance 15%. Noise 10%. Warranty 10%. Installation accessibility 10%. Feature set 10%. Price-to-performance 10%. Documentation and parts availability 5%.

Normalization rules per category

Each category first calculates a raw value from sourced product facts, then normalizes that raw value to a percentile rank from 0-100. A score near 50 means the system is around the middle of the chosen peer group; a score near 90 means it is around the 90th percentile. Higher-is-better categories move upward as performance improves; lower-is-better categories, such as noise and price-to-performance, are inverted before percentile ranking.

  • Cooling efficiency: capacity-bucket percentile rank; SEER2 primary, EER2 secondary (70/30 blend).
  • Heating efficiency: capacity-bucket percentile rank; HSPF2 primary; COP at 47F secondary when published.
  • Cold-weather performance: capacity-bucket percentile rank; composite of capacity retention (rated heating at 5F divided by rated at 47F), minimum operating temperature, and published low-ambient rating existence. Missing all three means the category is not scored, never imputed.
  • Noise: capacity-bucket percentile rank; indoor minimum dB(A) (60%), outdoor dB(A) (40%); lower is better.
  • Warranty: catalogue-wide percentile rank; compressor years plus parts years, registered terms; dealer-only warranty restrictions are noted and penalized for DIY-class units.
  • Installation accessibility: catalogue-wide percentile rank; DIY status, precharged line set, included line-set length, 120V availability, and weight class. Professional-install systems are marked not applicable because DIY accessibility is not a valid comparison for them.
  • Feature set: catalogue-wide percentile rank; count of verified supported features, weighted (inverter, Wi-Fi, low-ambient kit, dehumidify, etc.).
  • Price-to-performance: capacity-bucket percentile rank; median current price record divided by capacity and efficiency blend, inverted. Stale prices (older than 60 days) are excluded. Professionally installed quote-only systems with no retail price are marked not applicable instead of being penalized.
  • Documentation and parts: catalogue-wide percentile rank; manuals on file, spec sheet on file, error-code coverage in the model database, parts-channel availability flag.

Applicability

Categories that do not apply to a unit type are excluded from the overall roll-up and disclosed on the model page. Their weight is redistributed across the scored categories. Missing required data is different: it is marked insufficient data and can keep a system from showing an overall score.

Quote-only professional systems are the main price example. If there is no retail equipment price to evaluate, price-to-performance is excluded and the model page says that installed quotes are the right comparison point.

Model page review snapshot

The buyer review snapshot is generated from the same normalized model record used by the score. It is not a separate hands-on review. Fit notes use the model's capacity, zone count, install path, voltage, cold-weather rating, and efficiency rating. Install difficulty weighs the refrigerant connection, DIY evidence, voltage, included line-set information, multi-zone layout, and electrical requirements.

Alternatives are same-zone models from other brands with similar capacity, a higher score, and a current exact retail destination. Availability notes use the most recent verified retailer price when one is current; older prices are labelled reference-only with the date last seen. Online listings are shown as context when a brand publishes a warranty or sales-channel rule that affects buying online.

Source dates come from the imported model record, page review metadata, and retrieval dates stored with source-backed facts. Warranty tiers show standard coverage and the best available registered or qualified coverage when those terms are published.

Confidence gating

  • Each category has required fields. A category with missing required fields is shown as insufficient data and excluded from the overall roll-up.
  • Overall score displays only if at least 5 applicable categories are scored and overall field completeness is at least 80%. A system must never look mediocre because data is missing; it looks unrated.
  • Score history is retained internally so methodology changes can be audited.

Governance

  • Weight or normalization changes require a new methodology version, a changelog entry on the public methodology page, and owner sign-off.
  • Sponsored status is invisible to scoring by construction; sponsorship fields are not included in scoring inputs.
  • Editorial "best for X" picks may consider scores but are labeled editorial judgment and never presented as computed output.