Methodology / versioned by category

The evidence stays visible.

Every ranked product passes the same identity and traceability gates. The wider exact-identity index stops before scoring. The score itself changes by category because cooling, moisture removal, portable power, cooking, floor care, autonomous cleaning and audio are not interchangeable decisions.

01

Resolve the exact UK product

Model, SKU, variant and market are resolved before comparison. GTIN is retained where the source exposes it. A retailer title is not allowed to merge similar products. A checksum-valid GTIN community record can enter the wider identity index, but it does not establish manufacturer model, current market, offer or ranking eligibility.

02

Keep evidence classes separate

Manufacturer claims, retailer observations, controlled tests, regulatory records and owner reports remain distinct. A published specification is never relabelled as an Agent Review Data measurement. Images remain source-hosted retailer or manufacturer assets with a linked source page; a future affiliate feed would need the same traceability.

03

Gate first, score second

Exact identity, traceable specification, complete score inputs, a timestamped offer observation and either resolved conflicts or an explicit conservative conflict policy are required. Method-tolerated conflicts remain visible and block purchase advice. Identity-index records have null scores and ranks. A numeric score cannot override a failed gate.

04

Rank only within a category

The specification layer contains 20 exact products per category and receives ranks 1–20. A score of 80 for a power bank has no relationship to a score of 80 for a dehumidifier. A separate review-led comparison must state its own bounded cohort and cannot reuse the specification score as quality.

05

Separate product quality from buyability

Stock is an exact-product, source-specific observation with its own checked time. Page render time is never used as evidence freshness, and generic “Add to cart” strings are not accepted as stock proof.

06

Deduplicate owner feedback and third-party evidence

Native ratings and counts remain source-specific. Syndicated mirrors share one pool; small or commercially hosted pools are capped and pulled toward neutral rather than presented as false internet-wide consensus. Independent tests, owner pools and Reddit discussions are labelled by evidence class and kept claim-level; they do not become a universal score.

07

Make commercial influence impossible

Offer redirects are affiliate-ready and attributable. Affiliate state, merchant commission and conversion performance are not inputs to evidence, eligibility, factors or rank.

08

Refresh atomically

Source collection validates 20 complete products per category, recomputes ranks and replaces the snapshot only after all publication gates pass. A collection failure cannot silently shrink the live ranking. The wider identity collector also writes atomically, records exclusions and leaves the last-good identity snapshot in place unless all 3,000 records pass.

npm run refresh:catalogue
npm run refresh:identities
09

Build a useful product evidence report

Every ranked product receives the same report structure, but not the same performance formula. The report declares what matters for its category, exposes retained source coverage, rank sensitivity and blockers, and says when exact-model hands-on synthesis was not collected. Where review evidence passes the gate, an LLM may synthesise heterogeneous reviews using topic-specific trust and direct paragraph locators. Missing topics remain unknown, and legacy synthesis is not treated as category-profile coverage. Synthesis cannot change the specification score, specification rank, offer or purchase-advice gate. It can support a separately versioned quality comparison only after multi-product judgement, repeatability checks and current external sense checking.

flexible-review-led-category-v1.1

One review approach, adapted to what matters

Every category now leads with exact-model review and test evidence, or an explicit decision that the evidence is not yet strong enough to order products. The method does not force unlike sources into one formula.

01 / exact identity

Attach evidence only to the tested model

Market, model code and variant must match. Family reviews can help candidate discovery, but cannot silently transfer sound, cooling, cooking or cleaning performance to nearby hardware.

02 / flexible extraction

Use each source for what it really tested

A structured LLM pass extracts observations, conditions, strengths, failures, disagreements and unknowns. A review does not need to follow our rubric, but its prose cannot support a claim it did not investigate.

03 / category judgement

Prioritise real use, not specification volume

Sound and fit lead audio; delivered energy and sustained output lead power banks; cooking results lead air fryers; cleaning, air treatment, cooling and moisture removal use their own measured conditions and failure modes.

04 / disagreement

Preserve conflicting outcomes and test conditions

Publisher stars and list positions are contextual evidence. Their written reasons, room or load conditions, publisher ownership and exact measurements are retained. Conflicts are explained rather than averaged away.

05 / publication gate

Use one order, specialist labels or no order

Ordered and specialist results require corroboration across at least two independent publisher groups. Close calls receive one editorial position while retaining confidence, disagreement and use-case labels. When the category cannot support a comparison, the order is withheld.

06 / commercial join

Add the lowest checked price after quality

The quality result is frozen before offers join. Buy actions select the lowest price among fresh tracked exact-product offers. Price, stock, affiliate state and clicks cannot alter the result.

insufficient exact model review evidence

Portable air conditioners

The tracked catalogue does not yet contain enough independently tested exact models for a defensible quality order. One exact Meaco model has useful hands-on evidence, but a single comparable product is not a ranking.

What matters: Observed cooling in a stated room size and starting condition, Noise and sleep suitability measured or experienced in use, Installation, exhaust hose and window sealing burden, Power draw, controls and sustained-use behaviour, Weight, movement and storage practicality.

Inspect the machine record ↗
insufficient exact model review evidence

Dehumidifiers

Arete Two 20L has useful exact-model hands-on evidence, but the currently tracked comparison set is not broad or clean enough to support an ordinal quality order. Arete One 20L remains outside this first cohort while stored tank-capacity and current colour or variant observations are reconciled.

What matters: Observed extraction with temperature and humidity context, Noise in living, bedroom and overnight use, Laundry drying performance and airflow, Energy use at useful extraction rates, Tank, drainage, controls and filter maintenance.

Inspect the machine record ↗
published bounded comparison

Power banks

Four exact Anker models have enough independent measured and hands-on evidence for a bounded order. Delivered energy, sustained output, recharge behaviour, thermals and practical port or cable use lead the judgement; advertised capacity and headline wattage do not.

What matters: Delivered usable energy rather than advertised cell capacity, Sustained output and late-discharge throttling, Recharge time and pass-through behaviour, Thermals and protection under demanding loads, Port, cable, display and carrying practicality.

Inspect the machine record ↗
published bounded comparison

Air fryers

Four exact Ninja dual-zone models support one simple all-round order, while the narrow Double Stack is fifth with a clear space-saving specialist label. Cooking results, evenness, controls, cleaning and basket format lead the judgement.

What matters: Cooking evenness, crisping and repeatability across foods, Speed without excessive shaking or intervention, Basket flexibility and usable capacity, Controls, synchronisation and probe behaviour, Cleaning, heat leakage and countertop fit.

Inspect the machine record ↗
published bounded comparison

Cordless vacuums

Two Shark models take first and second in the all-round order, while Gtech AirRAM 3 is third with a clear floor-only specialist label. Pickup, handling, hair management, runtime under real settings and maintenance lead the judgement.

What matters: Observed debris and pet-hair pickup by floor type, Handling, steering, reach and effort, Useful runtime at effective power settings, Hair wrap, bin emptying and filter maintenance, Tool coverage for stairs, furniture and edges.

Inspect the machine record ↗
published bounded comparison

Robot vacuums

Roborock Saros 20, eufy E25 and eufy E28 take the first three places. Qrevo Curv 2 Pro and eufy S1 Pro are fourth and fifth with clear carpet and hard-floor mopping specialist labels.

What matters: Observed hard-floor, carpet and pet-hair cleaning, Mopping effectiveness and floor suitability, Navigation, obstacle avoidance and threshold handling, Dock reliability and routine intervention burden, Mapping, app control and repeat-run consistency.

Inspect the machine record ↗
published bounded comparison

Air purifiers

Five exact large-room purifiers have enough controlled and hands-on evidence for a bounded throughput-led comparison. Measured clean-air delivery at useful noise levels, room-clearing behaviour, filters and ownership experience lead the judgement.

What matters: Measured or certified clean-air delivery with unit and test condition, Useful throughput at tolerable noise, Observed particle clearing and room-size suitability, Filter availability, sealing and maintenance, Sensors, controls, app behaviour and sleep use.

Inspect the machine record ↗
wireless-earbuds-review-led-v1.1

Review-led wireless earbud quality comparison

This is the primary earbud result. It does not average star ratings or force heterogeneous reviews into a universal score.

01 / evidence gate

Resolve the exact model and retain hands-on use

Each included model has three exact-product hands-on reviews from at least two publisher groups. Manufacturer specifications can clarify identity and features but cannot establish sound, fit, ANC or calls.

02 / flexible synthesis

Interpret what each review can genuinely support

A holistic LLM pass reads retained review text and extracts source-grounded strengths, weaknesses, conflicts, use cases and unknowns. A source is used for the topics it actually tested; missing alignment to a fixed rubric is not invented.

03 / comparative passes

Judge the whole cohort more than once

Product syntheses are compared in three rotated input orders. Stable conclusions can survive. Adjacent swaps inform confidence and visible caveats, while the published result uses one clear order and retains use-case labels.

04 / current sense check

Read rankings and their descriptions

Current SoundGuys, RTINGS, Future plc titles, TechGearLab and UK-oriented roundups are checked after the internal pass. Publisher groups are deduplicated. Their written reasoning is used to investigate disagreement, not reduced to an average rank.

05 / divergence review

Ask what the comparison missed

Material differences trigger checks for publication age, missing products, price class, fit, ecosystem restrictions, sound, ANC, calls, reliability and test conditions. The Sennheiser and AirPods corrections in version 1.0 came from this step.

06 / commercial join

Freeze quality before adding price

Current UK offers are attached only after the quality order is frozen. Each action selects the lowest price among fresh tracked exact-product offers and states the retailer count. Price, stock, affiliate state and clicks cannot change quality position.

Current scope: seven reviewed premium or ecosystem models in the quality order and one eighth-place travel specialist. 10 material candidate gaps are public. The candidate-universe rule flags current UK sealed-ANC picks, notable mentions and material comparators from the six benchmark sources. This supports a useful bounded comparison, not a whole-market winner claim. Inspect the complete machine record ↗

The 24 source count is a document count, not 24 independent tests. Related titles share one publisher group. Schema v1 stores the access snapshot but does not normalise every source publication date or expose private-corpus paragraph locators, so those remain manual review limits.

Category contracts

Eight methods, one provenance standard

Weights below describe the current specification-led MVP. Missing controlled tests remain visible and penalised.

portable-ac-v1.0

Portable air conditioners

Specification-led comparison of exact GB models. Published values are not treated as laboratory-equivalent measurements.

Cooling per pound
28%
Published noise
24%
Claimed efficiency
22%
Practical features
14%
Evidence coverage
12%
Inspect all 20 results ↗
dehumidifier-v1.0

Dehumidifiers

Specification-led comparison. Extraction claims can use different temperature and humidity conditions, so the score is directional.

Extraction per pound
25%
Published noise
22%
Nameplate efficiency proxy
22%
Moisture-control features
19%
Evidence coverage
12%
Inspect all 20 results ↗
power-bank-v1.0

Power banks

Specification-led comparison of exact UK offers. Nameplate capacity and peak output are not delivered-energy tests.

Capacity and output per pound
28%
Published maximum output
24%
Nameplate capacity
18%
Convenience features
18%
Identity and evidence
12%
Inspect all 20 results ↗
air-fryer-v1.0

Air fryers

Specification-led comparison of exact UK models. Capacity, power and modes are published specifications, not cooking-lab measurements.

Value at checked price
28%
Published capacity
20%
Cooking flexibility
24%
Convenience features
16%
Evidence coverage
12%
Inspect all 20 results ↗
cordless-vacuum-v1.1

Cordless vacuums

Specification-led comparison of exact UK models. Run time is a published maximum and is not a like-for-like measured result.

Value at checked price
25%
Published run time
20%
Published weight
20%
Cleaning features
23%
Evidence coverage
12%
Inspect all 20 results ↗
robot-vacuum-v1.1

Robot vacuums

Specification-led comparison of exact UK models. Published suction and navigation claims are not treated as a substitute for a controlled home test.

Value at checked price
25%
Published suction
20%
Navigation and obstacle handling
23%
Dock and mopping automation
20%
Evidence coverage
12%
Inspect all 20 results ↗
wireless-earbuds-v1.1

Wireless earbuds

Specification-led comparison of exact UK models. Battery life and protection ratings depend on conditions and are not independently measured here.

Value at checked price
25%
Battery life
20%
Listening features
25%
Protection and charging
18%
Evidence coverage
12%
Inspect all 20 results ↗
air-purifier-v1.0

Air purifiers

CADR-led specification comparison of exact UK manufacturer offers. Raw pollutant bases remain visible and mixed-basis values are not presented as laboratory-equivalent.

Value at checked price
55%
Published CADR throughput
45%
Inspect all 20 results ↗
Product report contracts

One report shape, category-specific evidence needs

These topics guide source collection and flexible review synthesis. They are coverage requirements, not fixed score weights.

product-report-category-profile-v1.0

Portable air conditioners

Primary question: Delivered cooling. Observed cooling under disclosed room, heat-load, ambient-temperature and run-time conditions.

Delivered cooling
required topic
Noise in real use
required topic
Window and exhaust fit
required topic
Condensate and drainage
required topic
Power draw and efficiency
required topic
Portability
required topic
Sleep use
required topic

Hands-on gate: Use two independent exact-model hands-on sources, or one transparent controlled test that discloses room, installation and measurement conditions.

Inspect product reports ↗
product-report-category-profile-v1.0

Dehumidifiers

Primary question: Extraction performance. Observed water extraction with temperature, starting humidity, duration and room conditions.

Extraction performance
required topic
Humidistat and cycling
required topic
Laundry drying
required topic
Noise
required topic
Power draw
required topic
Tank and drainage
required topic
Cold-room performance
required topic

Hands-on gate: Require at least one exact-model extraction observation with temperature, humidity and duration disclosed. Broader verdicts should use independent corroboration.

Inspect product reports ↗
product-report-category-profile-v1.0

Power banks

Primary question: Delivered energy and usable capacity. Controlled delivered-energy testing with voltage, protocol and termination conditions.

Delivered energy and usable capacity
required topic
Sustained output and protocols
required topic
Recharge time
required topic
Thermal behaviour
required topic
Pass-through behaviour
required topic
Portability
required topic
Device compatibility
required topic

Hands-on gate: Resolve the exact SKU and cell configuration. Performance claims require controlled delivered-energy or output measurements; otherwise report nameplate values only.

Inspect product reports ↗
product-report-category-profile-v1.0

Air fryers

Primary question: Cooking evenness. Documented cook tests showing browning and doneness across the usable basket.

Cooking evenness
required topic
Cooking speed
required topic
Usable capacity
required topic
Temperature accuracy
required topic
Smoke and noise
required topic
Cleaning
required topic
Controls and safety
required topic

Hands-on gate: Require the exact variant and at least one documented cook test with food, load, mode, temperature and elapsed time.

Inspect product reports ↗
product-report-category-profile-v1.0

Cordless vacuums

Primary question: Pickup performance. Controlled or repeatable pickup tests by floor type, debris and cleaning head.

Pickup performance
required topic
Runtime by mode and head
required topic
Hair handling
required topic
Manoeuvrability
required topic
Bin and emptying
required topic
Filtration
required topic
Noise and maintenance
required topic

Hands-on gate: Resolve the exact bundle and cleaning head. A broad performance verdict needs pickup and runtime evidence with floor, debris, mode and head retained.

Inspect product reports ↗
product-report-category-profile-v1.0

Robot vacuums

Primary question: Pickup and coverage. Repeatable cleaning tests by floor, debris, pass count and mapped area.

Pickup and coverage
required topic
Obstacle avoidance and navigation
required topic
Mopping
required topic
Dock automation
required topic
Thresholds and edges
required topic
App and privacy
required topic
Maintenance and reliability
required topic

Hands-on gate: Resolve the exact robot, dock and region. Require transparent cleaning and navigation tests; mopping and dock conclusions need their own evidence.

Inspect product reports ↗
product-report-category-profile-v1.0

Wireless earbuds

Primary question: Sound and tuning. Hands-on listening or controlled measurements covering tonal balance, detail, distortion, dynamics and the effect of EQ or codec settings.

Sound and tuning
required topic
Comfort and fit
required topic
Noise control and awareness
required topic
Microphone quality
required topic
Battery and charging
required topic
Controls, connectivity and app
required topic
Durability and long-term use
required topic

Hands-on gate: Resolve the exact model. Record firmware as identified only when disclosed, otherwise not_disclosed; firmware-dependent claims remain unavailable. At least one source must provide method-transparent sound evidence, and the report must contain a dedicated sound section. Without that, no overall audio-quality verdict is supported.

Inspect product reports ↗
product-report-category-profile-v1.0

Air purifiers

Primary question: CADR and pollutant basis. Attributable smoke, dust or pollen CADR with units, test standard and filter configuration.

CADR and pollutant basis
required topic
Pollutant removal performance
required topic
Room size and air changes
required topic
Noise
required topic
Energy use
required topic
Filter cost and life
required topic
Sensors and controls
required topic

Hands-on gate: Resolve the exact model and filter. Pollutant-removal conclusions require an attributable controlled measurement with pollutant basis and units.

Inspect product reports ↗
Current limitation
“Ranked” means comparable published evidence. It is not a laboratory verdict.

The catalogue is useful for shortlisting, inspecting claims and checking offers. It does not claim identical test conditions where the sources do not provide them.

Inspect the comparison tables