Wiki / concepts / wiki

Manufacturer-study skepticism

A vendor-published accuracy study is a starting point for testing, not a result. dc-rainmaker's formulation: "anytime you have a study sponsored by the company that's releasing the study, yeah, who knows about the accuracy of the study" โ€” while still crediting Apple for publishing one at all and naming its comparison set (Garmin Forerunner 970, Samsung Watch, Whoop band, Oura Ring).

Two specific critiques from the apple-watch launch coverage:

  • Methodology asymmetry โ€” Apple's study compares its 5-second HR sampling frequency against competitors that update every 1 second. Not apples-to-apples, and Ray wishes more details had been published.
  • Undisclosed formulas โ€” the same skepticism extends past sensors to derived metrics: health-age's variable weighting isn't published, so users can't tell what moves the number (rob-ter-horst asked Apple directly).

The remedy in both reviewers' hands is the same: reference-device-validation against ECG chest straps, EEG, and manual counters, run independently. Related: claims-vs-testing.

Vindicated a week later: the in-depth review found the study's core flaw โ€” quality-gated-data-omission. Apple graded only the readings its own confidence gate chose to record, excluding the gaps where the watch was implicitly wrong, while competitors were scored on every second. The 5-second-vs-1-second concern from the hands-on turned out to be the visible edge of a deeper asymmetry. Also confirmed: nine pages, 1,300+ participants, no raw data, not peer-reviewed.

Sources: DC Rainmaker Ultra 4 hands-on, Series 12 hands-on, Quantified Scientist analysis

Linked from

DC Rainmaker (Ray)Quality-gated data omission