Manufacturer-study skepticism
A vendor-published accuracy study is a starting point for testing, not a result. dc-rainmaker's formulation: "anytime you have a study sponsored by the company that's releasing the study, yeah, who knows about the accuracy of the study" โ while still crediting Apple for publishing one at all and naming its comparison set (Garmin Forerunner 970, Samsung Watch, Whoop band, Oura Ring).
Two specific critiques from the apple-watch launch coverage:
- Methodology asymmetry โ Apple's study compares its 5-second HR sampling frequency against competitors that update every 1 second. Not apples-to-apples, and Ray wishes more details had been published.
- Undisclosed formulas โ the same skepticism extends past sensors to derived metrics: health-age's variable weighting isn't published, so users can't tell what moves the number (rob-ter-horst asked Apple directly).
The remedy in both reviewers' hands is the same: reference-device-validation against ECG chest straps, EEG, and manual counters, run independently. Related: claims-vs-testing.
Vindicated a week later: the in-depth review found the study's core flaw โ quality-gated-data-omission. Apple graded only the readings its own confidence gate chose to record, excluding the gaps where the watch was implicitly wrong, while competitors were scored on every second. The 5-second-vs-1-second concern from the hands-on turned out to be the visible edge of a deeper asymmetry. Also confirmed: nine pages, 1,300+ participants, no raw data, not peer-reviewed.
Sources: DC Rainmaker Ultra 4 hands-on, Series 12 hands-on, Quantified Scientist analysis