Customer feedback 50 %
Reviews and ratings from real buyers – the most direct indicator of actual product quality in everyday use.
The Trufl Score is a holistic quality rating for products on a scale from 40 to 100. It combines four independent signals – customer feedback (50 %), brand reputation (20 %), bestseller rank (20 %), and technical features (10 %) – into a comparable rating within the same product category.
The Trufl Score answers one central question: "How good is this product compared with other products in the same category?"
Unlike simple star ratings, which can be misleading across different product categories, the Trufl Score ensures a fair comparison: a highly rated puzzle is not compared with a security camera, but only with other products in its category.
The score is built from four independent rating pillars, each capturing a different perspective on product quality. Combining multiple signals balances distortions from individual data sources and produces a robust overall rating.
Each pillar captures a distinct aspect of product quality. The weighting reflects the reliability and relevance of each data source.
Reviews and ratings from real buyers – the most direct indicator of actual product quality in everyday use.
The brand’s quality history within the category, plus public perception on social media.
Market validation through real purchase decisions – the “wisdom of the crowd” as an objective reality check.
Objective analysis of category-specific product attributes such as material quality, feature set, and build quality.
Customer feedback is the core of the Trufl Score and has the strongest influence on the overall result. It captures real buyer experience and is composed of multiple signals for a nuanced picture.
Instead of relying only on the average star rating, the Trufl Score analyses several dimensions of customer feedback:
Ratings from platforms other than Amazon (e.g. Google Shopping) are not taken at face value. For each category we calculate a platform offset from overlapping products and correct for systematic rating differences. That keeps comparisons fair – regardless of where the rating originated.
Verified professional test results can further strengthen the customer-feedback pillar: products that score especially well in independent expert reviews receive a moderate bonus of up to 4.5 % within this pillar.
Different product categories naturally have very different rating levels. The Trufl Score corrects those distortions for customer feedback, brand reputation, and bestseller rank so only products within the same category are compared fairly. Technical feature analysis is the exception: it is deliberately not normalised relative to the category, but scored on absolute capability.
Brand reputation captures a manufacturer’s quality history within a given product category. It has two components:
How have all of a brand’s products performed within the category? Brands that consistently deliver high quality receive a higher score. Products with more reviews are weighted more heavily – a single niche product influences the brand score less than a bestseller with thousands of reviews.
We also analyse public perception of the brand on social platforms such as Reddit and Twitter. We capture not only tone (positive vs. negative), but also the quality and relevance of discussions. Product-related criticism counts more than general complaints about shipping or a website.
Leading brands in a category receive an additional bonus based on their overall performance within that product group.
Bestseller rank reflects the “wisdom of the crowd”: which products do consumers actually buy most often? The data foundation is the Amazon Sales Rank.
Amazon Sales Rank is collected for every product in the category
Misclassified products and accessories are filtered out of the bestseller list
Rank is converted into a percentile relative to the cleaned category size
The result feeds into the overall score as market validation
Not every product has an Amazon Sales Rank. For products sold mainly on other platforms, we calculate a bestseller proxy from rating activity on that platform. The value is adjusted by category and weighted conservatively so market relevance outside Amazon is still represented fairly.
Amazon bestseller lists often include misclassified products (e.g. computer games in a baby-monitor category). Those outliers are removed systematically so only relevant products influence the rating.
This signal provides an objective reality check: products that have proven themselves in the market receive corresponding recognition in the score. A category winner among 10,000 competitors receives the highest possible value.
Every product category has specific quality attributes that can be assessed objectively. Relevant properties are extracted with AI from technical datasheets and product descriptions, then scored:
For each product category, relevant product attributes are identified automatically from technical data and descriptions and stored in the database.
The identified attributes are ranked and weighted by their relevance to customers and product quality.
Each product is scored on its attributes to identify the product with the strongest features in the category.
Unlike the other pillars, technical features are not scored via relative category normalisation. Actual capability is used instead – so simple products are not artificially pushed up or down.
If ratings are not yet available for individual attributes, a neutral baseline is used. Products are not penalised simply because data is missing for some aspects.
To keep the Trufl Score meaningful, strict quality requirements apply to the underlying data:
Products with few reviews could receive extreme scores by chance alone. The Trufl Score uses statistical smoothing that gently pulls new-product scores toward the category average. As more data accumulates, the score is driven increasingly by real feedback.
A product must have at least one real signal (e.g. customer reviews or bestseller rank) to receive a Trufl Score. Products with no data basis are excluded from scoring – preventing empty records from distorting the overall distribution.
The Trufl Score is calculated on a scale from 40 to 100 points. The weighted results of all four pillars are combined for each product and normalised within the category. We then apply an S-curve (sigmoid) transform so similar top products are easier to tell apart.
The higher a product’s Trufl Score, the better it performs versus other products in the same category. The rating is always relative to the category – a score of 85 for headphones is not directly comparable with a score of 85 for suitcases.
Analysis of product ratings, product-focused review sentiment, price perception, and malfunction reports – including platform correction and an optional expert-review bonus.
Review-based brand quality history within the category, plus social-sentiment analysis from relevant discussions on platforms such as Reddit and Twitter.
Market validation via Amazon Sales Rank and a conservative bestseller proxy for products on other platforms.
AI-assisted extraction and objective scoring of category-specific product attributes based on absolute capability.
Studies and ratings are conducted independently and published editorially. The Trufl Score is calculated fully from data, without influence from manufacturers or retailers. Positively rated providers may, after the study is complete, obtain a licence to use test seals for advertising. Licensing has no influence on the methodology or the results.