Training Example: Wikipedia – Review the Data, Give Your Score & Compare to the Real AI Evaluation

Industry Context — Common BS Fingerprints in Media, News & Publishing
Generic Claims: trusted news source, unbiased reporting, the truth, delivered, journalism that matters…
Red Flags: no named editorial staff, sponsored content without clear labelling, no corrections or complaints policy, ownership and funding not disclosed…
Semantic Drift Patterns: claims editorial independence but content is sponsored, claims fact-checked but no corrections policy visible, homepage says investigative but content is aggregated wire stories, claims community voice but no local reporting staff…
Proof Expectations: named journalists and editorial staff, published editorial standards and ethics code, corrections and complaints policy, ownership and funding transparency…

Wikipedia

(https://wikipedia.org) 📸 Data Snapshot: June 20, 2026

Analyze the raw signals below. How would a machine score this business’s credibility?

Here are the exact signals captured from up to six pages of the site — the same raw inputs the evaluation engine analyzed. They are grouped by signal type so you can weigh each the way the machine does.

🏗️ Semantic Structure — heading hierarchy & page identity (Info Density · Commodity Fingerprint)
HOMEPAGE Wikipedia (https://wikipedia.org)
Title

Wikipedia

Meta

Wikipedia is a free online encyclopedia, created and edited by volunteers around the world and hosted by the Wikimedia Foundation.

H1 Wikipedia The Free Encyclopedia
H2 1,000,000+ articles
H2 100,000+ articles
H2 10,000+ articles
H2 1,000+ articles
H2 100+ articles
H3 We owe you an explanation.
📝 The Narrative — clean text per page (Info Density · Semantic Coherence)
HOMEPAGE (https://wikipedia.org) Wikipedia
[H1]
Wikipedia
The Free Encyclopedia
[H3] We owe you an explanation.
You deserve an explanation, so please don't skip this 1-minute read. Our fundraiser won't last long, and we need some help to reach our goal. Less than 2% of our readers donate, but if everyone who saw this message gave $2.75, we'd hit our goal in a few hours. The rare few who donate do so because Wikipedia provides them with useful knowledge. If that sounds like you, please donate $2.75. Any contribution you make today helps.
We ask you, sincerely: don't skip this. Be one of the rare readers who gives.
There are no small contributions: every edit counts, every donation counts. Thank you.
Donate now
I already donated
? Thank you for donating recently! ?
Your support means the world to us. We'll hide banners in this browser for the rest of our campaign.
Close
838 chars
🛡️ Trust Signals — reviews, proof links, trust-theatre flag (Trust & Proof)
0Review mentions (all pages)
0External proof links (all pages)
PageReviewsProof links
/ (home) 0 0
🔗 Identity & Technical Layer — schema JSON-LD: identity chains, entity gaps (Identity & Authority)
Homepage — no schema detected (entity gap)

Your Diagnosis

Before revealing the machine’s verdict, predict the BS score for each signal. Higher = more BS (more fluff, less verifiable substance). Drag each slider, then submit to compare your judgment against the engine.

Information Density 0 / 30
Read the Narrative & headings: do hard facts (prices, dates, numbers) outweigh fluff power-words?
Semantic Coherence 0 / 20
Compare the homepage promise against the sub-page reality. Do they hold the same line?
Trust & Proof 0 / 20
Weigh review mentions against actual external proof links. Claims without verification = theatre.
Commodity Fingerprint 0 / 15
Check headings & narrative against the industry clichés in the setup above.
Identity & Authority 0 / 15
Inspect the schema: is there real Organization/Person identity with sameAs links, or gaps?
Your predicted BS score 0 / 100
💡 Stuck? Reveal the heuristic lens — how the deterministic page-auditor reads each signal (no AI, pure pattern rules)

These are the structural rules a local, deterministic auditor applies — the same lens you can use to judge each signal. They describe what to look for, not this company’s result.

Information Density

Classify each sentence as substantive or hollow. Grounding markers — numbers, currencies, dates, technical units, named entities — outweigh marketing adjectives. When fluff sits right next to hard evidence, the fluff is forgiven.

Semantic Alignment

Pull the main entities out of the H1, then check whether they actually recur through the body. A page that announces one thing and then talks about another drifts. Headings with no real sentences underneath read as pseudo-substance.

Trust & Proof

Count trust words (review, testimonial, rating, verified) against real outbound proof links (Google, Trustpilot, Clutch, G2, Yelp). Lots of trust language with zero verification links is trust theatre. Unlinked logo galleries count against it.

Commodity Fingerprint

Look at how much sentence length varies. Natural writing varies its rhythm; templated or mass-produced copy is statistically uniform. Very low variation reads as commodity content — unless unique named entities break the pattern.

Identity & Authority

Inspect the JSON-LD. Is there an Organization or Person schema, and does it carry sameAs links to real external profiles (LinkedIn, socials)? Missing schema or no identity declaration signals an anonymous entity.

Want to apply this lens yourself? The free BS Indicator Chrome extension runs these heuristic checks live on any page. Bear in mind it is a single-page, deterministic tool — it relies only on pattern rules for the page in front of it and does not perform the cross-page semantic correlation this audit uses, so its readout is a starting lens, not the full verdict.

B
BS Level
Media, News & Publishing
34.7 Avg BS

Based on 831 businesses audited.

BS Detector

Media, News & Publishing BS: Wikipedia (wikipedia.org)

https://wikipedia.org 📍 Industry: Media, News & Publishing
13 BS / 100

Wikipedia is a masterclass in anti-BS communication, using quantitative scale and transparent financial appeals instead of industry jargon. It avoids almost every red flag in the Media category by focusing on its utilitarian function. The only measurable BS factors are technical omissions in schema and the inherent anonymity of its volunteer-led authority model.

Info Density Power-words vs. Substance ratio.
3
10% BS
Semantic Coherence Homepage promise vs. Sub-page reality.
1
5% BS
Trust & Proof Verifiable evidence vs. Trust Theatre.
3
15% BS
Commodity Fingerprint Detection of industry clichés/templates.
0
0% BS
Identity & Authority Expert verifiability & Schema depth.
6
40% BS

Implement Organization schema with sameAs links to the Wikimedia Foundation to anchor digital identity. Add a specific link to published Editorial Standards or Neutral Point of View policies to satisfy industry proof expectations. Clearly name the governing entity in the H3 or footer to bridge the identity gap. Convert quantitative article counts in H2 into links to the respective language category pages to provide immediate proof paths.

The site fits the Media, News & Publishing category as a digital-first publishing entity. The content confirms its status as a free encyclopedia and information aggregator rather than a traditional newsroom, focusing on volunteer-driven content distribution.

“The score of 13 is driven primarily by the 'Identity and Authority' pillar (6 points) due to missing schema_json and named leadership, and 'Trust and Proof' (3 points) due to the absence of external outbound proof links in the snippet. Information Density (3 points) reflects the high use of specific numbers over fluff. Commodity Fingerprint (0 points) confirms the site's unique, non-generic positioning.”

Verified Analysis Date: June 20, 2026 © 1EuroSEO Independent Evaluator — Non-Sponsored Result
Brand AI Reputation