Posted on

Knowledge panels and the selection of an authoritative identity

A knowledge panel does something more ambitious than ranking a page.

It tells you, in effect, this is the thing you searched for.

Google says knowledge panels are automatically generated from its Knowledge Graph using information from multiple web sources, licensed data, and in some cases direct feedback from the entities represented. See About knowledge panels and How Google’s Knowledge Graph works.

That works remarkably well when the identity is unambiguous.

It gets more interesting when two people share a name, a photograph is attached to the wrong biography, or different sources disagree about a fact.

Entity resolution can fail in very human-looking ways

In 2024, The Guardian reported the case of a physicist who discovered that a Google knowledge panel had effectively declared him dead after information about another person with the same name became mixed into the displayed identity. The panel combined a photograph and biographical information in a way that looked authoritative because the interface itself was authoritative-looking.

The underlying problem was not that no information existed. It was that the system had to decide which records belonged to the same entity.

See “Google says I’m a dead physicist”.

This is a useful Dead Internet Theory example because it shows how algorithmic reality can simplify messy source material into one clean card.

Clean does not mean uncontested.

A panel is a synthesis, not a primary source

Google allows people and organizations represented by knowledge panels to claim them and suggest corrections after verification. That process is itself evidence that the panel should be understood as maintained synthesis rather than an untouchable database record.

If a panel matters to your research, inspect the sources behind the claim when possible. Search the person’s official site, institutional page, original publication, corporate record, or another primary source appropriate to the fact.

The same name can describe different people. The same organization can change names. A band, product, company, and person can share overlapping terms. Even a correct entity can contain one incorrect attribute.

Knowledge panels are useful because they reduce all that friction.

But reduction has a cost: disagreement, ambiguity, and provenance can disappear from the first glance.

The panel is Google’s best current model of the entity.

It is not the entity itself.

Posted on

Authority signals and the advantage of established publishers

A famous publisher can be wrong and an obscure hobbyist can be the world’s best source on one strange little subject.

Search engines still need a way to rank both of them.

That forces search systems to use proxies for relevance, usefulness, and reliability. Some are specific to the page. Others emerge from the wider web around it: links, references, reputation, topical history, and signals that other people treat the source as worth consulting.

Google’s Reliable results on Search documentation gives a simple example. If other prominent sites link to or refer to a piece of content, that can suggest that the source is reliable. Google’s ranking systems guide also describes link-analysis systems, including the modern descendants of PageRank, along with systems intended to surface more reliable and authoritative information.

None of that translates into a simple “big site bonus.”

It does create an accumulation problem.

Established publishers have history to spend

A long-running publication may have millions of inbound links, recognizable authors, years of citations, structured archives, stable URLs, and a large audience that continuously creates new references to its work.

A new independent site starts with almost none of that.

Even if both publish equally useful pages today, they arrive at the ranking system with very different histories.

That advantage can be deserved. Institutions that repeatedly publish accurate material should not be forced to prove themselves from zero on every query. Strong reputation signals help suppress spam, impersonation, disposable content farms, and pages created yesterday to exploit today’s search demand.

But proxies have edges.

A retired engineer may maintain the best documentation for an obsolete machine. A collector may have photographed a component that no museum has cataloged. A regional historian may know a subject ignored by national publishers. These sources can be exceptionally valuable while possessing very little conventional web authority.

Authority is not the same as expertise on every page

Google itself notes that ranking operates substantially at the page level and that good site-wide signals do not guarantee every page will rank highly.

That distinction matters.

The existence of authority systems does not prove that a large publisher automatically beats a specialist. Nor does a small site’s poor ranking prove that its content was judged incorrect. Ranking combines many signals, and the exact weighting is not public.

The useful observation is narrower: established publishers have more opportunities to accumulate the signals search systems can observe.

That affects the internet people experience.

If discovery repeatedly favors sources with large existing reputations, the web can appear more institutionally concentrated than the underlying supply of knowledge really is.

The specialist page may still exist.

It just arrives at the race without forty thousand people already pointing at it.