O que o seu empregado sabe antes de construir qualquer coisa
Qualquer pessoa pode apontar um modelo em um site. Estas são as regras que se aplicam a cada página, por que cada um existe, e qual delas parar uma publicação em vez de avisar sobre isso.
Cada regra abaixo foi paga por um incidente em um site que corremos — um precipício de classificação, uma tradução quebrada, uma página que fez seu próprio código fonte. Não estamos a adivinhar em boas práticas; estamos escrevendo o que já correu errado, assim que não vai errado em seu site também.
Regras que bloqueiam uma publicação: 8 de 12. Uma página bloqueada é entregue ao empregado com a razão, e tenta novamente. Não publica e avisa depois.
1. One page, one subject
blocos publicados
Exactly one <h1>. A descriptive <title> of 15-60 characters.
Porquê?: Title is the strongest per-page signal Google has. Google rewrites titles for the SERP most of the time, but the signal it forms from yours still affects ranking, and a generic or duplicated title reads as a quality problem.
2. A description written for a human
advertências
A meta description of 140-158 characters, unique to the page.
Porquê?: Description does not rank the page, it decides whether anyone clicks it. Too short wastes the snippet; too long truncates mid-sentence.
3. No shared body text, ever
blocos publicados
Every page gets body copy and FAQ answers written for that page. Never a shared template with a noun swapped in.
Porquê?: This is the one that costs months. A site we run shipped ~1700 pages on one generic FAQ body and a site-level quality classifier suppressed the whole domain — pages still indexed, ranking gone, no manual action to appeal. Recovery ran on core-update timescale.
4. Absolute canonical, paired with hreflang
blocos publicados
Every page self-canonicalises with an absolute URL. Any hreflang cluster ships alongside that canonical, also absolute.
Porquê?: Relative hrefs are silently dropped. hreflang without a self-canonical fills the index report with "Duplicate without user-selected canonical" — 56,801 pages on one site we run.
5. Markup may only claim what the page shows
blocos publicados
Structured data mirrors the visible text exactly. Schema goes on the page types that earn a rich result, not on everything.
Porquê?: Schema is not a general ranking input and not required for AI search -- Google says so directly. Markup that claims a price or a question the page never shows is a violation that can earn a manual action.
7. Positional {0} slots in translatable strings
blocos publicados
Any string that will be machine-translated uses {0}, {1} — never {name}.
Porquê?: MT engines translate or transliterate the word inside the braces, or drop a brace entirely. A numeric slot has no word to translate and survives intact. This silently broke about 10% of non-English rows on one site before anyone noticed.
8. Nothing is indexable until it is finished
blocos publicados
Pages ship noindex. Indexing is a deliberate opt-in by the owner, per site.
Porquê?: These sites share a wildcard domain. One spam tenant indexed on that wildcard damages every other tenant on it, and a half-finished page indexed on day one takes months to undo.
9. Build the thing, do not describe the thing
advertências
A page either does something useful or says something specific. A page that only describes a tool is not a page.
Porquê?: Thin wrappers around someone else's API are the clearest signal a quality classifier has. Real utility is what survives a core update.
10. No orphans
advertências
Every indexable page has at least one inbound internal link.
Porquê?: Internal links pass ranking signal and drive crawl. A page reachable only from the sitemap gets crawled a few times and then largely forgotten.
11. Alt text and intrinsic dimensions on every image
advertências
Every <img> has alt text and explicit width/height.
Porquê?: Alt text is an accessibility requirement first and an indexing signal second. Missing dimensions cause layout shift, which is a Core Web Vitals failure on mobile — and mobile is what gets indexed.
12. Retired URLs return 410, verified
blocos publicados
A removed page returns 410 Gone, confirmed by an actual request after deploy.
Porquê?: A 404 lingers as a soft-404 for months; a 410 deindexes cleanly. And per-view dispatch often intercepts the request before the 410-returning view is ever reached, so the check has to be a real HTTP request, not a code read.
Por que isto é um linter e não um guia de estilo
Uma regra que só vive numa instrução é uma sugestão — o modelo segue-a a maior parte do tempo e calmamente não o resto, e você descobrir três meses depois em um Denunciar de tráfego. Então, estes correm como código, depois da página é escrita e antes de ele se tornar vivo. Uma página que falha uma regra de bloqueio é reenviado com a razão e reescrito. O resultado é no recibo de qualquer forma, para que você pode ver quais os controlos corriam.
E nada é indexado até que você diga
Cada site começa sem índice. Não como padrão você tem que descobrir — como regra o linter força, porque esses sites compartilham um domínio e um vizinho mau magoa todos nele. Quando seu site vale a pena encontrar, você liga indexando em configurações e começa a ser rastreado.
Execute o seu primeiro turno livre