O que o seu empregado sabe antes de construir qualquer coisa

Qualquer pessoa pode apontar um modelo em um site. Estas são as regras que se aplicam a cada página, por que cada um existe, e qual delas parar uma publicação em vez de avisar sobre isso.

Cada regra abaixo foi paga por um incidente em um site que corremos — um precipício de classificação, uma tradução quebrada, uma página que fez seu próprio código fonte. Não estamos a adivinhar em boas práticas; estamos escrevendo o que já correu errado, assim que não vai errado em seu site também.

Regras que bloqueiam uma publicação: 8 de 12. Uma página bloqueada é entregue ao empregado com a razão, e tenta novamente. Não publica e avisa depois.

1. One page, one subject

blocos publicados

Exactly one <h1>. A descriptive <title> of 15-60 characters.

Porquê?: Title is the strongest per-page signal Google has. Google rewrites titles for the SERP most of the time, but the signal it forms from yours still affects ranking, and a generic or duplicated title reads as a quality problem.

2. A description written for a human

advertências

A meta description of 140-158 characters, unique to the page.

Porquê?: Description does not rank the page, it decides whether anyone clicks it. Too short wastes the snippet; too long truncates mid-sentence.

3. No shared body text, ever

blocos publicados

Every page gets body copy and FAQ answers written for that page. Never a shared template with a noun swapped in.

Porquê?: This is the one that costs months. A site we run shipped ~1700 pages on one generic FAQ body and a site-level quality classifier suppressed the whole domain — pages still indexed, ranking gone, no manual action to appeal. Recovery ran on core-update timescale.

4. Absolute canonical, paired with hreflang

blocos publicados

Every page self-canonicalises with an absolute URL. Any hreflang cluster ships alongside that canonical, also absolute.

Porquê?: Relative hrefs are silently dropped. hreflang without a self-canonical fills the index report with "Duplicate without user-selected canonical" — 56,801 pages on one site we run.

5. Markup may only claim what the page shows

blocos publicados

Structured data mirrors the visible text exactly. Schema goes on the page types that earn a rich result, not on everything.

Porquê?: Schema is not a general ranking input and not required for AI search -- Google says so directly. Markup that claims a price or a question the page never shows is a violation that can earn a manual action.

6. Django comments are {% comment %}

blocos publicados

Never {# #}. Not even on one line.

Porquê?: Django's template lexer regex has no DOTALL flag, so {# #} only closes on its own line. A multi-line one renders as visible page text, and if it contains a {% %} substring the lexer treats it as a real tag and every render crashes. Both have happened in production.

7. Positional {0} slots in translatable strings

blocos publicados

Any string that will be machine-translated uses {0}, {1} — never {name}.

Porquê?: MT engines translate or transliterate the word inside the braces, or drop a brace entirely. A numeric slot has no word to translate and survives intact. This silently broke about 10% of non-English rows on one site before anyone noticed.

8. Nothing is indexable until it is finished

blocos publicados

Pages ship noindex. Indexing is a deliberate opt-in by the owner, per site.

Porquê?: These sites share a wildcard domain. One spam tenant indexed on that wildcard damages every other tenant on it, and a half-finished page indexed on day one takes months to undo.

9. Build the thing, do not describe the thing

advertências

A page either does something useful or says something specific. A page that only describes a tool is not a page.

Porquê?: Thin wrappers around someone else's API are the clearest signal a quality classifier has. Real utility is what survives a core update.

11. Alt text and intrinsic dimensions on every image

advertências

Every <img> has alt text and explicit width/height.

Porquê?: Alt text is an accessibility requirement first and an indexing signal second. Missing dimensions cause layout shift, which is a Core Web Vitals failure on mobile — and mobile is what gets indexed.

12. Retired URLs return 410, verified

blocos publicados

A removed page returns 410 Gone, confirmed by an actual request after deploy.

Porquê?: A 404 lingers as a soft-404 for months; a 410 deindexes cleanly. And per-view dispatch often intercepts the request before the 410-returning view is ever reached, so the check has to be a real HTTP request, not a code read.

Por que isto é um linter e não um guia de estilo

Uma regra que só vive numa instrução é uma sugestão — o modelo segue-a a maior parte do tempo e calmamente não o resto, e você descobrir três meses depois em um Denunciar de tráfego. Então, estes correm como código, depois da página é escrita e antes de ele se tornar vivo. Uma página que falha uma regra de bloqueio é reenviado com a razão e reescrito. O resultado é no recibo de qualquer forma, para que você pode ver quais os controlos corriam.

E nada é indexado até que você diga

Cada site começa sem índice. Não como padrão você tem que descobrir — como regra o linter força, porque esses sites compartilham um domínio e um vizinho mau magoa todos nele. Quando seu site vale a pena encontrar, você liga indexando em configurações e começa a ser rastreado.

Execute o seu primeiro turno livre