Referential Cohesion in Text: How It Actually Works
Cohesive referential chains are what keep a text from falling apart when you move past a single sentence. When you write about someone and then switch to "ele", the reader doesn't have to search for who you mean. That's the basic mechanism. But the real work happens when things get messy, which they always do in practice.
O que é coesão referencial e por que a maioria dos textos falha nisso
Referential cohesion is the network of relationships that link noun phrases to their referents across a text. Pronouns, demonstratives, syntheses, and ellipsis all participate. Unlike lexical cohesion, which relies on semantic repetition or synonymy, referential cohesion tracks identity through the text. The reader constantly resolves "who or what" is being talked about at any given point. In practice, I've seen this break down repeatedly in technical documentation, especially when writers shift between third-person narration and direct address mid-paragraph. The pronoun doesn't change form, but its antecedent becomes genuinely ambiguous. A developer reads a manual where "o sistema pode falhar" and has no way to know whether "sistema" refers to the server cluster, the application layer, or the entire deployment. The text is grammatically correct. The cohesion is broken.
The workaround I settled on years ago is simple and nobody writes about it: when introducing any entity that might be confused with another, anchor it immediately with a restrictive clause or parenthetical label, then use only that form going forward. Instead of "o servidor" followed by "ele", write "o servidor de banco (MySQL)" on first mention, and then commit to one referential chain for that entity. Stop mixing "o sistema", "o servidor", and "ele" across three sentences referring to the same thing.
The Mechanics Behind the Scenes
There are specific devices you'll encounter: pronominal reference (he, she, it, they, seu, sua, isso), nominal substitution (replacing a noun with a synonym or hypernym), elliptical reference (omitting an element recoverable from context), and anaphoric versus cataphoric direction. Anaphora points backward to something already established. Cataphora points forward, usually for stylistic or structural reasons, like "Quando chegou o relatório, todos ficaram calmantes" where "todos" is anticipatorily introduced before we know who it is. The counter-intuitive part most people miss is that longer texts don't automatically need more cohesive devices. In fact, dense referential chains in long documents cause more problems than they solve. Readers lose track of who "seu" belongs to when the referent is eight sentences back. Shorter texts benefit from explicit repetition of the full noun phrase at key transitions.
👉 Clique no botão abaixo para saber mais sobre o assunto!
Another thing that trips people up: possessive adjectives in Portuguese (seu, sua, nossos, deles) are notoriously ambiguous without context. "Maria encontrou o notebook de João e seu teclado estava quebrado." The "seu" could belong to Maria or João. This isn't a style preference. It's a genuine resolution problem. The fix is restructuring the sentence so the possessive attachment is unambiguous, or replacing the possessive with an explicit noun phrase.
Where It Completely Fails
Machine translation tools handle referential cohesion poorly because they process sentence by sentence rather than paragraph by paragraph. Google Translate and DeepL will frequently drop pronouns, substitute the wrong antecedent, or render a possessive as a demonstrative. If you're working with technical content in Portuguese, run any translated document through a pass where you verify every pronominal reference against the original. This typically catches the most damage in about 20 minutes for a 10-page document. Automated plagiarism checkers also ignore referential cohesion entirely. They look at surface-level lexical overlap. Two texts can share zero identical phrases and still be cohesionally identical in structure. That's a limitation worth knowing if you're doing any kind of text comparison or originality verification.
Practical Application
When building a text, establish your referents early and maintain consistency. The first time you mention an entity, use its full name. Subsequent mentions can use pronouns or shorter forms. Don't introduce a new referential chain without a clear marker. If you need to shift which entity you're discussing, do it in a new paragraph or with an explicit transition, not mid-sentence. For academic writing in Portuguese specifically, the "seu/sua" ambiguity is the single biggest cohesion failure point I see. Replace with "de + nome próprio" or restructure the clause whenever there's more than one possible antecedent in the preceding context. It adds words but it removes reader errors. In my experience, this small change reduces revision requests on draft documents by roughly 40 percent because the ambiguity complaints disappear entirely.
Ellipsis works well in tight contexts but degrades fast outside them. If you omit a subject and the reader has to backtrack more than two sentences to recover it, the ellipsis has failed. When in doubt, include the explicit form. The cost is minimal. The clarity gain is significant.