04 — Prior Art: the published record on docs.diniscruz.ai
Ten core articles, 68,846 words, published February – October 2025. A year before the __Send material, in the founder's public voice, already indexed and already circulated on LinkedIn.
This is the site's strongest single asset. It means newsroom.sgit.ai does not launch as a manifesto — it launches as the continuation of a two-year, dated, publicly-checkable thread. Lead with that.
Every URL below was resolved live on 21 August 2026. Machine-readable equivalents, including repo paths and git first-commit dates, are in sources__docs-diniscruz-ai.json.
Before republishing anything here, read 02__source-provenance-and-attribution.md. The contract is: original date, original link, original co-authors, honest curation label.
Core articles
Authors: every core article credits Dinis Cruz and ChatGPT Deep Research except where the JSON records otherwise. Preserve the credit — see 02 §5.2.
Adjacent and index pages
| First published | Title | Words | Why it is here | Link |
|---|---|---|---|---|
| 2025-05-27 | LETS (Load, Extract, Transform, Save): A Deterministic and Debuggable Data Pipeline Architecture ↗ | 9,362 | The LETS method, referenced but never defined in the __Send corpus | source ↗ |
| 2025-06-13 | Technical Briefing: Web Content Filtering Project ↗ | 10,204 | Adjacent: content filtering and transformation | source ↗ |
| — | The Future of news ↗ | 110 | The existing hub page — the site's own prior IA for this topic | source ↗ |
What each one gives the site
| Article | What it carries that nothing in __Send does |
|---|---|
| Micro and Nano Payments (10,729 w) | The largest single treatment of news monetisation anywhere in the corpora, 15 months before the x402 rail existed. Pair it with the Aug 2026 rail research and the pairing itself demonstrates correction-propagation: the argument held, the mechanism arrived. |
| Time as a Calibrator of Credibility and Trust (13,486 w) | The longest and newest piece — statements as evolving entities whose trustworthiness is calibrated by time and accumulated evidence. It is also the answer to a gap flagged in the graphs pack (time as a first-class dimension). |
| Personal Content Rights (9,720 w) | Deepfakes and AI cloning. The only treatment of individual content rights; the __Send CC-Signed brief covers the licensing stick but not the personal-rights case. |
| Strengthening Trust in News: Identity Graphs (6,940 w) | Author and source identity as a graph — the bridge page to pki.sgit.ai and nhi.sgit.ai, already written. |
| From Free Scraping to Fair Compensation (5,660 w) | Cloudflare's crawler charges. The commercial context for the whole rights-and-payment argument, and it dates the thread against a real market event. |
| Journalists' Challenges with Digital Content Provenance (4,403 w) | The practitioner-facing framing — written from the journalist's problem, not the architecture. The site badly needs one page in this register. |
| Monetising Trust and Knowledge (3,549 w) | Personalised semantic graphs for news providers. The earliest statement of the thesis, Feb 2025. |
| Building Trust Through Fact Provenance (3,665 w) | The origin article. ⚠️ Currently renders a literal {{title}} heading — fix at source before citing as canonical. |
| Project InsightFlow (6,082 w) | Regulatory and news feeds transformation — the workflow ancestor of the newsroom briefs. |
| Personalised Briefing for Dan Raywood (4,528 w) | The thesis explained to a named journalist. ⚠️ Names a real person; confirm consent before republishing. |
This document is released under the Creative Commons Attribution 4.0 International licence (CC BY 4.0).
== briefs/05__site-architecture.md