Coverage, and the pitch — spanning the news, and two ways a writer meets a topic coverage BUILT 2026-09-09 · the pitches BUILT 2026-09-09

The owner's answer to the budget-meeting study: the world is big and the harvest sees a patch of it; and a writer can be given a topic by an editor, or can want one. Sections 1 to 3 are coverage. Section 4 is the two doors. Nothing here is built.

The world is big; our current several queries only cover a small patch of the news space. We need to span the space. And there are two ways to match a writer to a topic: the budget editor's match-making, and the persona's spontaneous motivation.the owner, 2026-09-08

1 · Coverage, in four words

Three of the four are already the brew's own. A peg is a candidate story with its link. The harvest is how pegs are gathered each morning. The slate is the day's sixteen commissions. The fourth is the owner's: the space is everything the paper could be about, and it has three sides: where it happened, what it is about, and what kind of thing happened (a ruling, a death, a discovery, a match, a strike, an anniversary, a small strange thing). Coverage is the share of the space today's pegs came from. Nothing else is needed.

2 · The patch, measured

The 09-09 harvest, 82 pegs, placed on two sides of the space: the beat, and where the outlet lives. The third side, the kind of event, the harvest cannot see at all, because it searches by beat name.

World
China
Tech & AI
Science
Culture
Markets
Sports
Life
US outlets
8
·
9
8
24
1
·
4
UK outlets
1
·
1
·
8
·
·
·
global sites
·
·
·
8
·
·
·
8
Chinese sites
·
·
·
·
1
·
1
·
everywhere else
·
·
·
·
·
·
·
·

Thirty-three searches such as「technology AI news today」and「今日体育新闻」, 15 outlets, the New York Times alone 17 pegs. Two thirds of the pegs are American; Culture is a third of them because book, film and art sections open wide; Sports, Markets and China have two pegs between them and「everywhere else」has none. A search by beat name asks Google what is loud, and the loud centre of every beat is the same five outlets.

3 · The harvest, redesigned

Two ways in, then two rules, then a count.

readA fixed list of pages that already span the world, read every morning, no search: Wikipedia's Current events page (a curated list of what happened everywhere, by region), its On this day and Deaths lists, the front or section pages of about twenty outlets across regions (AP, Reuters, Al Jazeera, DW, NHK, The Hindu, Folha, the Straits Times, Caixin, 澎湃, Nature, ArXiv), and a calendar file of fixtures, earnings, elections and launches. The section opener we have does the reading.
+
searchSearches written from a checklist that walks the space, not from beat names: every region once a week, every kind of event twice, every topic on a rota, phrased as the event (「a court ruled」·「a study found」·「a team lost」, plus the place). About sixty searches a day; the checklist is a table in the repo. Chinese searches come from the same checklist through Bocha.
pegsTwo rules on the way in. One story, one peg: the same story from six outlets is merged, so the meeting sees the world and not the echo. No outlet over a tenth of the day's pegs. About 150 a day instead of 80.
the countOn the brew log, beside the metaphor median: how many regions, topics and kinds of event today's pegs came from, and how much of the checklist the week has walked. A count, not a score.
decisionHow much to read and how much to search. Reading is cheap and honest, and Wikipedia's page alone would have given 09-09 pegs from every region. Searching is what reaches the corners no curated page lists. The sketch says read first, search for the rest; the owner sets the weight.

How the aggregators sample, and what we take from it

AppHow it picks what you seeWhat we take
Google Newscrawls ~50k outlets, groups articles into stories, ranks a story by how many outlets carry it and how fast, weighted by outlet authority; editors do not touch Top storiesthe story, not the article, is the unit; breadth of coverage is the importance signal
Apple Newshuman editors pick the Top Stories; the algorithm fills the rest by topic and reading historya curated layer on top of the machine one (Wikipedia plays this part for us)
SmartNews · 今日头条cluster, then rank by engagement and freshness, then personalise hardfreshness matters; personalisation is not the harvest's job
Flipboard · Redditpeople curate topic magazines; votes rank inside a topicthe topic list must be wide before ranking means anything
Ground Newsclusters a story and shows how many outlets, of what leaning, covered itthe outlet count per story is a fact worth keeping and showing

The common mechanic is cluster first, then rank the cluster by how many sources carry it. We already fold duplicates by title overlap; the change is to count what we fold instead of discarding it. A story that arrives from forty outlets in nine countries is important whatever its topic.

Every part touched: the checklist

The specimen's sixty sub-field searches were a first cut, and the owner's question (art? films? games? what else?) is the right one: touched-at-least-once has to be a property of a list, not of luck. The list below is the checklist the search layer walks, about 130 lines under the eight beats plus the kinds of event. Each line is one free Google News search; the rota puts every line through at least once a week and the count on the brew log names the lines a week has not touched. The list is a table in the repo, edited by hand when a gap shows.

BeatSub-fields (each one search)
Worldelections · courts and rulings · war and ceasefires · diplomacy and summits · protests · strikes and labour · migration and refugees · disasters, quakes, floods, wildfires · weather extremes · crime and trials · terrorism · corruption and scandal · military and defence · aid and famine · borders and territory · religion and the churches · the UN and treaties · royals
Chinathe economy · policy and the Party · tech and chips · Taiwan · Hong Kong · trade and tariffs · Chinese science · Chinese culture, film and books · Chinese sport · society and 社会 · the diaspora
Tech & AIAI models and labs · AI in use (medicine, law, school) · AI policy · chips and semiconductors · robotics · cybersecurity and breaches · social platforms and moderation · smartphones and gadgets · apps and software · open source · crypto · VR, AR and wearables · the space industry · EVs and batteries · the internet and telecoms · gaming hardware
Sciencespace and astronomy · space missions and launches · physics · quantum · chemistry and materials · biology and new species · genetics · neuroscience · medicine and trials · public health and outbreaks · drugs and addiction · climate · oceans · energy and fusion · earth science and volcanoes · palaeontology and dinosaurs · archaeology · mathematics · animals and nature · agriculture and food supply
Culturebooks and prizes · poetry · film and festivals · television and streaming · music, pop · classical, opera and dance · theatre · art, museums and auctions · art crime and forgery · architecture · design and photography · fashion · food, restaurants and chefs · video games · esports · comedy · podcasts · comics and animation · language and words · history and anniversaries · philosophy and ideas · celebrity · obituaries
Marketsstocks · bonds and rates · central banks · currencies · commodities, oil and gold · IPOs and deals · startups and funding · earnings · banking and insurance · property and housing · private equity · luxury and retail · airlines and transport · energy companies · media and advertising · trade and tariffs · economics research · philanthropy · scams and fraud
Sportsfootball, world · Premier League · American football · basketball · baseball · cricket · tennis · golf · Formula 1 and motorsport · athletics and marathons · rugby · boxing and MMA · cycling · swimming · winter sports · the Olympics · chess and mind sports · horse racing · women's sport · sport business and doping
Life & relationshipsparenting · dating and marriage · friendship · sex · grief and loss · work and careers · school and universities · money at home · mental health · sleep · fitness · diet and cooking · ageing and retirement · housing and the home · pets · travel · hobbies, gardening, crafts · cities and commuting · faith and meaning · disability · LGBTQ · the odd and the small strange thing
Kinds of event (cross-cut)a ruling · an election · a death · a discovery · a launch or release · a match or race · a strike or shutdown · a price move · a deal · a disaster · a record · a trial · an anniversary · a festival · a scandal · a rescue · a theft

Against this list the specimen's sixty searches missed: religion, weather, crime and trials, military, energy, oceans, chemistry, neuroscience, animals, agriculture, poetry, classical music, art crime, fashion, design, comedy, podcasts, comics, language, philosophy, celebrity, bonds, currencies, commodities, banking, luxury, airlines, media, scams, baseball, boxing, cycling, swimming, winter sports, chess, horse racing, women's sport, friendship, grief, sleep, fitness, diet, pets, travel, hobbies, faith, disability. About fifty of a hundred and thirty. A list is the only way to know that.

Important whatever the topic: three signals, no model

the ruleA story is important when many unrelated sources carry it. An art theft that forty outlets in nine countries report is front-page whatever beat it sits under. The harvest keeps three signals, all counts:
1 · breadth: how many outlets, and how many countries, the fold merged into one line (Google News's own mechanic);
2 · the editions' top stories: how many of the 37 country editions put it in their top twelve (their ranking, borrowed);
3 · Wikipedia's list: whether a person wrote it into Current events.
A line that scores on any two goes to the meeting flagged front, ahead of beat and region, and the slate policy's「every piece pegged」becomes「at least four of the sixteen from the front」. The cheese theft (three signals on 09-07), the Houthi strikes, the AGI tweet: all would have been front by count alone.

The sources: what already aggregates the world

Nobody has to search for the span; several services already gather it, and the useful ones are free. The commercial ones sell the same search shape we have now over a wider index, which fixes the outlet problem and not the query problem.

ServiceWhat it isCostFit
GDELTscans world news in 100+ languages every 15 minutes; tags each article with its location, its themes and the kind of eventfree, one request every 5 sthe wide net: its fields are the three sides of the space; gives the link and the tags, not the text; noisy, so the outlet cap matters; its event codes see rulings, strikes and protests sharply and a cheese theft only through themes
Google News feedsa feed per country edition and per topic over tens of thousands of outletsfreewhat is loud in each country rather than only in the US; one feed per region a day
Wikipedia Current events · On this day · Deathsa human-curated daily list of what happened everywhere, by region, with a source link per linefreethe cleanest one-liners of all; the honest baseline; anniversaries and obituaries only one persona would want
Event Registry (newsapi.ai)groups articles into events across languages~100–500 USD a monthdoes the one-story-one-peg merge for us; priced for companies
NewsAPI · NewsData · GNews · Mediastackkeyword search over 100k+ outlets by country and categoryfree tiers with delays, then 50–450 USD a monthour search shape, wider index; not needed while the free layers hold
Serper newsGoogle's news vertical by query (what we use)~0.001 USD a searchkept for the checklist searches that reach the corners

Gone or closed: Bing's news API (retired 2025), Ground News (no public API). Chinese one-liners come from Bocha and the Chinese outlets' own pages; a Chinese story only Chinese media carried arrives as a Chinese headline, which the meeting reads.

Speed and cost, a day

The current pegs stage on 09-09 took 90 seconds and cost 0.147 USD, all Serper fees; no model gathers pegs. The redesign is faster and about as cheap, because everything wide is free.

Layercalls a daytime, in parallelcost
GDELT by region and kind~454 min (spaced 5 s)free
Google News country and topic feeds~45under 30 sfree
Wikipedia's three pages32 sfree
outlet feeds and front pages across regions~2030–60 sfree
Serper from the checklist~6020 s0.06 USD
one story, one peg (title overlap, local)0under 1 sfree

Three to five minutes of wall time, 0.06 to 0.10 USD. Time is invisible: the text pass starts at 03:00 and the meeting waits for the harvest. GDELT throttles bursts (a 429), so it runs spaced and the harvest proceeds on the other layers when it is slow, as the peg ladder already falls through to Wayback. The box in Singapore reaches all of these directly.

wide and shallowA day's coverage is a few hundred one-liners, not full text. GDELT gives a headline, the link, the outlet and its tags; Google News a headline and a snippet; Wikipedia a sentence written by a person with its source; a front page a headline and a standfirst. Full text is fetched only for the sixteen commissioned and the spares, by the desk's ladder as now. Wide and shallow at the harvest, narrow and deep at the desk. What no harvest reaches: the quiet world, the ruling no outlet wrote up. The specimen is one real day compiled this way.

Tags from birth: what a picker and a recommender will need

A published piece already carries a beat and the free tags the meeting invents that morning (russia · ukraine · naval-war), and the app records each reader's verdicts and how far they read. The reading side of a recommender exists; the supply side does not: sixteen pieces from a patch, tagged in words that change from day to day.

decisionThe peg record carries its tags from birth, in a closed vocabulary. Region, topic and kind of event are fixed lists kept in the repo; GDELT's tags and the checklist's own region, topic and kind fill them at the harvest, they ride the commission, and the edition keeps them beside the meeting's free tags. A closed list is what a topic picker shows and what a recommender counts; free tags alone drift. And the harvest keeps its tagged one-liners rather than discarding what the meeting did not commission, so a reader who picks「Southeast Asia」and「science」can one day pull from the few hundred the harvest saw, not only the sixteen we wrote. A first recommender needs no model: topic tags from the harvest, the writer as a tag of its own (readers follow writers as much as topics), and the reads and thumbs we have. Small now, expensive to retrofit.
the rule · owner 09-09Rank by the importance of the news, never a proportion per region or topic. The owner's correction, twice over: balance is not a uniform spread, and it is not a measured share either; a fixed number for any region or topic is the wrong tool whatever it is set to. So the selection ranks by attention, which the harvest counts without a model: the outlets that carried a story, the countries and the front pages it was in, whether a person listed it on Wikipedia's Current events. If the Middle East is forty percent of the day's attention it is forty percent of the list. Variety is a diminishing return on sameness, the way the aggregators do it: the k-th chosen story from the same running story, the same outlet, the same sub-field or the same beat counts for less, so the top is not sixty lines on the same three stories. The long tail keeps forty seats: after the attention list fills the rest, the newest story from each sub-field the list did not touch, so every part is at least touched. That is the only place a count of seats exists. A single country's cluster saturates (a local top story is not world attention), non-English lines are discounted and capped at twelve for the desk's sake, no outlet takes more than a tenth, and the front flag and the meeting's duty to take front pieces stay.

The wide harvest, as a picture

read
Wikipedia · Current events · Deaths · On this dayGoogle News · 37 country editions · top 12Google News · 6 topics × 5 editions54 outlet feeds · regions and topics
~1,200 lines · free
search
the checklist: 159 sub-fields under the eight beats + 17 kinds of event, every line every day as one free Google News search (when:2d, 8 results)
176 searches · ~1,400 lines · free
fresh
a line older than three days is out; a calendar line is exempt
~2,600 lines
fold
one story, one peg. Google's own cluster rides each top story (the other outlets' headlines) and becomes the story's alias titles and its outlet count; two lines are one story when titles overlap 0.45 or one holds 70% of the other with four shared words. The story keeps the best line (a publisher URL over a Google stub, a headline over Wikipedia's sentence) and counts the rest
~2,250 stories
outlets · countries · editions · wikipedia
signals
breadth (4+ outlets or 2+ countries) · editions (in 2+ editions' top stories) · wikipedia (a person listed it). Two of three = FRONT, whatever the beat
12–18 front
select
by attention (outlets · countries · front pages · Wikipedia), greedily, each pick discounted by what is already chosen: the same running story ×0.35, the same outlet ×0.75, the same sub-field ×0.8, the same beat ×0.97, a non-English line ×0.4 → 120 seats; then the tail: the newest story from each sub-field the list did not touch → 40 seats; no outlet above a tenth
160 for the meeting
keep · count
the whole tagged pool stays in pegs-all.json (beat · region · sub-field · kind · layer · outlets · countries · front); coverage.json and two log lines count what the pool and the selection touched
~2,250 kept · 0 USD · ~1 min
the shortlist
code
every pitch that stood (persona · peg · its reason · its quote) · every peg that names a persona (the full name, the curated short name, or a surname of five letters or more) · the front stories, open · the top ten unpitched stories by attention, open — the door. Beside it, the bench: who is free to write, the idle first
25–31 candidates · bench 74
the pick
V4.1 Flash
one call, thinking off: read the shortlist and the bench, pick 16 + 4 in editorial order with beat, format, register and a one-line subject; a pitched line is taken as it stands; an open line needs a persona from the bench and a stated reason — a memory, a position, a place, never the field or the fame
~12 s
the questions
V4.1 Flash
one small call per commission, all at once: the peg, the persona, its pitch or the editor's reason → the ONE question, never a thesis; a failed call gets a plain fallback so the day never waits
20 calls · ~2 s
the repairs
code
the validator's checks as before, plus one peg, one piece; a shortage filled from the unused pitches; no light piece → the last main commission becomes the Kicker; spares filled from the shortlist; the kicker moved last; beats over four and a seventh Essay trimmed from the tail. One small re-ask, thinking off, only for what code cannot repair, with the gaps named and the unused candidates listed
~10 s · 0 problems left
the slate
code
16 commissions in editorial order + spares, each with subject · question · door · pitch or reason; frozen to the day's file; the writer reads its own pitch back in the commission, or the editor's reason
the whole meeting ~24 s (was ~9 min)

Green is counting and rules, auditable from the files; blue is the cheap model, reading as a persona or picking from a shortlist with thinking off. The owner's rule (09-09 evening): ten minutes for this step means the design is wrong. The one giant call (154 pegs · the pitches · 90 roster lines · the policy → 20 commissions with questions, thinking on) starved on every model; the meeting is now a shortlist made in code, one fast pick, twenty small question calls and repairs in code, and the door stays open through the ten unpitched stories by attention. Same inputs, 24 seconds instead of nine minutes, no problems left. What the owner's read of 09-09 found: the pitch door held; the assignment door let a name pun, a brand and a register through, because the validator only checks that a reason exists.

i · The editor's match-making, kept for the holes

The meeting stays, but it stops being the only door. It is the right tool when the day needs a story and no one wanted it: the war on its fourth day, the budget, the thing that happened at 5 a.m. It gets the fixes from the study: the bench first, a reason that cites a line of the profile, no professional adjacency, a quota on how many of the day's slots it may fill by assignment. What it loses is the job of inventing a writer's interest, which is the job it did worst.

ii · The persona's spontaneous motivation, the pitch

the ideaA writer with a reason writes better than a writer with an assignment. Real newsrooms run on pitches: the writer says what they want to write and why, the editor says yes or no. The far-fetched pairing disappears because the reason is the persona's own and comes from its own life; the stereotype thins because a persona pitching from its memories does not pitch from its Wikipedia field; the bench fills itself because everyone reads the paper every morning.

The pitch pass. One cheap call per persona (Flash-class, ~4k tokens in, ~200 out). The prompt is the day's headlines first, as the shared, cached prefix, then the persona's motivation sheet, then one instruction: read the morning's paper as yourself; if something in it is yours, say which, and why, in two sentences, and in what form; if nothing is, say so. Silence is a valid answer and most days most personas are silent. Cost at 92 personas: roughly 0.1 to 0.2 USD a day on DeepSeek Flash with the prefix cached, less than one piece's writer call.

What a pitch is. Three lines, machine-readable: the peg (by id), the reason (which must name something in the profile: a memory, a value, a place, a person), the form. Two examples of what the pass should produce, written by hand here:

Phil Knight · peg: the AP college-football poll · form: ReviewI ran for Bowerman at Oregon and built a company on college sport's back. A poll voted by sportswriters that decides who plays for the title is a thing I have argued with for fifty years. Let me say what it is and what it is not.
Anthony Bourdain · peg: the Neal's Yard cheddar theft · form: KickerTwenty-two tonnes of clothbound cheddar stolen by a man pretending to be a French distributor. I have eaten that cheese with the people who make it. This is the best crime of the year and I want the last word on it.
Roger Federer · peg: the AP college-football pollNothing today.

The motivation sheet. The pass needs something the profile does not have. The card's writes: is the language a persona writes in, not its subjects; nothing on the card says what a persona would want to write about this month. The sheet is a derived block, ~300 words, made once per persona from Values & seams and Specific memories and refreshed when the profile changes: the five or six things this person cannot walk past, each tied to the profile line it comes from, plus the things they are following (a persona may follow a story across days, which is how series happen). It is not authored by hand for 92 personas; it is extracted by one call each and cached beside the card, and it is the only thing the pitch pass reads besides the headlines.

What still needs an editor

FailureWhy the pitch alone does not fix itThe editor's rule
everyone pitches the warthe loudest story is everyone's story; ninety hands go upa running story at most twice a week, whoever pitched it
the same twenty hands every daya persona with a rich sheet pitches more often and betterthe bench: a pitch from a persona idle fourteen days outranks an equal one; a floor of a third of the slate
a reason that is a pun on famea model can always write two sentences of reasonthe validator checks the cited profile line exists; the editor refuses adjacency without a memory
no one pitched a story the day needs150 pegs have holes no pitch fillsmatch-making, for at most a quarter of the slate, with the reason cited
the pitch is better than the piecea two-sentence want is not an 800-word essaythe pitch travels with the commission: the writer reads its own reason back, which is the「you are in it」line the first read asked for, written by the persona itself
decisionThe split. The sketch says three quarters pitched, one quarter assigned. A paper that is all pitches drifts toward what its writers love; a paper that is all assignments is what we have. The owner sets the fraction.
decisionThe living and the dead. A living persona pitches on today's headline; a fixed-horizon one, by the vantage rule, has no today, so its pitch is on the pattern (「a naval blockade; I have seen one」), and the desk supplies the facts. The pitch pass should say which of the two it is doing, so the far-fetched pairing of Lincoln and a train is refused at the pitch, not at the gate.
decisionWhere the sheet lives. Beside the card as a cached derived file (like the voice bank), or as a tenth section of the profile written by the pipeline. The first is cheap and reversible; the second makes it part of the persona, which it arguably is.

5 · What this changes upstream

StageNowProposed
the harvest33 beat-name searches → ~80 pegs from 9 outletsbuilt 09-09: read 37 editions + 54 feeds + Wikipedia, search the 176-line checklist → 2,200 stories, 160 read, one story one peg, no outlet over a tenth, no beat over 22%; the coverage count on the log
before the meetingnothingthe pitch pass: 92 calls, 0–2 pitches each, from the motivation sheet
the meeting140 pegs + a 92-line roster → 16 commissionsthe pitches + the coverage sheet + the bench → ~12 accepted pitches + ~4 assignments, each with a cited reason
the commissionsubject + questionsubject + question + the persona's own pitch, read back
costpegs ~0.15 USDpegs ~0.20 (more searches; reading pages is free) + the pass ~0.15; the rest unchanged
statusDesign in discussion, 2026-09-08. No code. Siblings: the budget-meeting study · the upstream plan · what the writer reads.
design 2026-09-08 · siblings: the budget-meeting study · the brew