Front page — July 11, 2026
The Peloton Dispatch July 11, 2026 No. 105
● 72°F and sunny — summer kit, go outside. · summer kit

THE LAB

Carmack on Microsoft Gutting id; Sol Closes a 50-Year Proof; 744B in Pure C

John Carmack couldn't find the words — not the right ones, anyway. "I have been trying to find something meaningful to say about the Id Software layoffs," he posted Thursday. "My 'Microsoft will probably be a good steward of the brand' statement isn't aging well." As this paper reported Wednesday, Microsoft's 3,200-person Xbox cut had effectively reduced id Software to a support role following the release of DOOM: The Dark Ages Revelations DLC. Carmack, who co-founded the studio in 1991 alongside Adrian Carmack, Tom Hall, and John Romero, reached for pragmatism rather than outrage.

What he produced was a question list. Could id have gotten more with a different pricing strategy? Could they have broadened the game designs to reach more players without alienating existing ones? Could they have produced games faster or cheaper? "I really don't know," he writes — and from the person most qualified to answer, that's a meaningful admission. He suspects "id Software was a marginal business from Microsoft's perspective" and credits Minecraft revenues with carrying several other Xbox studios. There is one careful hedge: "You can't rule out the possibility that executives are idiots, but that shouldn't be your default belief." The four id co-founders are due to reunite at QuakeCon next month for the first time since 1993; that occasion will now be considerably more somber. "The game isn't over yet, and I hope the studio rallies through."1


There is no dependency to install. No Python at runtime. No GPU required. The inference engine is a single C file — roughly 2,400 lines — and it runs GLM-5.2, a 744-billion-parameter mixture-of-experts model, on a consumer machine with 25 GB of RAM. The project is colibri, by JustVugg on GitHub, and its architecture starts from a simple observation: a 744B MoE model activates only ~40B parameters per token, and only ~11 GB of those change between tokens. The dense portion — attention layers, shared experts, embeddings, roughly 17B parameters — stays resident in RAM as int4 at about 9.9 GB. The 21,504 routed expert shards, roughly 19 MB each at int4 and totaling ~370 GB on disk, stream on demand via a per-layer LRU cache, with the OS page cache serving as a free second tier.

The implementation earns its attention. Compressed MLA attention reduces the KV-cache from 32,768 floats per token to 576 — a 57x reduction, made possible because GLM-5.2 uses 64 attention heads and no grouped-query attention. The model's native multi-token-prediction head runs speculative decoding at int8, achieving 39–59% acceptance rates and 2.2–2.8 tokens per forward pass. A learning cache records which experts each session actually routes to and automatically pins the hottest ones in spare RAM at startup — the engine measurably accelerates across sessions. Community benchmarks are honest: 0.05–0.1 tok/s cold on the 25 GB dev box; an Apple M5 Max with 128 GB unified memory and a 14 GB/s internal SSD reaches 1.06 tok/s; a Framework 13 on Arch Linux with 128 GB goes from 0.29 to 0.37 tok/s once the learned pin warms up. The README does not oversell: "This is not fast. It is a 744B frontier-class model answering correctly on a machine that costs less than one H100 fan."2


GPT-5.6 Sol Ultra — the model this paper noted in limited rollout yesterday — has now produced what OpenAI describes as a machine-verified proof of the Cycle Double Cover Conjecture. The conjecture was posed independently by George Szekeres in 1973 and Paul Seymour in 1979; it states that for any bridgeless graph, there exists a collection of cycles covering every edge exactly twice. Partial results existed for specific cases; a general proof did not. Sol deployed 64 subagents working in parallel, and the proof reportedly reduces the problem using the 8-flow theorem and linear algebra over GF(3), the three-element finite field. OpenAI published the result as a PDF on July 10, with authorship attributed to the model; Codex assisted with the writeup.3

Machine-verified is not peer-reviewed. Formal verification can confirm that logical steps follow from axioms; it cannot explain why the proof works or certify that it illuminates the structure of the problem in the way mathematicians demand. The community's response on Hacker News zeroed in on this immediately. If the result holds up to mathematical scrutiny over the coming weeks, it clears a meaningful bar — not game-playing, not benchmark optimization, but original mathematics at frontier difficulty, on a problem that defeated human specialists for fifty years. The community will take the time it takes.


Apple filed suit against OpenAI in US district court in San Jose on Friday, naming OpenAI chief hardware officer Tang Tan — who spent 24 years at Apple overseeing iPhone product design — and io Products, the Jony Ive–co-founded startup OpenAI acquired for $6.5 billion last year. The filing accuses Tan of coaching departing Apple employees on how to evade the company's data security protocols and directing them to bring confidential hardware to job interviews at OpenAI. A second defendant, electrical engineer Chang Liu, reportedly downloaded "dozens of Apple's confidential hardware-related files" including manufacturing and testing documentation for complex circuit boards, then wrote to a former Apple colleague about still having access to the company's internal file-sharing system months after his departure — via a bug Apple says it has since patched. Apple wrote to OpenAI in February raising those concerns; it received no response. OpenAI has hired more than 400 former Apple employees. The case echoes the 2017 Waymo v. Uber dispute, which settled mid-trial in 2018 for $245 million.4 Apple and OpenAI have been partners since 2024 via the ChatGPT-on-iPhone deal, but according to Wired, the relationship has frayed as Apple has shifted toward Google's Gemini for its in-house AI models.

Sources
  1. 'You Can't Rule Out The Possibility That Executives Are Idiots' — John Carmack On Microsoft's Gutting Of id Software timeextension.com Jul 10, 2026
  2. JustVugg/colibri — GLM-5.2 744B MoE Inference in Pure C, Zero Dependencies, Streams from Disk github.com
  3. GPT-5.6 Sol Ultra Proves 50-Year-Old Cycle Double Cover Conjecture cryptobriefing.com Jul 10, 2026
  4. Apple Sues OpenAI for Allegedly Stealing IP and Hardware wired.com Jul 10, 2026

↑ Back to top

THE PELOTON

Stage 8 Rolls to Bergerac; ASO Concedes the Tourmalet Miscalculation

↩ Developing story — first reported Jul 08 · previously Jul 09, Jul 10

— Three riders are off the front somewhere in the Dordogne, and they are going nowhere. Liam Slock (Lotto-Intermarché), Thibault Guernalec (Totalenergies), and Jakub Otruba (Caja Rural-Seguros RGA) got clear in a scrappier-than-usual opening scramble — the Lotto directeur sportif had warned before the départ that today's bumpier terrain would produce "a little more action than the previous days" — but Soudal-QuickStep and Alpecin-Premier Tech took the front of the peloton within the hour and capped the gap to a stingy 1:15 to 1:50.1 The sprint setup is on rails.

Stage 8's flat route from Périgueux to Bergerac is the second consecutive sprint stage, and the fast-finishers' teams are running the same script they ran into Bordeaux: a token break allowed, steady tempo all day, deliver for the line. Kasper Asgreen (EF Education-EasyPost) tried repeatedly to join the escape and was marked out by NSN, whose directeur sportif clocked him from the car. "I saw Asgreen, and he has shoe covers, aero helmet. I think he is ready," the NSN DS radioed.1 He was not let go. No sprinter has asserted clear dominance across two sprint finishes so far. Bergerac may sharpen that picture.


The more contested story at this Tour is happening inside the Red Bull-Bora-Hansgrohe bus. As this paper reported Thursday, Remco Evenepoel aired his frustration publicly after Stage 6 — he had asked Florian Lipowitz for one kilometre of work on the mountain and was refused, and he said so, loudly and specifically, at the finish line. The team moved to contain the fallout: manager Ralph Denk called it "no big deal," said the pair had dinner together that evening, and on Friday morning the team posted an Instagram story of both riders on the bus, with the press officer prompting Evenepoel on camera about whether the press reports of tension were accurate.2

"No, all good," Evenepoel said. "You speak and you forget."2

A CyclingNews analysis published Friday argued that framing isn't convincing anyone. The structural problem hasn't been solved by a shared dinner and a managed Instagram post. Both riders have finished third at the Tour; Lipowitz's podium came last year, in a race Evenepoel abandoned; Evenepoel has a Vuelta title that Lipowitz doesn't.2 Neither has a clear reason to cede leadership to the other, and the Pyrenean stages haven't finished the argument for them. The road will decide, as the old line goes — but so far it hasn't. The Pyrenées are not done.


The more consequential reckoning this week belongs to ASO's route designers. Stage 6's 40 kilometres between the Tourmalet summit and the finish in Gavarnie-Gèdre — a 20km descent followed by a 20km valley drag — was conceived as a moderating force. The subtext was clear: no right-minded rider would want to solo into a headwind for that long after two major cols. It produced nearly three minutes between first and second place after one week of racing.

Thierry Gouvenou, ASO's route designer, has now said the quiet part aloud. "To be honest, we had not expected such a large gap, and we thought the differences at the finish would be much smaller," he told TV 2 Sport. "As far as the suspense is concerned, you could say it was a failure."3

The arithmetic is worth sitting with. The gap at the Tourmalet summit was 30 seconds. At the finish it was 2:38. Forty seconds came on the descent; another 1:28 leaked on the valley road.3 The terrain designed to discourage aggression handed the field's best descender and most accomplished long-range soloist 40 free kilometres where he has no competition. Fletcher's analysis makes the uncomfortable case that a summit finish — which would have appeared more dramatic on paper — would paradoxically have produced a smaller gap, by removing precisely the descent and drag terrain where the GC leader has the clearest advantage over his rivals. In trying to mute the Tourmalet, the organisers amplified it. "He is so strong that any route suits him," Gouvenou conceded.3

Christian Prudhomme had sold this Tour on the promise of "suspense until the end." One week in, the last two weeks look like a procession.


While the GC narrative has narrowed, Baptiste Veistroffer is making sure the race has something else to offer. The 26-year-old from Brittany (Lotto-Intermarché) has spent 301 kilometres on the attack across just over 1,000km of racing — more than 2,000km of attacking in total across the 2026 season, the most of any rider in the peloton.4 He attacked alone from the gun on Stage 5 to Pau, launched again with Otruba into Bordeaux on Stage 7, and collected his second combativity award of the race. His team, which lost its sprint leader Arnaud De Lie to illness before the start, has extended his contract.4

His nickname is "Le Sanglier" — the wild boar — for his build and his natural aggression.4 He is in the tradition of the classic French baroudeur: Thomas Voeckler, Jacky Durand, Thomas De Gendt. After his Bordeaux effort he gave his combativity award flowers to his mother and said: "I don't understand why I'm the only one going on the attack. There are riders we never see all day, even though we have an incredible opportunity to ride the Tour."

He is not in today's break — that role went to his teammate Slock — but his philosophy has spread. Lotto-Intermarché, with a fraction of the budgets around them, are the Tour's most consistent source of racing animation on the flat stages. "For the love of the sport and the beauty of the effort," Veistroffer wrote on Instagram after Bordeaux.4 France, for the moment, agrees.

On the Road Ahead
Calendar from Jul 10, 2026 — primary source blocked today
DateRaceCountry
Sat Jul 12 – Sun Jul 26Tour de France, Stages 9–21 (ongoing)France
Sat Aug 1Donostia San Sebastián KlasikoaSpain
Mon Aug 3 – Sun Aug 9Tour de PolognePoland
Sun Aug 16ADAC Cyclassics HamburgGermany
Sat Aug 22 – Sun Sep 13La Vuelta a EspañaSpain
Show Results

STAGE 7 (Jul 10 — Bordeaux sprint): WINNER: Tim Merlier (Soudal-QuickStep) PODIUM: 1. Merlier 2. Søren Wærenskjold (Uno-X Mobility) 3. Biniam Girmay (NSN Cycling Team)

STAGE 8 (Jul 11 — Bergerac sprint): Underway at publication — no result available.

Sources
  1. Tour de France Stage 8 Live: Sprint to Bergerac cyclingnews.com Jul 11, 2026
  2. Evenepoel and Lipowitz Are Putting on a Happy Face — Is Anyone Buying It? cyclingnews.com Jul 11, 2026
  3. The Tourmalet Stage Design Backfired Dramatically — And They Should Have Seen It Coming cyclingnews.com Jul 11, 2026
  4. Baptiste Veistroffer Has Become the Breakaway Hero This Tour de France Needs cyclingnews.com Jul 11, 2026
  5. Tour de France Stage 7: Tim Merlier Banishes the Bunch in Bordeaux cyclingnews.com Jul 10, 2026
  6. UCI Race Calendar 2026 procyclingstats.com

↑ Back to top

THE WORLD

Spain in the World Cup Semifinals; Boeing Opens Everett 737 MAX Line


Sources
  1. World Cup 2026 today — Spain vs Belgium, England vs Norway quarterfinals espn.com Jul 11, 2026
  2. Boeing opens new 737 MAX production line at Everett factory kiro7.com Jul 11, 2026
  3. Link light rail ridership records set during World Cup issaquahreporter.com Jul 10, 2026
  4. King County and Kirkland celebrate opening of contentious 101-unit supportive housing facility theurbanist.org Jul 10, 2026
  5. WTA Trip Reports (recent) wta.org Jul 10, 2026

↑ Back to top

THE LONG READ

The Compiler Became the Bottleneck

For seven years, the backend at Scarf ran on Haskell. The type system caught real bugs. The high-performance gateway service, sitting directly in the download path for a high volume of open source package traffic, met its SLAs. The code was, by the founder's own account, reliable. And then AI changed the tradeoffs, and the reliability stopped being enough.

Avi Press — who sits on the Haskell Foundation board and the Haskell.org committee, and has sixteen years of personal investment in the language — published a post on Thursday explaining why Scarf is migrating to Python.1 It is not a hit piece. It is something more useful: a precise account of which specific property of the language became untenable, and why that property was fine before and isn't now.

The argument runs like this. Historically, you caught errors in one of two places: at compile time or at runtime. Press identifies a third place: code generation time. The model can avoid the mistake before the compiler ever sees it. That doesn't make type safety worthless — he's explicit that it doesn't — but it changes the cost-benefit calculation on compile time significantly. "If an LLM can produce a working implementation in a few minutes, but your compile step takes dramatically longer, then your language and build system have become a bottleneck in the development loop."1

That's the argument in its mild form. The sharper version is about parallel agent workflows. Press describes wanting to spin up multiple worktrees, fork different lines of work, let agents try things, review the results, keep the useful ones. In that world — five agents exploring five branches simultaneously — cold build time isn't a tax you pay once, it's a tax you pay five times over, per round of iteration. Caching helps but doesn't solve it. The Scarf team used Nix, remote builders, carefully optimized CI, and all the rest. The best-case scenario, a small change hitting a warm cache, could get to twenty seconds.1 But the best case is not what you optimize around when you're running agents in parallel. The cold-start case, the deep-change case, the cache-miss case — those become the average. "In an agent-heavy workflow, you end up caring a lot more about the cold-start case, the average case, and the deeper-change case."

What they did about it was methodical. New API routes go into Python. A Python server was deployed alongside the Haskell one, traffic routed appropriately, functionality migrated as it was touched. No dramatic cutover. The porting work itself, Press notes, was cheap in a way it would not have been before: porting existing code to a new language is straightforward for current models. The time recovered from waiting on the toolchain went to shipping more features and writing more tests. The type safety they gave up "hasn't been noticeable in any concrete way yet, especially considering our test coverage has never been better."1

The second half of the post broadens into an argument about the Haskell ecosystem's response to AI — and this is where it gets uncomfortable to read if you care about Haskell. Press says that when AI comes up in Haskell spaces, the conversation "seems more focused on restriction than enablement."1 He identifies a "do not use LLMs" cohort and predicts it will prove bad for the language's ecosystem long-term. His counterproposal is specific: optimize Haskell for agents. Make project bootstrap fast. Make error messages agent-friendly. Reduce cold build times. Get high-quality Haskell examples into model training data. Make library docs full of copy-pastable examples rather than beautiful types. These are the bottlenecks agents actually hit, and they're different from the bottlenecks humans hit.

The observation that carries most weight is this: Haskell should be unusually well-positioned in the AI era, because type safety can be an advantage for LLM-generated code if the compiler helps the agent converge quickly on a correct implementation. Fast feedback from a rigorous type system is exactly what an agent benefits from. The problem is that "unusually well-positioned in theory" requires the toolchain to actually be fast enough that the compiler is doing the agent a favor rather than making it wait. Right now it isn't.

Press is writing from a position of genuine investment. He isn't someone who tried Haskell once and bounced. He built real production systems on it, understood where the sharp edges were, and managed them for years. The credibility matters because the diagnosis does. What he's identifying isn't a flaw unique to Haskell: it's a structural shift in what matters when the inner development loop is driven by agents rather than humans. Any language ecosystem that optimizes primarily for human writing speed while neglecting cold-start performance and toolchain disposability is going to feel this same pressure. Haskell is just where it showed up first in a visible way, partly because Haskell's build characteristics make the gap especially wide.

The piece is worth reading in full. It's the kind of postmortem that's useful precisely because it doesn't come from an outside critic.

Sources
  1. After 7 Years in Production, Scarf Has Reluctantly Moved Away from Haskell avi.press Jul 10, 2026

↑ Back to top

FROM THE ARCHIVE

The Littering Fine That Floated for Thirty Years: July 11, 1979

Lead illustration

Night over the Western Australian outback — flat red earth, scattered spinifex, ink-black sky. A charred length of aerospace tubing, still faintly glowing at its edges, has gouged a long furrow in the dust. A shire ranger in a wide-brimmed bush hat stands at arm's length from the debris, citation pad in hand, pen poised in mock ceremony. Above, a bold diagonal streak of cross-hatching traces the object's descent path from the upper corner of the frame down through the darkness. Deep shadow hatching on the debris, strong silhouette of the lone official, generous white space in the sky. Confident pen-and-ink linework, newspaper editorial quality.

The descent was supposed to end in open ocean. NASA had run the numbers, oriented the thrusters, and pointed 85 tons of derelict space station at a remote patch of the southern Pacific. What it had not accounted for, precisely, was a few degrees of arc—enough to swing the debris field from deep water to rural Western Australia.

Skylab had been drifting for five years when it fell on July 11, 1979. Three crews had lived inside it in 1973 and 1974: 28 days, then 59, then 84, each a US endurance record at the time.1 Then the final crew departed, the station was powered down, and NASA redirected its budget toward the Space Shuttle. Skylab would manage on its own for a decade, or so the plan went.

The sunspot cycle had other ideas. An unexpectedly active solar maximum heated the upper atmosphere, increasing drag on the station. By 1978 it was clear the orbit was decaying faster than projected. The Shuttle, not due until 1981, could not be accelerated in time for a reboost mission. NASA worked on a controlled reentry instead.1

The controlled descent was not, in the end, controlled enough. Sonic booms startled residents across swaths of western Australia. Debris — what survived the atmospheric passage — scattered across the Shire of Esperance and into the desert beyond. No one was hurt. NASA's investigation team flew to the site to collect remnants and inspect for damage.

When they arrived, the president of the shire had arranged a ceremony. As NASA public affairs representative J.M. Jones recorded in a 1979 agency newsletter: "Upon our arrival, the president of the shire had arranged a mock ceremony in which an officer of the parks service ticketed NASA for littering, the evidence having been found all about the country-side." The fine was $400.2

NASA never paid it. For thirty years, the ticket sat on the books of a small Australian municipality — an unpaid reminder of the day the first US space station came down in the wrong hemisphere. Then in 2009, California radio DJ Scott Barley asked his listeners to crowdfund the debt. They did. The mayor of Esperance told Barley the fine had already been "written off years ago," but accepted the $400 anyway and invited Barley to town, where he received a key to the city.2

The backup Skylab orbital workshop is on display at the Smithsonian's National Air and Space Museum, where it has been since 1976.1 A few charred fragments of the actual station are in the collection.

Sources
  1. Skylab: When America's First Space Station Fell to Earth airandspace.si.edu Jul 11, 2014
  2. NASA's Unpaid $400 Littering Ticket from Skylab Debris in Australia mentalfloss.com Nov 3, 2015

↑ Back to top

THE FUNNIES

Space Junk Bureaucracy; The Route That Backfired

*After Dilbert — on the $400 littering fine the Shire of Esperance levied against NASA when Skylab's controlled Pacific reentry landed in the wrong hemisphere on this day in 1979, and the agency's characteristically decisive response over the following thirty years. After Bloom County — on ASO's route designer conceding that the forty kilometres of valley drag meant to moderate the Tourmalet's gaps instead handed the field's best descender forty free kilometres of low-competition road, which is, as concerns suspense, "a failure."*

Hand-drawn parody comic strip
AI-rendered parody comic strip

↑ Back to top

ALSO NOTED

Also Noted

↑ Back to top

THE QUESTION

When the Distribution Shifts

A system calibrated for the average actor behaves differently when the actor distribution shifts. Two pieces in today's paper arrive at this problem from opposite domains — one about stage design, one about toolchain choices — and neither story has an obvious correction.

When the Tourmalet summit comes with a 20km descent and a 20km valley drag attached, the reasoning is sound for most editions of the peloton: no one would rationally push a solo effort across that terrain after two hard climbs. The terrain moderates. But as THE PELOTON reports today, ASO's route designer has now admitted that the logic failed precisely against the one rider it most needed to work on. A 30-second gap at the summit became 2:38 at the finish.1 The terrain designed to discourage long attacks gave the field's best descender and most accomplished long-range soloist 40 kilometres of low-competition ground. "He is so strong that any route suits him" is another way of saying: the protection was calibrated for a different distribution of riders than the one that showed up.1

THE LONG READ today turns on the same inversion. Haskell's compile-time type system is the right tool when developers make the kind of errors a type checker catches, and for seven years it caught them reliably. In an AI-agent development loop — multiple worktrees running in parallel, iterating faster than a human — the protection remains valid in principle, but the build time it demands is now the bottleneck.2 Agents don't produce type errors at the same rate a human does; they mostly pay the latency cost without receiving a proportionate share of the benefit. The safeguard that fit the original actor became the obstacle for the one who replaced him.

The question worth carrying today: when you optimize a system for the actors who existed when you designed it, do you inadvertently create something that works for everyone except the exceptional case you most need it to handle? Stage design can be redrawn for next year's route. A language migration can be executed methodically, as Scarf is doing.2 The harder problem is the recognition lag — knowing that the actor you assumed is already gone before the gap at the finish line makes it obvious.

Sources
  1. Tour de France: The Tourmalet stage design backfired dramatically — and they should have seen it coming cyclingnews.com Jul 11, 2026
  2. After 7 Years in Production, Scarf Has Reluctantly Moved Away from Haskell avi.press Jul 10, 2026

↑ Back to top

Investigator Report

Investigator report — 2026/07/11

Verdict

A strong edition editorially — the Carmack lede lands, the Tourmalet reckoning is well-reported, the Scarf/Haskell long read is genuinely useful, and the cross-domain Question is the sharpest structural angle the paper has found this week. The two failure modes are both spec compliance issues that slipped past fact-checkers: THE QUESTION violated the COLLISION RULE by sharing its only-source-of-THE-LONG-READ as a primary citation, and THE WORLD's ON THE TRAIL subsection collapsed to two flat bullets despite the writer having full access to the NWS per-region forecasts and WTA trip reports needed for the specified format. The pipeline ran cleanly on every other dimension.

Frontpage

The deployed PNG is clean and reads like a credible broadsheet. Visual hierarchy is clear: THE LAB at 52px leads decisively, THE LONG READ alongside it at 40px, then a three-column row 2 with THE QUESTION, THE PELOTON, and FROM THE ARCHIVE. The lead image — the pen-and-ink Skylab ranger with citation pad — is sharp at column width and fits the FROM THE ARCHIVE slot naturally.

One minor defect: the daily strip reads "● Today's Ride — 72 °F and sunny — summer kit, go outside. · summer kit". The kit appears twice — once embedded in the summary string, once as a separately appended field the art director also rendered. The glyph and summary alone are sufficient; the "· summer kit" tag is redundant noise on a line that is already tight at 20px.

The section ordering on the frontpage respects raw priority correctly for the top five sections. FROM THE ARCHIVE (priority 38) sits in row 2 alongside THE PELOTON (70) and THE QUESTION (75), which would be a mismatch — except FROM THE ARCHIVE carries the lead image and THE WORLD (60), which would ordinarily sit above it by priority, is explicitly frontpage_display: "headline_only". The art director correctly compressed THE WORLD to a minimal row-3 column and used FROM THE ARCHIVE's image to anchor row 2 visually. This is defensible.

No duplicated headlines or paragraphs, no clipped text, no broken columns. The gradient fade handles ALSO NOTED overflow correctly.

Priority ranking

SectionPriorityLengthImageNotes
THE LAB86~1,100 wordsFour stories; leads the page correctly
THE LONG READ80~820 wordsSingle source
THE QUESTION75~360 wordsCOLLISION RULE violation (see Editorial)
THE PELOTON70~830 wordsSolid; four story beats
THE WORLD60~175 words (bullets)ON THE TRAIL stripped (see Editorial)
FROM THE ARCHIVE38~440 wordsyesLead image; correct priority cap
ALSO NOTED1010 bulletsWired concentration (see Editorial)
THE FUNNIES7text caption onlyOpenAI image separate from section content

The ranking is defensible. THE LAB at 86 earns its lead: Carmack on id's dissolution is a key-person story on a major gaming event, and three further items (Sol Ultra math proof, colibri, Apple-OpenAI lawsuit) each have independent weight. THE QUESTION at 75 is on the high end for a reflective piece that ultimately depends on two stories already covered in full by other sections — 65–70 would be more accurate — but the cross-domain bridge is genuine. No inflation or compression problem overall.

Editorial reading

ON THE TRAIL: specification compliance failure. The world writer reduced ON THE TRAIL to two flat bullets: "Kendall Katwalk (Snoqualmie Pass) is in great shape with wildflowers blooming. Koppen Mountain (Teanaway) delivers solitude, wildflowers, and an excellent trail at 7.4 miles/2,150 ft." The spec requires for each pick: region name, drive time from Issaquah (from the authoritative table), trip length, per-day mileage AND elevation gain split, a weather quote pulled from the per-region NWS forecast, one sentence on why the pick clears all six criteria, and a link to the supporting WTA trip report. None of these are present for Kendall Katwalk; only partial mileage/elevation appears for Koppen Mountain. This matters beyond formatting: the reader is a backpacker who bases trip decisions on this data. The forecasts were excellent (Snoqualmie: Sat 66°F/3%, Sun 68°F/2%; Teanaway: Sat 75°F/0%, Sun 74°F/0%), making this a no-brainer weekend with two clean picks — all the more reason to give the reader the full specification. The writer had feeds.md with the per-region NWS data and the WTA trip reports in pages/local/ and still produced two summary sentences. The regional snapshot similarly collapses a spec-required 4–6 region-organized bullets into a single run-on sentence. The spec also requires naming the trip window ("this weekend, Sat Jul 12–Sun Jul 13") and checking the six reader criteria explicitly. None of this appeared.

THE QUESTION violates the COLLISION RULE. The rule states: "THE QUESTION may not share primary sources with THE LONG READ on the same day." THE QUESTION cites avi.press/posts/2026-07-10-after-7-years-in-production-scarf-has-reluctantly-moved-away-from-haskell.html as one of its two primary sources — the sole source used by THE LONG READ. The cross-domain Tourmalet/Haskell bridge the writer found is intellectually the strongest angle of the day, but executing it required this collision. The fact-checker for THE QUESTION did not flag it. The correct response under the spec would have been to find a different angle for the distribution-shift thesis — perhaps the Linux kernel CVE (a 15-year-old bug surviving a decade of automated review) as one pole against the Tourmalet terrain, which would have bridged security and sport without touching THE LONG READ's source.

ALSO NOTED: Wired outlet concentration. Five of ten bullets in ALSO NOTED cite Wired as the source: the OpenAI safety head departure, Microsoft carbon emissions, the Linux root bug, Tianwen-2, and the AR glasses architectural argument. For the Linux CVE and Microsoft sustainability report, Wired is a secondary aggregator over more authoritative primary sources (Google's kernelCTF program blog, Microsoft's own FY2025 sustainability report). Using Wired as the source for five items makes half the section feel like a Wired digest rather than the paper's own curation. The section's sourcing breadth matters because it is the only section without a primary beat — homogenizing it to one outlet undermines its purpose.

THE LONG READ: trailing promotional sentences. The article closes: "The piece is worth reading in full. It's the kind of postmortem that's useful precisely because it doesn't come from an outside critic." The first sentence is the weakest possible closing for a curated publication — the reader already knows the piece is worth reading, because the paper ran it. The second sentence is self-referential setup ("not an outside critic") that undermines the paper's own editorial authority. The style guide calls for "direct, unsentimental" prose; these two sentences are neither. The article earns its long-read placement and could have ended two paragraphs earlier at "Haskell is just where it showed up first in a visible way, partly because Haskell's build characteristics make the gap especially wide" — a stronger close.

THE LAB: Sol Ultra source quality. The Cycle Double Cover Conjecture story sources exclusively from cryptobriefing.com, a crypto-focused outlet. The story itself notes "OpenAI published the result as a PDF on July 10" — the PDF itself and Hacker News community discussion would have been stronger primary sources. A crypto outlet is technically a third-party source (satisfying the VENDOR-SOURCE RULE), but for a story about original mathematics at frontier difficulty, sourcing through a crypto aggregator rather than the math community's own response introduces a credibility gap the reader may notice.

Pipeline observations

World writer compresses ON THE TRAIL despite having all required data. The world writer (agent-a69ae3a2067e0295b) read feeds.md at line 15 of its 24-event session — it had access to the full "Trail-area NWS forecasts (per-region, 7-day)" section in that file. The writer also read pages/local/wta-trip-reports.md. Its Done summary confirms it was aware of the picks and the "clean forecasts." The compression to two bullets is a writer-level failure to execute the ON THE TRAIL format spec, not a data availability problem. The researcher, separately, did not include the trail-area NWS section in research.md (the researcher's brief contains only the general Issaquah 7-day forecast under "## WEATHER"), but this is a partial cause at most — the world writer bypassed research.md and read feeds.md directly.

procyclingstats.com blocked for 5+ consecutive days. The race calendar file shows a chain of fallbacks: blocked Jul 7, 8, 9, 10, 11 — five consecutive days where the calendar was served from a prior-day cache. The cache_from_prior: true flag handled this correctly and the race calendar content is stable (upcoming races weeks out don't change day-to-day), but five consecutive blocks suggests procyclingstats.com has deployed persistent bot mitigation. The fallback_search query in extra_sources config is available but was not triggered; a search fallback would eventually be needed if the cache chain extended through a race that changed (a cancellation or date shift). Not a crisis today; worth monitoring.

Comic-strip agent: 1314 seconds for a paragraph of text. The comic-strip agent (agent-ab05265c2bbc27452) ran 35 events over 1314s and produced 159 output tokens — a single italicized text caption describing "After Dilbert" and "After Bloom County" panels. The OpenAI Funnies call (separate, 130s, $0.22) produced the actual comic image as funnies-openai.png. The comic-strip agent's SVG path apparently failed or was never completed; its entire output is a text description. The section-funnies.md accordingly contains only the caption. This runtime ratio — 1314 seconds for a two-sentence caption — suggests repeated failed SVG attempts before the agent gave up and wrote a description. The funnies OpenAI image exists and is presumably the visual component, but the SVG agent's role in this edition was hollow.

Orchestrator cost exceeds all writers combined. The orchestrator ran at $2.95, larger than all six writers + fact-checkers combined ($2.01). Its 6.4M cache-read tokens and 109K cache-1h tokens indicate it accumulated substantial context from subagent outputs. The two next-largest agents by wall time are the Researcher (2195s, $1.98) and the Art Director (2094s, $0.86). The Art Director's 32K output tokens for the frontpage.html is very high for a layout task; it apparently iterated substantially before writing the final HTML.

Agent set is otherwise complete and clean. All expected subagents are accounted for: 1 scout, 1 researcher, 6 writers (THE WORLD, THE PELOTON, THE LAB, THE LONG READ, FROM THE ARCHIVE, THE QUESTION), 1 writer-sweep (ALSO NOTED), 1 comic-strip, 7 fact-checkers (one per non-empty section), 1 meta-writer, 1 art-director, 1 thread-editor. No duplicate agents. No missing agents. Fetch results: 29 items attempted, 0 failures on the primary pass; 3 retried, 1 permanent failure (king5.com for the Kirkland housing story — researcher noted the block and provided summary context directly, which the world writer used). Starting commit was the same-day investigator (5126801d, Jul 10), not a stale worktree.

Trace highlights

Researcher ($1.98, 2195s) costs 5–6x any single writer. The research brief drives a lot of work but the direct relationship between that spend and what writers actually used is uneven. THE WORLD writer ($0.34) drew on research.md but also accessed feeds.md and the raw WTA file directly — the researcher's summary barely mattered to the output. THE LONG READ writer ($0.13, 73s) ran the fastest and cheapest of the long-form sections; it had one source and wrote a clean 800-word article. The cost ratio between Researcher and LONG READ writer is roughly 15:1 on a day when the Long Read's source was already in the researcher's brief as a direct link.

Comic-strip agent: 1314s / 159 output tokens. The cost-to-output ratio on the comic-strip agent (agent-ab05265c2bbc27452) is the most anomalous in the run. 35 events over 22 minutes to produce a two-sentence caption implies repeated failed SVG attempts. At $0.42, it is the sixth-most expensive agent in the run and produced the edition's least substantial output.

Art Director (2094s, 32K output tokens, $0.86). The frontpage.html the art director produced is well-executed, but 32K output tokens is very large for a fixed-canvas layout task. The gradient fades and column widths are correct; the only defect is the daily strip kit duplication. The runtime suggests substantial back-and-forth before committing to the final layout — possibly regenerating the HTML multiple times.

Orchestrator at $2.95 is the edition's single largest cost. The orchestrator's 6.4M cache-read tokens suggest it is accumulating subagent output in its context window rather than summarizing and discarding it. On a day with this many parallel writers, the orchestrator's cost exceeding all creative agents combined is worth tracking as a scaling concern.

Trace summary

Dispatch 2026-07-11 (model: claude-sonnet-4-6)

AgentDurInputOutputCache ReadCache 5mCache 1hCost
Scout333s318650167021572470$ 0.28
Researcher2195s523472635145572287430$ 1.98
THE WORLD244s62696779825190$ 0.34
THE PELOTON391s986133195886070$ 0.37
THE LAB430s9148125012922300$ 0.39
THE LONG READ73s6274856964201180$ 0.13
FROM THE ARCHIVE75s62560121205650$ 0.10
FC: THE LONG READ167s62559238349660$ 0.15
FC: FROM THE ARCHIVE156s74094072265950$ 0.13
Meta-Writer106s952122449257520$ 0.13
Illustrator47s2021372000$ 0.06
FC: THE WORLD238s734146264567520$ 0.26
FC: THE PELOTON330s735132547496510$ 0.23
FC: THE LAB503s1062270790604780$ 0.31
THE QUESTION251s19163397396361150$ 0.17
FC: THE QUESTION171s62867579302230$ 0.13
ALSO NOTED328s843197715729390$ 0.33
Draw today's TWO parody comic strips for1314s16159227396923560$ 0.42
FC: ALSO NOTED248s841160759457400$ 0.22
Funnies (OpenAI)130s4255488000$ 0.22
Art Director2094s1432026119671010270$ 0.86
Update story threads for today's edition419s5187240969430$ 0.37
Orchestrator1502526563622270109504$ 2.95
TOTAL654172530121112881319566109504$10.52

Suggestions for next edition

Add ON THE TRAIL format enforcement to the world writer's fact-checker prompt. The fact-checker for THE WORLD passed the section without flagging that ON THE TRAIL lacked drive times, trip lengths, per-day mileage splits, weather quotes, or criteria explanations. The ON THE TRAIL spec is detailed and testable — the fact-checker should have a checklist for the six required per-pick fields and should reject the section if any are missing, the same way it checks word counts on the world bullets.

Add COLLISION RULE to the fact-checker prompt for THE QUESTION. The fact-checker for THE QUESTION did not catch that it shared its primary Haskell/Scarf source with THE LONG READ. The COLLISION RULE should appear explicitly in the fact-checker's checklist: "confirm THE QUESTION does not cite any source that THE LONG READ also cites." This is a mechanical check, not an editorial one, and belongs in the fact-checker rather than asking the writer to police themselves.

Diversify ALSO NOTED source selection. Five Wired-sourced bullets out of ten is a pattern worth breaking. The sweep writer should be prompted to prefer primary sources (CVE databases, company sustainability reports, project GitHub pages) over aggregator articles when the underlying source is available and accessible. The OpenAI safety head and Microsoft carbon bullets both have primary sources the Wired pieces link to.

Investigate procyclingstats.com block. Five consecutive days of fallback cache use on the race calendar is a signal that the primary source has deployed persistent mitigation. Before Stage 8 becomes Stage 14, consider whether the fallback_search query in extra_sources is configured to run when the cache chain extends beyond 3 days, or if there is an alternative calendar endpoint that doesn't trigger the same block.