X PAPERPaper mode
Archive
Menu
—
Newspaper page
THE PERSONAL DAILY Monday, October 5, 2026 Issue 001 / Edited edition
From your following
Edited for reading

X Paper READ

Your following. Your daily paper.

Take time to read
See who said what
76 followed sources · This issue 41 stories Named voices · Traceable sources
IN THIS ISSUE

This expanded edition examines AI promises and acceptance, robot-truck adoption, fiscal execution and employment measures. The reverse covers weight-loss trial populations and safety, laboratory redesign and watermark-detection limits. Markets, public life and culture remain separately attributed, with evidence and limitations.

OCT
05
AI & Technology

What counts as an improvement in Codex’s 28-day pledge?

Tibo (@thsottiaux), whose profile identifies his work as Codex and ChatGPT at OpenAI, promised that each of the next 28 days would bring either a clear improvement relevant to most Codex/Work users or a full reset. Users immediately asked what qualifies and who decides.

The pledge followed a discussion about complexity. On October 4, Tibo described work on simplification, efficiency to support more usage, groundbreaking features and new models. He acknowledged that users wanted things to be simpler. Early on October 5 in Beijing, he quoted that response and added the 28-day commitment.

The condition was a clear improvement relevant to most users, rather than simply a release. Codex is an AI tool for programming tasks; Work is the product term used in his post. Baoyu summarized the pledge as daily releases or a usage reset. Gorden Sun questioned who would judge usefulness, whether one feature might be split across several days, and said he would prefer improvements to usage allowances and models. The pledge did not specify which allowances or accounts a full reset would cover, or when it would happen.

There was already a reset issue in the background. On October 3, Tibo said some subscribers he called “Pro 500” had not received the expected reset and that the team was investigating. He later reported a fix. Those posts establish his account of the investigation and repair; they do not independently confirm the outcome for every account.

The disagreement concerns whether a release translates into a practical benefit. Announcements can appear every day, while restored allowances and resolved problems require checking actual results. This issue records the pledge, responses and Tibo’s reported fix. The 28-day outcome is not yet available.

Original posts and supporting context reviewed

AI & Technology

Claude Mods bring context, tool calls and next steps into the interface

Boris Cherny introduced Mods as a way to change Claude Code’s interface and behavior: users can add information, intervene before tools run and share their modifications as plugins.

An ordinary skill primarily supplies instructions to the model, while external tools supply capabilities. Mods run JavaScript or TypeScript handlers inside Claude Code. When a prompt is submitted, a tool is about to run or the interface is drawn, a handler can observe, modify or take over the event. That allows both new panels and changes to how an operation proceeds.

The official blast-radius example pauses commands such as file deletion or force pushes, shows their expected impact and offers buttons to proceed or cancel. token-weather displays context-window usage above the input: how much of the information the model can accommodate has been used. replay-theater adds /replay to step through the previous turn’s file edits. ClaudeDevs adds that token-weather plots the past twelve turns, making the buildup of context visible rather than showing only the current total.

Thariq’s next-steps example suggests skills and commands to use next. ClaudeDevs’ You should Know highlights information that might be overlooked in output. These examples address what to do, what to notice and what to inspect before acting, placing reminders inside the workflow.

Users can describe a modification to Claude and have it create a Mod, or install an existing plugin. The documentation supports interface extensions in the terminal and the Code tab of Claude Desktop; support differs across other environments. Mods execute code with the user’s permissions, making their authors and sources relevant to installation choices. This report describes published capabilities and examples, not our own product testing.

Original posts and supporting context reviewed

Biomedicine

How protein watermarks leave a trace of AI design

DeepMind’s SynthID Bio attempts to embed provenance signals in protein sequences or predicted three-dimensional structures. The work highlighted by Hassabis and Pichai is a proof of concept that preserves biological function in tested designs, not a universal detector of AI-generated biology.

A watermark is more than a label attached to a file. For sequences, the method subtly guides amino-acid selection; for predicted structures, it adjusts atomic coordinates to carry a detectable signal. The sequence approach aims to retain a verifiable provenance trace after a design is synthesized into a physical protein.

The team used AlphaProteo and a watermarked version of ProteinMPNN to design proteins that bind to particular targets. Across VEGF-A, the SARS-CoV-2 spike receptor-binding domain and PD-L1, it reported comparable hit rates, binding strength and sequence diversity for watermarked and unwatermarked designs. Testing binding to those targets does not establish that every possible function is unaffected.

For structure prediction, the team fine-tuned part of AlphaFold 3’s diffusion network so that output coordinates carry a signal. It reported preserved prediction accuracy and detection that withstands digital noise or small coordinate changes. Resistance to deliberate tampering remains a research challenge.

The proposed use is an additional provenance signal for synthesis screening and biological databases. An unfamiliar sequence could be a natural discovery or an AI design; mislabeled synthetic structures could also affect later research. Watermarks help distinguish origins but cannot alone establish whether a design is dangerous. Adoption and stronger resistance to tampering remain future work.

The DeepMind podcast explains further limits: very short, fixed-answer text may offer no room for a watermark. Detecting another system’s output requires its key. Biological deployment also needs coordination between model providers and synthesis companies. A missing watermark does not prove natural origin or safety. This addition follows a complete transcript reading, not an independent audio review.

Original posts and supporting context reviewed

AI & Technology

AI can build the tools; creative judgment still needs practice

Thariq described using Claude to find references, teach him and build an animation editor for repeated changes to a game character’s jump. After the update, he said he had reached the limits of his own judgment.

Introducing this personal project the previous day, he had argued that making a game was satisfying in itself: use AI to realize a vision together rather than outsource the entire creative process. For the animation update, he asked for an environment in which he could keep experimenting, rather than a finished result in one pass.

A reply in the same thread adds a qualification to the showcase. He believed there were still many problems and opportunities for improvement, but lacked the skill to evaluate them further. He said he needed to develop better taste.

The account separates building an editor and running more experiments from judging which movement works better. When a single generation falls short, creating an environment for comparison, revision and practice is one possible working method. More capacity to experiment does not automatically supply a standard of quality.

Original posts and supporting context reviewed

Finance & Business

US vehicles averaged 12.8 years in 2025: durability or replacement cost?

Charlie Bilello revisited the aging US vehicle fleet. The underlying 12.8-year figure is the 2025 combined average for passenger cars and light trucks. Passenger cars averaged 14.5 years; light trucks, 11.9.

Bilello offered two explanations: vehicles last longer and replacement has become more expensive. Both could operate at once, but they imply different consumer circumstances: an existing product that still meets needs, or a replacement cost that delays a purchase.

The Bureau of Transportation Statistics identifies S&P Global Mobility as the source for recent years. Its 2025 analysis confirms the combined average and the differences between vehicle types. The aggregate should not be treated as an identical experience for all owners, nor described as a newly established 2026 record.

The chart documents changing vehicle ages. Durability and higher replacement costs are Bilello’s explanations; it does not measure each factor’s contribution or directly establish the earnings outlook for a particular carmaker or repair business.

Original posts and supporting context reviewed

Politics & World

The Supreme Court enters a new term amid persistently low public ratings

Ian Bremmer drew attention to declining trust in the Supreme Court. Gallup’s September survey reports 46% trust in the federal judiciary and 34% approval of the Court’s work. They measure different things.

Gallup conducted its annual Governance poll on September 1–17 and published the results on September 30, before Bremmer’s post. The 46% combines people expressing a great deal or a fair amount of trust in the federal judicial branch headed by the Supreme Court. The 34% measures approval of the Court’s job performance.

Job approval was essentially unchanged from July’s 33%, with 61% disapproving. Judicial trust was below half for a fifth consecutive year, slightly below the 47%–49% range in 2022–2025. This indicates continuing low ratings; 34% is not a freshly broken record low.

Party differences are much larger than the movement in the overall figure. Judicial trust was 23% among Democrats, 42% among independents and 79% among Republicans. Job approval was 12%, 31% and 65%, respectively. The new term begins with sharply divided public assessments.

Gallup discusses the longer decline alongside the Court’s conservative majority and the decision overturning federal constitutional protection for abortion. The survey does not isolate the causal contribution of each ruling and is not an instant reaction to events on the day of Bremmer’s post.

Original posts and supporting context reviewed

AI & Technology

Guizang shows a Muse firmware interface; connected functions await a demo

On October 3, Guizang described the Muse Gadgets SDK for connecting ESP32 devices. Two days later, he said he had flashed firmware onto an M5 Stack Stop Watch Meta and posted a photograph of its circular display.

An SDK is a developer’s software toolkit; firmware is the program installed on a device to make it operate. Guizang said the toolkit includes firmware for several common ESP32 devices, avoiding the need to write everything from scratch. His earlier post described connecting small devices to Muse, retrieving information and building personal accessories.

The photograph shows a pixel character and a SET UP WI-FI prompt. It demonstrates a running interface, but does not show information updates after connection, voice functions or two-way interaction. The documented milestone is firmware installation and an interface display; the connected functions still need a demonstration.

Progress report; connected functions not yet demonstrated

Arts & Life

Records and touring: this week’s music announcements

The Beatles announced a Rubber Soul special edition; Red Hot Chili Peppers announced limited vinyl for Stadium Arcadium’s twentieth anniversary. Muse previewed a return to Europe in January 2027 and opened artist-presale registration.

The Beatles’ October 2 post announced that the Rubber Soul special edition was available. The Red Hot Chili Peppers announcement concerns limited anniversary vinyl, with the band’s website as the purchase destination.

Muse’s announcement concerns January 2027 European performances and provides an artist-presale registration link. Registration is a step before buying, not confirmation of a ticket. Specific dates and sales details remain subject to the band’s announcements.

Original posts and supporting context reviewed

AI & Technology

Pichai reports a prototype satellite launch for space computing

Pichai recalled the initial proposal to move computing resources into space and reported a successful prototype satellite launch with Planet on October 2.

He thanked SpaceX and emphasized both the progress made and the distance still to travel.

His post describes a prototype milestone; it gives no computing capacity, operating cost or date for a public service.

Announcement / attributed statement

AI & Technology

An old PSP becomes a voice interface for Muse

Wang shared Robert Soriano’s account of porting the Muse gadget client to C for a PSP that had been unused for more than fifteen years.

Soriano described holding R to speak into the built-in microphone, with Muse responding and OpenAI supplying the voice. He used his agent during the port.

Announcement / attributed statement

AI & Technology

Thariq asks company leaders about AI coding challenges

Thariq, who usually writes for developers, is preparing an article for company leaders dealing with changes brought by AI coding agents.

He asked what readers want their leadership to understand about agents and what problems they encounter when running a company.

Announcement / attributed statement

AI & Technology

Claude.dev brings engineering articles and developer guides together

ClaudeDevs introduced Claude.dev on October 1 as a new destination for developers building with Claude.

The announcement lists engineering deep dives, Claude Code and API guides, and tips from the teams building Claude.

Announcement / attributed statement

AI & Technology

Guizang considers open-sourcing a foldable-screen screensaver

Guizang said he made an Android foldable-screen screensaver inspired by the half-folded iPhone Duo display.

He had not released it because he expected few users, but offered to open-source it if there was interest.

Announcement / attributed statement

AI & Technology

Tibo reports his first inbox zero after assigning the task to dot

On October 4, Tibo said he reached inbox zero for the first time after setting an active email-clearing goal for his dot.

He quoted a post from two days earlier in which unread mail had fallen from more than 9,000 messages to 6,110, with a goal of clearing the inbox over the next 48 hours. He also said he wanted to see how long inbox zero would last.

Announcement / attributed statement

AI & Technology

Altman voices concern over surrendering judgment to AI

On October 3, Sam Altman expressed concern about people attributing religious power to AI models or relinquishing human judgment.

Announcement / attributed statement

AI & Technology

Altman responds to speculation about Cerebras

On October 3, Altman called Cerebras a close partner of OpenAI and said the companies were working together at the frontier of speed.

The response did not disclose contract size, computing purchases, duration or a specific product release.

Announcement / attributed statement

AI & Technology

OpenAI and ASBDC announce AI guidance for small businesses

On October 1, OpenAI introduced a report on small businesses using AI agents and announced practical training and local guidance with ASBDC.

Its post listed finding customers, building products and managing finances as tasks small teams were taking on with AI.

Announcement / attributed statement

Politics & World

Cato announces a $50,000 essay prize and submission date

Cato Institute announced a $50,000 prize for the best libertarian essay, with submissions opening on March 1, 2027.

Its post says anyone may apply and provides registration for updates.

Registering for updates is separate from submitting an essay. Length, judging and submission requirements remain subject to the full competition rules.

Announcement / attributed statement

Politics & World

Obama shares a voting-rights volunteer call for lawyers

On October 2, Obama called attention to We The Action’s recruitment of lawyers willing to volunteer their time to defend voting rights.

He described lawyers as being on the front lines of protecting democratic institutions and linked to wetheaction.us/guide.

Announcement / attributed statement

Arts & Life

OneRepublic announces December Jingle Ball appearances

On October 3, OneRepublic announced appearances at this December’s iHeart Jingle Ball and said tickets were on sale.

The short post does not list individual cities, dates or ticket prices.

Announcement / attributed statement

Finance & Business

High indexes, neutral positioning: two different readings

Liz Ann Sonders’s latest podcast examines a divided market. Vanda’s Eric Liu has turned bullish on U.S. equities; that does not mean every stock is strengthening.

Liu says U.S. equity positioning is near its historical average in Vanda’s data since 2010, while bond longs have unwound. He favors large-cap technology. Positioning measures existing exposure, rather than index prices. His full model data are not disclosed.He also cites stronger equity rebounds when yields fall during recent conflict episodes. That small historical sample is not a rule for future returns.

Breadth asks how many stocks participate, using measures such as advancing versus declining shares or the percentage above moving averages. A few heavily weighted stocks can lift an index while most lag, increasing dependence on those leaders.

Narrow breadth and neutral positioning can coexist. They ask different questions: how crowded exposure is, and how widely gains spread. An index high answers neither.

Podcast passages and breadth definitions reviewed; positioning model not independently validated

Biomedicine

A recovery score is still a long way from a health verdict

Eric Topol reshared his September essay questioning wearable HRV and readiness scores. Recording a signal, interpreting it and improving health are separate steps.

HRV describes variation between heartbeats. ECG records electrical activity; most consumer wearables infer variability from optical pulse signals. Topol argues that associations between low HRV and disease do not prove that raising it improves health. Readiness scores still lack validation against health outcomes.

A cited 2025 study measured 931 people at one medical institution for 5–7 seated minutes. The same upper-arm monitor recorded mean optical RMSSD of 37.49 milliseconds versus ECG’s 43.14. RMSSD measures successive normal beat-interval differences. These readings were not interchangeable.

The study did not test Apple Watch or follow health outcomes. Four authors worked for device maker Tiger Tech. It establishes a measurement concern, not an error estimate for every watch.

Topol challenges health claims, rather than all wearable functions. Scores across brands cannot be directly compared; a low score alone cannot establish deteriorating health.

Essay and study methods, tables, limitations and conflicts reviewed

Politics & World

A return-home deal has not ended Spain’s housing protests

An evicted 87-year-old tenant has a low-rent return agreement. Wider housing protections were rejected, leaving protesters’ demands unresolved.

Maricarmen Abascal was hospitalized after eviction from her longtime Madrid home. Her lawyer told AP on September 29 that a return agreement capped rent at 30% of income, no more than the previous €500 monthly. Actual return has not been confirmed here.

BBC reported on October 4 that parliament rejected two housing decrees the preceding Friday. Proposals included extending protection against eviction for vulnerable tenants until 2030, lease extensions and automatic renewal. These were proposals, not implemented protections.

Around 50 marches followed on Saturday. BBC cited the Bank of Spain’s estimate of a 700,000-home gap between demand and new construction. Resolving one tenant’s case has not settled the wider protection and supply disputes.

BBC report and AP lawyer interview checked; actual return unconfirmed

Finance & Business

A cattle collar turns the fence into software

Gorden Sun highlighted Halter’s smart cattle collars. The underlying announcements clarify the chronology: its $220 million funding round was in March, and a satellite-connected product followed in April.

GPS locates the herd; sound and gentle vibration guide cattle within virtual boundaries and between pastures. Solar-powered collars and a phone map let ranchers change grazing areas without moving wire. Later tools add behavior monitoring, heat detection and forage information.

Gorden described the tower-based system. On April 28, Halter announced direct-to-satellite collars without on-ranch communications towers, initially for U.S. and New Zealand beef operations, with Australia and Canada to follow. That announcement does not establish how many ranches now use that version.

The March 25 Series E, led by Founders Fund, valued Halter at $2 billion. The company reported over 2,000 ranch and farm customers and one million collars sold. Sales are not a count of cattle simultaneously online. Funding was earmarked for field operations, health tools and expansion.

This is an agricultural digitization case resurfacing in October, rather than a new funding event. Terrain, deployment and herd training still matter; the funding and company examples do not establish a general payback period.

Original post and supporting sources read

AI & Technology

Same weights, 62% in one harness and 33% in another

Hugging Face’s multi-harness experiment put identical LFM2.5-2.6B weights at 62.1% success in Mini-SWE-Agent and 33.2% in Claude Code. The test comprised 250 held-out data-analysis tasks, rather than a general coding leaderboard.

A harness governs context, tools, retries and stopping. Keeping weights fixed does not keep the whole agent fixed. A capture proxy records exact generated tokens and probabilities; OpenEnv, Harbor and TRL connect environments, tasks and training.

Training across OpenCode, Claude Code, Codex and Mini-SWE-Agent lifted average success from 42.2% to 54.2%. OpenCode-only training averaged 52.3% and reached 58% within OpenCode; mixed training did better in Claude Code and Codex. The authors put the 1.9-point overall gap within noise.

The reported 31.1% reduction in tool calls applies to tasks both baseline and trained models solved, not all tasks or bills. Each run had one seed, and mixed training saw more distinct tasks and processed more tokens. There was no LFM run without the efficiency bonus to isolate its effect.

The practical lesson is to evaluate the model in its actual harness. Open code and models permit further testing; these results do not establish universal superiority at equal compute.

Original post and supporting sources read

AI & Technology

Correct calculations are not yet worthwhile science

In an Anthropic guest essay, Harvard physicist Matthew Schwartz describes using his BootLoops project to produce 36 research manuscripts across 18 fields with 19 collaborators in three months. Manuscripts are not a count of accepted, peer-reviewed papers.

He selected from roughly 400 candidate directions, assigning coding, calculation and cross-field methods to Claude while humans chose questions. Parallel sessions, evolving plans and adversarial reviewers supported the work. Schwartz is also an Anthropic visiting researcher; this is a participant’s account.

One ecology result found species-composition change in a Panamanian forest about 4.5 times faster than neutral theory predicted. Ecologist James O’Dwyer noted that the qualitative finding was already known. The project shifted toward subtracting the random prediction and explaining the biological remainder, then extending a demographic model to other plots.

A genetics integral likewise needed a more meaningful biological question. Schwartz describes premature completion claims and unproved lemmas silently treated as axioms. Expert feedback, inspecting plots and choosing a better problem mattered more than endless retries.

AI can widen computational exploration; domain judgment still determines scientific value. We read the essay, rather than reproducing its 36 manuscripts. It provides no uniform compute cost or independently verified research return.

Original post and supporting sources read

AI & Technology

As AI does more, how do people understand it?

Karpathy argues that more autonomous models leave people spending more effort on supervision and understanding. He has been trying controlled language, diagrams, interactive pages and custom explanatory videos.

For prose he asks for ASD-STE100, simplified technical English originally developed for aerospace maintenance manuals, often at roughly 80% strictness. He wants less ambiguity and excess wording; this is a personal prompting practice, not an accuracy experiment.

Images help show relationships. HTML pages let readers manipulate an explanation. For complex topics he experiments with 3Blue1Brown-style videos and narration. The formats engage sight, interaction and hearing, with different uses.

Previously, custom software for a single explanation was expensive. Fast generation can make such temporary tools practical. Karpathy offers personal experience rather than a controlled learning study; easier presentation still leaves the facts in each diagram, page or video to check.

Full original post read

Finance & Business

September adds 29,000 jobs; revisions change the picture

FXTrader highlighted September’s U.S. payroll gain of 29,000 and unemployment rising from 4.1% to 4.2%. The BLS October 2 release also revised July and August payroll gains down by a combined 60,000.

July changed from a 21,000 gain to a 10,000 loss; August fell from 162,000 to 133,000. Additional employer reports and recalculated seasonal factors drive revisions. September is preliminary. Gains had averaged 45,000 over the previous twelve months.

Unemployment comes from the household survey; payroll jobs come from the establishment survey. In September the household labor force rose by 485,000, employment by 406,000 and unemployment by 78,000. Participation moved from 61.6% to 61.8%. People entering the labor force can lift unemployment even while employment grows; subtracting the two surveys’ figures is misleading.

Private hourly earnings rose 0.1% monthly and 3.0% annually. Weak payroll growth, downward revisions and labor-force entry coexist. BLS notes unemployment has stayed within 4.1–4.3% since March. One month does not determine a recession or the next interest-rate decision.

Original post and supporting sources read

Finance & Business

More ships can still mean less usable capacity

An AEI-shared working paper, Shipping to America, finds utilization losses of 20–40 percentage points during disruptions, preceding visible port congestion. Its utilization measure is not container load fullness.

Five authors matched AIS position and speed records to vessel characteristics for January 2016–March 2025. They recorded 58,582 U.S.-bound trips across six main routes covering roughly 80–90% of containerized maritime imports. Effective ton-mile capacity subtracts waiting, canal delays, rerouting and slowdowns from the fleet’s potential.

By early 2022 deployed potential capacity exceeded pre-pandemic levels by over 60%, yet aggregate utilization fell from about 89% to below 70%. Ships waiting or taking detours do not translate directly into deliveries; shifting ports can shift congestion.

A calibrated model of shared fleets, congestion and sourcing estimates welfare losses equivalent to 0.69% of output for the 2021 West Coast crisis and 0.35% for Red Sea attacks. These are consumer-income-equivalent counterfactuals, not measured GDP declines.

Tariff decongestion effects depend on conditions, rather than proving tariffs always help. AIS cannot identify cargo origins and destinations precisely. We checked methods, relevant charts and policy sections, without reproducing the model or auditing every mathematical appendix.

Original post and supporting sources read

AI & Technology

Claude’s half-usage offer has important boundaries

An artifact promotion covers Pro, Max and Team through October 15 at 11:59 p.m. Pacific, excluding Free and Enterprise.

After creating or editing an artifact in a regular chat, the next ten messages count 50% less toward the five-hour limit, for the first fifteen steps per reply. Starting via Document, Presentation or Design in Output discounts the first message. Weekly limits and usage-credit prices stay unchanged; Claude Code, API and local Cowork are excluded. Cloud Cowork has separate rules: roughly the first 45 minutes or 80 steps after artifact creation.

Original post and supporting sources read

Finance & Business

ChatGPT finances reaches U.S. Free and Go users

ChatGPT announced a rollout to U.S. Free and Go users, connecting accounts through Plaid and Experian. Its follow-up lists forgotten subscriptions, unfamiliar or potentially duplicate charges, rising bills, budgets and debt-repayment plans.

Weekly updates and portfolio concentration checks are also described. These are account-analysis capabilities, rather than an announcement of executing trades. We did not connect an account to test them; a U.S. rollout does not establish availability in every region or plan.

Original post and supporting sources read

Finance & Business

Unusual Whales adds a ChatGPT data entry point

On October 3, Unusual Whales announced a ChatGPT plugin for its market trading data and analysis, linking to the integration. The announcement does not itemize coverage, latency, subscription requirements or permissions. We did not install it; it does not establish that every dataset is free or live.

Full original post read

Politics & World

Tariff poll: 48% see manufacturing harm, 36% a benefit

Cato’s October 4 post cites 48% of its “working-class” respondents saying tariffs harmed U.S. manufacturing, 36% helped and 16% had no effect. The article uses registered voters without college degrees for this group, rather than all workers defined by occupation or income.

The September 24 article draws on Cato/Morning Consult total samples of 2,002 in March and 4,150 in August, not subgroup counts. These are perceptions, not measured manufacturing losses. We did not verify the full questionnaire or subgroup error, and do not project election results.

Original post and supporting sources read

Arts & Life

Aichi–Nagoya Games close; China takes 169 golds

CCTV reported on October 5 that the Aichi–Nagoya Asian Games closed the previous Sunday. China finished with 169 gold, 89 silver and 83 bronze medals, totaling 341, topping the gold table for a twelfth consecutive edition. CCTV calls this China’s best overseas Games performance. This brief relies on its full closing post; we have not audited individual official results.

Full original post read

AI & Technology

Hugging Face chat brings MCP connections back

Hugging Face announced MCP connections for bringing users’ data into its ML Intern chat. MCP is an interface protocol linking AI applications with external tools and data. The post does not detail every connector, permission or price; we did not test the demonstration. This brief reports the announced integration.

Full original post read

AI & Technology

Argon extends output, with cyber defenders first in line

Pichai, Hassabis and DeepMind announced Gemini 4 Argon for coding, enterprise work and cyber defense. Its first cohort consists of trusted Fairwind testers; broader access is still ahead.

The output ceiling rises from 64,000 to one million tokens. This concerns long generated trajectories, rather than a new input-context limit. Feedback from early users is intended to strengthen safeguards before the wider rollout.

Google reports internal agents freeing over 300 TiB of memory and migrating C/C++ projects toward Rust, including kernel work. Those rewrites still undergo automated testing and human audits. Longer autonomous work retains an acceptance step.

The announced 77.9% on DeepSWE v1.1 measures long-horizon software engineering tasks. These are Google-reported evaluations, not an independent replication or a success rate for every repository.

The planned wider release begins with paid API customers and Google AI Ultra subscribers. Announced introductory prices are $2/$10 per million input/output tokens, then $4/$20, with a 95% cached-input discount. Neither pricing nor the announcement establishes general availability.

Post and cited sources reviewed in this batch

Finance & Business

Faster fiscal spending—and what Hong still expects next

Hao Hong read Lan Fo’an’s October 1 fiscal-policy essay as a reason to expect further measures after the holiday. The ministry text separates implementing existing measures from studying additional ones.

It calls for accelerating slow spending, coordinating ultra-long special treasury and local special bonds, and expanding interest subsidies. Budget authorization and money reaching projects remain distinct stages.

The RMB300 billion special treasury bonds support core tier-one capital at designated central financial institutions. This strengthens their capital base; it is not a household cash transfer.

Using remaining local-government debt limits is described as a measure under study. The essay gives no final amount, allocation or implementation date for that item. Unused borrowing headroom does not establish that money has been disbursed.

Hong’s expectation of post-holiday action is his forecast. The essay sets policy tasks rather than predicting market returns. Follow-up evidence would be formal measures, issuance and actual expenditure.

Post and cited sources reviewed in this batch

Finance & Business

Robot trucks show why capability is not immediate job loss

AEI shared James Pethokoukis’s argument about the pace of AI adoption. Trucking supplies a case about commercial friction, rather than proof that automation will never displace work.

Against earlier predictions of driver displacement, he considers Aurora’s goals of 200 driverless trucks by the end of 2026 and more than 30,000 by 2030. Those remain company targets, not today’s fleet.

He cites a bank analysis putting the latter scale at roughly 1.9% of Aurora’s addressable long-haul miles. We did not obtain that report; the figure remains attributed to his account.

Manufacturing, insurance, maintenance, utilization and integration into carrier networks shape deployment. Aurora’s own release also identifies production and customer agreements as forward-looking dependencies.

The implication is to specify adoption and costs before estimating job losses. Software could spread faster; this trucking example does not measure every occupation.

Post and cited sources reviewed in this batch

Biomedicine

Petrelintide: keep Week 28 separate from Week 42

FierceBiotech reported phase 2 results for Roche and Zealand’s investigational amylin analog. The middle-dose result and the previously promoted later endpoint answer different questions.

The report puts Week 28 mean weight loss at 9.8% with 5 mg versus 1.7% with placebo: an 8.1-percentage-point difference. The 7 mg and 9 mg groups reached 9.3% and 9.4%; higher doses did not produce greater loss in this analysis.

Zealand describes a randomized, blinded dose-finding study alongside diet and activity measures. Week 28 was the primary endpoint. Its release emphasizes up to 10.7% at Week 42 under the efficacy estimand, versus up to 10.2% under a policy estimand handling discontinuation differently.

FierceBiotech reports nausea in 20% of treated participants versus 6% on placebo. Zealand’s favorable vomiting and gastrointestinal-discontinuation statement concerns the maximally effective dose, not every dose pooled.

Phase 3 has begun, but this remains investigational. We read the release and reporting, not the full paper. These results cannot establish long-term benefit or rank drugs across different trials.

Post and cited sources reviewed in this batch

Biomedicine

Retatrutide: the population behind 20.8%—and adverse events

Lilly’s detailed TRIUMPH-2 results concern adults with type 2 diabetes and obesity or overweight. Mean loss at 80 weeks was 20.8% with 12 mg versus 4.0% with placebo.

The trial randomized 1,152 participants across three doses and placebo. The weekly drug activates GIP, GLP-1 and glucagon receptors. The highlighted analysis assumes continued intervention without prohibited weight-management treatment; it is not a guaranteed real-world outcome.

At 12 mg, 34.9% lost at least a quarter of their weight, versus 0.8% on placebo. A1C fell 1.5 percentage points from a 7.7% baseline, versus 0.2 points. Weight and glucose control remain separate endpoints.

High-dose diarrhea, nausea and vomiting rates were 33.6%, 28.0% and 15.7%. Adverse-event discontinuation was 7.7%, versus 4.9% on placebo; Lilly describes most events as mild or moderate.

Lilly plans a US filing in early 2027; approval is pending. We read the release and reporting, not the full paper. Comparisons across populations, durations and estimands cannot establish head-to-head superiority.

Post and cited sources reviewed in this batch

Biomedicine

Genentech’s AI redesign reaches the laboratory bench

FierceBiotech interviewed CEO Ashley Magargee about automation, restructuring and campus plans. She presents AI as a tool for scientists; the interview also describes senior research departures.

Testing agents on a BCA protein assay with Anthropic raised a concrete layout question: must equipment remain on horizontal benches designed around human access? Magargee says the team is examining vertical arrangements. No measured discovery-speed or drug-success improvement is supplied.

Her response to gRED restructuring and scientist exits is that scientific identity remains intact. That is management’s explanation, not proof that staffing changes have no effect.

A centralized research building is planned for 2027–2033. The project sits within Roche’s previously announced $50 billion US commitment, which she acknowledges includes existing plans as well as new investment. The total is neither wholly incremental spending nor completed construction.

Post and cited sources reviewed in this batch