mattwood.fyi

I'm Matt Wood, and this is For Your Information. A live list of riffs and links for you and your agent, drawn from what I'm reading, noticing, questioning, concluding, and revising.

Links indicate relevance, not agreement. How to use this site →

What Sort of Maths Are LLMs Good At?

Explores the mathematical capabilities and limitations of large language models following their recent breakthroughs in major open problems, analyzing whether they excel particularly at finding counterexamples versus proofs.

permalink8 · gowers.wordpress.com →

OpenAI Models Now Available on Amazon Bedrock

OpenAI's latest specialized cybersecurity models (Daybreak Red and Blue) are now available on Amazon Bedrock with enterprise security features, data governance, and pricing that matches OpenAI's first-party rates.

permalink11 · www.aboutamazon.com →
ChatGPT Desktop App for Linux Now in Preview

OpenAI announced a preview release of the ChatGPT desktop application for Linux, enabling users to access ChatGPT, ChatGPT Work, and Codex directly within their development environments and browser workflows on supported Linux systems.

permalink7 · x.com →
Emergent Introspective Awareness in Large Language Models

This paper investigates whether large language models can introspect on their internal states by injecting known concepts into model activations and measuring how this influences self-reported awareness. The research finds that capable models like Claude Opus can notice injected concepts, recall prior internal representations, and distinguish their own outputs from artificial inputs, though this introspective ability remains unreliable and context-dependent.

permalink5 · arxiv.org →
The Economics of Recursive Self-Improvement

A back-of-the-envelope calculation suggests that feedback loops are not currently strong enough to generate a self-sustaining acceleration, though they appear to be strengthening. We conclude by assessing the plausibility and implications of such an acceleration.

permalink8 · elasticity.institute →
Nemotron 3.5 Lightning and NeMo Switchyard for Agentic AI

NVIDIA introduces a lightweight open model and routing library that enables faster, more efficient AI agents with greater control over data and workflows across edge devices, PCs, workstations, data centers, and cloud environments.

permalink9 · blogs.nvidia.com →
Compression is prediction

Explores the fundamental relationship between data compression and predictive modeling, examining how compression algorithms work as predictors and the implications for AI and LLMs.

permalink8 · ngrok.com →
How Keras 3 Modernized Expedia's Lodging Ranking Stack

Expedia Group upgraded their lodging ranking system using Keras 3, improving their machine learning infrastructure for hotel search results.

permalink8 · medium.com →
Brunello Cucinelli: Documentary on Philosophy, AI, and Wealth

A documentary exploring luxury fashion designer Brunello Cucinelli's philosophical approach to business, wealth creation, and the intersection of artificial intelligence with modern commerce.

permalink6 · www.wsj.com →
The Creative Power of Invisibility

Explores how restraint, curation, and subtraction—exemplified by producer Rick Rubin's invisible hand in shaping iconic music—represent the true creative power in an era of algorithmic noise and constant visibility.

permalink10 · www.linkedin.com →
Write.md - Customizable Markdown Editor for macOS

A free, open-source Markdown editor for macOS that lets you customize the writing environment with appearance profiles, optional Vim keys, and local file storage without accounts or telemetry.

I love my growing collection of Markdown editors. This one looks fun.

permalink6 · writemd.app →
MiniMax H3 Inference Engine for Mac

An H3 inference engine implementation for Mac computers, providing MiniMax-based AI inference capabilities.

permalink10 · github.com →
OpenSSH 10.5 Release Notes

Recently the OpenSSH team have received a large number of security bug reports, many of which are findings from AI models or made with AI assistance. While many AI reports are determined not to have security impact when considered in the context of a realistic threat model, we very much welcome these reports, especially when combined with human triage, analysis, test-cases and particularly when accompanied by proposed fixes.

We have seen a number of cases where a security bug identified by AI tools is subsequently independently discovered by a different researcher. This suggests that adversaries who do not report bugs to OSS projects are likely to be able to discover these bugs too. Given this, the OpenSSH team will, for now, be making more frequent releases to get bugfixes into users' hands more quickly rather than batching them until the next planned release.

Makes sense.

permalink8 · www.openssh.org →
Introducing Grok Bot

Grok Bot is an AI agent platform that provides autonomous teammates capable of working 24/7 across apps and tools with their own cloud-based computer. The bots can sign into existing applications, complete multi-step workflows end-to-end, and communicate naturally like colleagues, now available in beta for select subscribers.

All useful assistants are now also computer-use agents.

permalink6 · x.ai →
11–16× Faster LLM Inference with llama.cpp

GPU passthrough in macOS virtual machines, covering how Apple Silicon's architecture handles GPU virtualization, the technical challenges involved, and how the Virtualization framework enables GPU resource sharing between host and guest macOS environments.

permalink14 · github.com →
LiquidAI/LFM2.5-2.6B · Hugging Face

LFM2.5-2.6B is a 2.6 billion parameter language model developed by Liquid AI, available on Hugging Face, featuring instruction-following capabilities with a chat template supporting system prompts and tool use.

permalink8 · huggingface.co →
Cactus Needle 2

Needle 2 is a 45-million-parameter, 14MB open-source language model designed for tool calling, device control, and structured extraction on low-cost edge hardware like microcontrollers, budget phones, and Raspberry Pis. It runs a full session in 28MB of RAM using CQ2-bit compression, achieving competitive performance against much larger small models on mobile device use benchmarks.

Feels like lots is all happening at once with small mobile-friendly models, but this is a space which has been making steady progress for months now. Encouraging.

permalink5 · cactuscompute.com →
What's the best programming language for coding agents?

An analysis and critique of claims that dynamic programming languages are more token-efficient than static languages for LLM coding agents, examining methodological flaws in the studies behind those claims. The piece argues that conclusions drawn from trivial benchmark problems don't generalize to real-world coding tasks.

permalink10 · danluu.com →
Parametron

The parametron, invented by Eiichi Goto in 1954, was a bistable circuit element using ferrite cores that enabled the development of early Japanese computers, including the PC-1. Its invention and application in computing had significant historical impact on computer engineering in Japan, influencing multiple major electronics manufacturers and nurturing a new generation of engineers.

Japan’s first university-built stored-program computer which became the nation's then-fastest in 1958

permalink7 · ethw.org →
Exploring Claude/GPT Knowledge Cutoffs

Methods for inferring hidden details about how large language models like GPT-5 and Claude were trained, including estimating parameter counts, dataset mixtures, and training timelines by probing models with carefully curated questions. Covers the three main stages of frontier model training (pre-training, capability fine-tuning, and post-training) as context for understanding what these probing techniques reveal.

permalink10 · blog.sshh.io →
Introducing Muse Glimmer: An Open Agentic Model That Runs on Your Device

Muse Glimmer is a 30-billion-parameter open-source AI model from Meta Superintelligence Labs, optimized for local agentic workflows including function calling, coding, and LLM-as-a-judge tasks, designed to run on consumer hardware without requiring cloud infrastructure. The model was developed using a combination of logit distillation from a larger teacher model, mid-training on agent-heavy data, and post-training with reinforcement learning, and is released under an Apache 2.0 license.

permalink11 · research.meta.ai →
50,000 boat names

An analysis of over 50,000 boat names extracted from NOAA's AIS vessel traffic data, exploring the creative, humorous, and pop-culture-inspired names that boat owners choose, alongside insights into who owns boats in the United States.

An important dataset, but the all-time best boat name of all time is clearly: Earn Trussed.

permalink6 · www.beautifulpublicdata.com →
How This Was Made

The piece explores how AI adoption momentum within organizations tends to concentrate among early adopters and struggles to spread more broadly, drawing on the example of Benjamin Franklin's Junto club to argue that making the process of learning visible — not just sharing outputs — is key to organizational change.

permalink13 · mattwood.blog →

Humanising LLM Outputs is Dumb

A critique of the trend of "humanising" LLM outputs through prompting techniques, arguing that this approach is the wrong abstraction for addressing verbosity and quirks in AI-generated text.

permalink8 · kuber.studio →

Herdr is joining Y Combinator. The runtime stays open.

Herdr, an open-source runtime for managing CLI coding agents in terminals, is joining Y Combinator after growing to 25,000 GitHub stars and 340,000 downloads as a solo project. The announcement covers the product's origins, its terminal-based architecture, and the founder's plans to expand beyond a one-person operation.

permalink7 · herdr.dev →