Weeknotes 2025-W17 [2025-W17]
- April 24, 2025
- April 25, 2025
- Jon Sterling
Weeknotes 2025-W17 [2025-W17]
- April 24, 2025
- April 25, 2025
- Jon Sterling
It has been a good but busy week. I have been moving more slowly than recently, as I did a tremendous number on my muscles and joints while working in my garden on the weekend. Hope to feel better soon.
1. We have AI at home… [01AS]
- April 25, 2025
- Jon Sterling
1. We have AI at home… [01AS]
- April 25, 2025
- Jon Sterling
On Tuesday, I travelled by train to Sheffield to take part in the Yorkshire and Midlands Category Theory Seminar #37 meeting, where I would be speaking about my paper that compares partial map classifiers with Sierpiński cones in synthetic (domain/category) theory, which I summarised previously.
1.1. A pleasant surprise [01AT]
- April 25, 2025
- Jon Sterling
1.1. A pleasant surprise [01AT]
- April 25, 2025
- Jon Sterling
I was preparing for my chalk talk when I realised that I could not remember the details of the proof of the main result and they couldn’t really be reconstructed from the abbreviated proof in the paper.
Luckily, I had actually formalised this result in Agda! I did not mention my formalisation in the paper because I do not think of formalisations as scientific contributions except in certain cases (that is a conversation for another day). But I did indeed formalise it because the proof was subtle enough that I needed computerised assistance back when I proved it the first time. The result I obtained was frustratingly weak, and seemed to require some annoying side conditions in order to go through; the formalisation helped me be certain that these side conditions were in fact sufficient.
Anyway, I was messing around with the code and what I realised was that I had missed a trick back then: one of the side conditions was actually unnecessary, and it seems kind of likely that the other one is unnecessary too. I am certain I would not have noticed this if I hadn't had the proof assistant, which made it easy for me to try something out and see if it worked. I should have time to update the paper to claim the strong result prior to the LICS camera-ready deadline next month.
1.2. Arise, symbolic AI! [01AU]
- April 25, 2025
- Jon Sterling
1.2. Arise, symbolic AI! [01AU]
- April 25, 2025
- Jon Sterling
There is a lot of discussion lately of the impact that some current machine learning techniques, marketed as “Artificial Intelligence”, can have on formalisation of mathematics in proof assistants. Some of the most esteemed members of the mathematical community have gone all in on this trend (is it a requirement of scientific fame and esteem that you begin to cause trouble in areas of research that you know nothing about?), but I think that evaluating LLMs on Olympiad questions is really missing the point of what computers can do to assist mathematicians. Olympiads are a good fit for LLMs, because kids who participate in Olympiads are behaving much more like LLMs than human mathematicians—the mathematics Olympiad is the ultimate feat of pattern-recognition without understanding, and they are certainly a good fit for the Might Makes Right approach being taken within AI today.
Agda (and Lean and Rocq and Isabelle) are “Artificial Intelligences” in the most progressive sense—they augment the limited context that a human can store in their mind at once, and are nimble tools for working mathematicians to check and verify their ideas, and (most importantly) they do not proceed by creating a fetish of illusion and misdirection that deceives the public. Their capabilities are limited, but well-circumscribed. I think often about how important it is to know in a definite sense what a tool can and cannot do, and I increasingly think that this is actually part of what makes something a tool. Some of my colleagues have compared LLMs to calculators, in order to make the case that we should get ready for them to be used as everyday tools; but LLMs are not simply tools in the sense that a calculator is a tool.
2. Progress on the Forester 5.0 Language Server [01AV]
- April 24, 2025
- April 25, 2025
- Jon Sterling
2. Progress on the Forester 5.0 Language Server [01AV]
- April 24, 2025
- April 25, 2025
- Jon Sterling
Kento Okura has made a lot of progress over the past week in getting Forester’s language server to the point where it can be used. The first editor that we will support is Neovim, which has good LSP support built-in. I think, however, that Kento had not realised quite what a huge amount of work it is to get a working Neovim configuration from scratch that exercises the features of the language server and actually works out-of-the box on other people’s machines. To address this problem, we will be providing a complete working configuration for anyone who wants to use it; experienced users of Neovim will of course prefer to set things up in their own way. Some preliminary code is available, but please stay tuned for further updates that take more advantage of the capabilities of Kento’s language server.
3. Lunch with Patrick Ferris [01AW]
- April 24, 2025
- April 25, 2025
- Jon Sterling
3. Lunch with Patrick Ferris [01AW]
- April 24, 2025
- April 25, 2025
- Jon Sterling
I had a very pleasant lunch in College with Patrick Ferris as my guest; we discussed many things, including the future of Forester and the importance of interop between different authoring tools on the World Wide Web. After lunch, Patrick and I wandered over to Espresso Lane where we had a coffee and a chat with Anil Madhavapeddy and David Allsopp.
Conspiring about the Open Web with my colleagues in the Energy and Environment Group here is making me feel scientifically alive again—there is much to do, and we intend to have fun doing it.
4. De-enshittifying Computer Lab infrastructure [01AX]
- April 24, 2025
- April 25, 2025
- Jon Sterling
4. De-enshittifying Computer Lab infrastructure [01AX]
- April 24, 2025
- April 25, 2025
- Jon Sterling
Not many people are aware that the Computer Lab has an old supplier agreement with Fastmail, which has persisted even after the (ill-advised!) transition to Microsoft Office 365 a few years ago. (There is a certain kind of person whom you can always trust to make poor and irreversible technical decisions, and argue for them on the basis of maintainability or security or liability or all of the above! Whenever you refute their technical arguments, there is always an unbounded source of further reasons why it is mandatory and inevitable that we enshittify our own infrastructure, at great cost of course!) Anyway, the savvier members of the Lab have been rocking Fastmail all this time while I have been suffering the constant outages and inconsistencies of Office 365, which is not only a horrible and unreliable product, but also interacts very poorly with standards-based clients other than Microsoft’s own unusable and bloated clients.
A month or two ago, Anil let me on the secret, and I quickly asked the Lab sysadmins to hook me up with a Fastmail account. The process was not completely trivial, as apparently nobody had asked to use this facility for a very long time. But in the end, Piete and Malcom were extremely helpful and I am now up and running with a Lab Fastmail account! PhD students, postdocs, and faculty are all entitled to use Fastmail if they choose, and I strongly recommend it. If you are one of those people, then you have access to this internal page, which contains the instructions for getting an account.
Although we will never be able to get professional services staff off of Office 365 and Teams (this is the sense in which such moves are irreversible), there is absolutely no reason why we have to use it too. I encourage everyone within the Lab to join me on Fastmail, which is extremely reliable and usable. And the more of us who depend on it, the stronger the insurance against forced enshittification in the future.
Fastmail is just the beginning. With Anil and other members of EEG, I am hoping that we can begin the process of taking back control of our internal infrastructure and making it work for us in the way it used to years before I arrived. I am spooked by recent proposals from University IT to drop talks.cam; it seems to me that taking over the administration and maintanance of such critical infrastructure would be a good fit for our capabilities, and I promise that we can do it for less than the millions that the University would pay a vendor to irrevocably enshittify this infrastructure.