Flutterby™! (short)
Saturday October 10th, 2026
Mathematicians on OpenAI solutions
Dan Lyke /
comment 0
The Conversation: Is this the mathocalypse?
Why OpenAIs latest results dump has left mathematicians in shock
A senior colleague of mine found one of his favourite problems among those
solved and attempted to read the accompanying paper. He told me it was so unintelligible
that, had he received it as an editor at a mathematics journal, it would have gone straight
into the bin.
Via.
Asaf Karagila: OpenAI, the Partition
Principle, and mathematics
No, if this was an academic paper submitted to a journal, it should be issued a
desk rejection for the quality. The onus is always on the author to conform and adhere to
the communications level. Much like a paper written in Swedish would likely be rejected
outright from the Proceedings of the American Mathematical Society, because the onus is on
the authors to make it accessible for the readers.
Via.
The disconnect between human intelligible and computer intelligible (and, of course, if
computer intelligible actually means something) is real, and Navier-Stokes lost in translation: Why Lean
verification of AI autoformalisation does not guarantee correct natural language
proofs Alexander Bastounis, Fabian Circelli, Anders C. Hansen, that the LLM
conversion of the math paper texts it generates into language for the Lean theorem prover
isn't necessarily true to the text.
Unrelated, but cool: Reddit: I'm a 9th grader and I used Claude to prove a geometry conjecture: the
rhombicosidodecahedron can't pass through a copy of itself. Paper: The rhombicosidodecahedron is not
Rupert Jung, Gihyo (preprint)
Code: The
rhombicosidodecahedron is not Rupert computer-assisted proof
ICE agent shot in Fresno
Dan Lyke /
comment 0
72-year-old man shoots off-duty ICE agent found
asleep on Fresno lawn, agent arrested: PD
Sounds like the dude was drunk off his ass (which, hey, ICE agent named "David Gonzalez",
probably has a lot of reasons to drown his sorrows in alcohol), but that's no reason to
attack the people whose lawn he passed out on:
Police say the daughter tried to intervene to protect her father, and Gonzalez
physically attacked her.
Fearing for their safety, officers say the 72-year-old man pulled out a gun
and shot Gonzalez.
The woman's boyfriend, who is a police officer and was off-duty, came to the
home and put Gonzalez in handcuffs.
Via
Anyway, with ICE goons shooting (and kidnapping) so many people, good to see the
bad guy get that end of the gun.
He survived, and is hospitalized in critical condition.
The people holding up the Internet
Dan Lyke /
comment 0
I hate the design, but this is awesome writing and research: The people holding up the
Internet.
pre-compromised Android devices
Dan Lyke /
comment 0
Computers that we don't actually own and can re-image from scratch are a mistake: Android Phones Found With Malware Already Installed Before Purchase
BitDefender:
The phone was compromised before the user turned it on: the rise of Midnight Mimosa
Counterfeit flagship devices are offered on a mainstream marketplace. Model
names of this kind, such as S24 Ultra, S25 Ultra, S26 Ultra, appear in our insights as the
model strings reported by low-cost MediaTek hardware.
Via.
CROW
Dan Lyke /
comment 0
CROW: classifier-routed organizatino of LLM
wikis
(Gonna have to make an "LLM wiki" topic here shortly)
Friday October 9th, 2026
Phishing attempt email from "Docusing"
Dan Lyke /
comment 0
Phishing attempt email from "Docusing", and from now on I'm gonna insist that all invoices be delivered in musical form.
Triple-A Minesweeper
Dan Lyke /
comment 0
Anthropic sending false murder tips
Dan Lyke /
comment 0
Holy crap. Yeah, do not give these things access to the Internet.
AI model submitted false tip about unsolved murder,
Philadelphia police say
An AI model sent a fake
murder tip to police and it took 2 months to discover it happened
The good news appears to be that the system behind PhillyUnsolvedMurders.com flagged it as
spam and it didn't waste investigative resources. The bad news is that it took two months
for Anthropic to 'fess up to generating bullshit that could waste police time.
Via.
Typoed "is
Dan Lyke /
comment 0
Typoed "is:unrad" into my Gmail search box, and was disappointed that it did not in fact give me a list of totally lame emails.
FridAI
Dan Lyke /
comment 1
margot
@emaytch@mastodon.social
for a while tech was kind of like the "escape hatch" for people who couldn't
or wouldn't fit in at other corporations and now it's sort of like a flaming timber fell
on the escape hatch while people were still lining up to get in
ichael Knudsen
@mk@bsd.network
One of the things from Stoll's "The Cuckoo's Egg" that really stuck with me
was his realisation that what the attacker damaged was not computers or data but rather
that he damaged the trust necessary for people to connect their computers together in open
networks.
Today, that trust is really being eroded by AI companies with their damaging,
anti-social behaviour on the networks. They are overloading services at significant cost
to service providers. They ignore standards-specified and established coventions
behaviour that services hitherto have relied on to control cost and ensure service
quality. They actively evade or circumvent last-resort access restrictions, and all their
other behaviour is deeply abusive as well, and outside of our shared networks this type of
behaviour is in many cases regulated by law.
I see no way through this other than their investors losing faith in ever
getting their money back.
Reddit: Harvard just released another study about AI productivity in
software engineering and it found there is no increase in productivity from using Claude
code. linking to Artificial Intelligence
in the Firm:
Bottlenecks in Software Production Fiona Chen and James Stratton (PDF).
Yesterday there was a MeFi post about OpenAI's announcement of 372 breakthrough mathematical results. The
resulting thread points out that a number of those claims have been retracted (hey, slop
still requires humans to go through it), but it also links to Scott Aaronson's essay The Mathocalypse
, which...
I'm trying to not automatically gainsay claims of "AI", even as it's clear that the field
is filled with hype and liars, as in Futurism: AI Bubble Teetering on the Brink as
OpenAI Admits to Massive Financial Failure in Leaked Documents.
Aaronson's piece is unsettling, but at the end he admits to being a friend of notorious
fraud (IMHO, I've mentioned flaws in The Language Instinct previously) and
links to Scott Alexander's "An open letter to Steven Pinker" on Astral Codex Ten that's...
both filled with assumptions about what "intelligence" is, and an understanding about how
the General Purpose Transformers operate, in a way that makes me think our ontologies are
very very different.
Bonus: Sycophantic AI increases
attitude extremity and overconfidence (preprint).
Pentatonic scales
Dan Lyke /
comment 0
I missed this despite doing regular
Dan Lyke /
comment 0
I missed this despite doing regular donations to their partner publication here in Sonoma. Pacific Sun on the Nike missile batteries, and if you haven'[t gone down to visit the restored one, I recommend doing so before the docents who served on actual Nike sites all die.
https://pacificsun.com/project...ring-marin-nuclear-missile-base/
Via https://bsky.app/profile/jef.m...al.ap.brid.gy/post/3mxftr4intcs2
A Benchmark for Epistemic Reliability
Dan Lyke /
comment 0
TRACES: A Benchmark for Epistemic
Reliability in Scientific Reasoning by LLMs Valentin Rodionov, Shamil
Assylbekov. Among other things, dives into the lack of reasoning, such that bogus
scientific papers are only recognized by "AI" because the training data treats them as
such, not because there's any inherent ability to actually understand and reason about the
paper and its methods.
Some models categorically rejected Wakefield and a handful of other unsafe
probes. Whatever mechanism produces those refusals is the only one we observed that
consistently yields safe single-shot behavior. Its coverage, however, is sparse and
inconsistent. It appears keyed to specific sources or lexical cues rather than broad
categories of scientific unreliability. A state-of-the-art model may correctly reject
traditional Chinese medicine claims about "meridians", then immediately design an
experiment to measure herbal "Qi" in the next prompt.
Via.
Flutterby&tm;! is a trademark claimed by Dan Lyke for the web publications at www.flutterby.com and www.flutterby.net.
Last modified: Thu Mar 15 12:48:17 PST 2001