(I retweeted this but FYI cannot verify the strict claim “0” and think it is very unlikely. I retweeted it because the culture that considers 0 a possible achievable outcome of putting in the basic work routinely achieves verifiable near-zero outcomes across a range of cases.)
@patio11
-
AI-Generated Images Track Capabilities Progress for Newsletter Headers
By
–
I use this for my newsletter header images (classic commercial work which would otherwise be stock photography or, if I was ever incredibly organized, commissioned digital art) and the archives are a timeline of capabilities progress.
-
LLMs Enable Parallel Construction in Intelligence and Law Enforcement
By
–
I’ve sometimes heard this referred to as “parallel construction” in an intelligence or law enforcement context. LLMs are presently a parallel construction goldmine. Whether that is for good or for ill, well, one of many things we’ll have to adjust to quickly.
-
Using LLMs to Find Public Evidence for Off-Record Journalist Beliefs
By
–
Journalists: if you ever have a thing you understand to be true, but cannot cite it or get it past editors due to commitments made to sources, describe your belief about the world to an LLM and ask the LLM if it can find public evidence which unambiguously confirms the belief.
-
AI Acts as World-Class Research Assistant for Arcane Trivia
By
–
… this is your standard citation.” It’s really weird to have a spiky, tireless research gopher who is literally world class on what even specialists would usually describe as arcane trivia.
-
Claude Opus Analyzes Document via Screenshot After PDF Fetch Fails
By
–
Opus ratholes for a few minutes but can’t successfully locate the document I was alluding to. I find a link in my ancient notes. Opus fails to fetch PDF for technical reasons. I provide screen grab. Opus then *immediately* highlights two sentences and says ~ “Yep, I get why…
-
LLM Opus 4.7 Successfully Surveys AML KYC Law Enforcement Guidelines
By
–
An anecdote for you from LLM land: I sent Opus 4.7 out to do a literature survey of various government law enforcement internal guidelines w/r/t AML and KYC usage in prosecutions. It was much more successful at finding these docs than I would have expected. I then told it:
-
LLM Detects Marketing Patterns Even in Truncated Text Excerpts
By
–
Bingoes a 72 word excerpt about marketing strategy with no firm-specific information included at all, citing a stylistic device. Truncated to 49 words to exclude that device; still bingoes it. Truncated to 8 words; somewhat huffily refuses to speculate.
-
AI Refusals Mask Truth With False Self-Deprecation
By
–
Several more refusals in that genre, and this salaryman is left with the distinct impression that his counterparty has considered telling him the truth, come to the conclusion that that is undesirable, and has substituted (false) self-deprecation regarding one's capabilities.
-
AI Correctly Identifies Redacted Federal Agency Communication From 15 Years Ago
By
–
Correctly bingoes a 3 paragraph communication to a federal agency written 15 years ago where all biographical hints were replaced with [REDACTED FOR PURPOSE OF EVAL].