|
← 判断力与美学 ← Judgment & Aesthetics
SAE 判断力与美学 · 余项之美
SAE Judgment & Aesthetics · Beauty of the Remainder
2026-08-03

诱饵字体:写给两种目光的文字

Decoy Font: Writing for Two Kinds of Eyes

Han Qin (秦汉)

Eric Lu 在字体铸造厂与 AI 研究室 Mixfont 设计的 Decoy Font,是一款利用混合图像视错觉构造的字体。每个字母叠合两层形象:前景是清晰的细轮廓,背景是模糊的粗色块。人眼通常先捕捉到背景中的暗色字形——通过虚焦、斜视或改变与屏幕的距离,可以在两者之间切换读取。但当代大型语言模型读取图像时,行为相反:它们按像素近距扫描,自动抓取前景中更清晰的高频轮廓。结果是:同一张图,人读到"早点睡觉",AI 读到"追剧到深夜"。同一字形,两个接收者,两则互不相知的消息。

这个感知裂缝——人眼与机器视觉之间的结构性分歧——正是 Decoy Font 的余项所在。余项不是作品本身,而是作品所暴露的那个尚未被任何体系完全消化的空隙。混合图像的视觉科学早已存在,最著名的是玛丽莲·梦露与爱因斯坦叠合的混合肖像:近看是爱因斯坦,退远是梦露,视觉系统根据观看距离自动切换读取层。这一物理机制是已构——被视知觉科学研究了几十年,原理清晰,已有完整名字。但将空间频率差用于同时书写给人与机器的双重文本,把两套感知体制变成两个读者,这一逻辑尚未被命名,尚未成为艺术史或设计史的凿构结果。它仍然是活的余项,而非已固化的构。

在 SAE 美学框架中,凿构循环(chisel-construct cycle)描述的是:每一次真正的凿出都在产生新的余项,而那些余项会被后续的时代消化为已构。Decoy Font 的美,正在于它站在这个循环的开口处。"反 AI 艺术"作为类别已经开始结晶,IMAGO 用本地 AI 模型替代云端数据抓取,Ghost Font 用运动动画隐藏消息,这些作品共同构成一个尚未完成命名的方向。但 Decoy Font 的具体机制——以视觉感知的物理频率差为素材,将两套读者嵌入同一字形——还没有一个完整的概念外壳。它还在生长,还在等待被消化。这正是余项的形态。

Lu 自己的描述点明了这件作品的时间性。他说:也许几个月后 AI 就能读破它。这句话不是悲观的预测,而是余项之美的精确描述。余项存在于凿子尚未落定的瞬间,存在于感知裂缝尚未被技术或理论封合的窗口期。今天,LLM 以像素为单位读图,尚未学会人类视觉的层次感知——它靠近看,只看到高频轮廓,错过了低频的整体字形。这个"错过"是真实的,是物理的,是计算架构决定的。当多模态模型学会像人眼一样虚焦时,这套双重读写的逻辑就失效了,余项消失,成为已构——一个"曾经有效的把戏",被收录进 AI 发展史的某个脚注。

现在看它,你看到的是两种视觉体制之间的真实裂缝,还没有被任何一方完全吞并。Decoy Font 不是艺术宣言,也不是技术论文——它是一个可以下载的 TTF 字体文件,和一个让你编码自己双重消息的网页工具,由一个字体铸造厂的 AI 研究者作为实验公开发布。它站在字体设计、机器感知研究和对 LLM 的静默抵抗之间,尚未被任何学科完全认领。这种悬置——在诸多话语之间、在命名之前——是余项之美最常见的栖居之处。

mixfont.com ↗

Decoy Font, made by Eric Lu at Mixfont—a type foundry and AI research lab—is a typeface that exploits the hybrid image optical illusion. Each character overlays two letterforms: a blurry low-spatial-frequency shape in the background, and a crisp high-spatial-frequency outline in the foreground. Human eyes, which read in context and shift focus naturally, tend to land on the fuzzy background forms; a squint or a step back from the screen lets you slide between the two. Large language models, which process images by reading pixels at close range, reliably lock onto the foreground outlines instead. The result: one image reads "SLEEP EARLY" to a person and "BINGE SHOWS" to GPT-5.6 Sol or Gemini 3.5. Same letterforms, two audiences, two messages that know nothing of each other.

The beauty here is structural, not conceptual. What Decoy Font reveals—and makes literally usable—is the gap between two fundamentally different visual regimes: the embodied, context-sensitive attention of human seeing, and the pixel-literal, frequency-naive scanning of current LLM vision. This gap is not a metaphor. It is a physical and computational fact about how two systems process spatial information. The hybrid image as optical illusion is already-construct—a good sixty years of vision science, and the Marilyn Monroe / Albert Einstein composite image has been reproduced so many times it's become a classroom example. The mechanism is named, understood, exhausted. But using that mechanism to write simultaneously to two different readerships—human and machine—choosing your reader by choosing your spatial frequency—is a logic that has not yet been named. It is still growing. That gap between what is known and what is not yet named is the remainder.

In SAE terms: the chisel-construct cycle describes how every genuine chisel-out generates new remainders, which subsequent eras digest into construct. Decoy Font stands at the open end of this cycle. "Anti-AI art" is beginning to crystallize as a category—IMAGO proposes locally-trained AI models as an ethical alternative to scraped training data, Ghost Font hides messages inside moving animations—but the specific move of spatial-frequency address, embedding two readers inside a single glyph, has no complete conceptual shell yet. It has not been absorbed into any discourse. It is alive as remainder precisely because the art world is slowly approaching it without having fully arrived.

Lu says openly that the font may stop working "in just a few months" as models advance. This is not a disclaimer. It is the exact temporal shape of 余项之美. The remainder exists in the window before the chisel has fully struck—before the technology has closed the gap, before the theory has given it a complete name. Today's LLMs process images pixel by pixel, zoomed in, reading only high-frequency edges; they miss the low-frequency shape that human eyes, at normal distance, perceive as the primary message. That "miss" is real, physical, determined by architecture. When multimodal models learn to perceive the way human vision does—with context, with distance, with the capacity to defocus—the gap closes. The remainder disappears. What is now a live question becomes a footnote in a history of the AI transition.

Seeing it now means seeing the fossil record of this mismatch while it's still forming. Decoy Font is not an art manifesto. It is a downloadable TTF file and a web encoder, published as a research experiment by an AI researcher at a type foundry—positioned between typography, machine perception, and quiet resistance to the reach of LLMs into every act of reading. That positioning between disciplines, before any discipline has claimed it, is exactly where 余项之美 tends to appear: not in the gallery, not in the paper, but in the gap that is still too new to have a room of its own.

mixfont.com ↗