Switch language한국어
Back to the list

Truth is not a direction: a Tarski attack on LLM probes | Hacker News

TL;DR AI

Key summary

2 min read
  1. A Hacker News discussion argues that phrases like “this sentence” do not literally refer to themselves, because a sentence is not a speaking subject.

  2. The post says apparent self-reference is a category error between the speaker and the written or spoken sentence.

  3. It invokes Tarski-style semantics to question whether self-referential probes for LLMs are conceptually sound.

  4. The broader point is that if reference is ill-defined here, some language-model evaluation prompts may be built on shaky ground.

Read the original