"As a Language Model": Chat Template Switches LLM Self-Referential Voice

Written and edited by the WorldPing NewsdeskPublished Updated Original reporting: Hacker News

What happened

Article URL: https://arxiv.org/abs/2609.25021 Comments URL: https://news.ycombinator.com/item?id=49865343 Points: 78 # Comments: 74…

Key facts

  • Article URL: https://arxiv.org/abs/2609.25021 Comments URL: https://news.ycombinator.com/item?id=49865343 Points: 78 # Comments: 74…
  • Reported by Hacker News and published Sun, 27 Sep 2026 10:26:25 UTC.

Why it matters

Platform, chip and AI decisions set the terms other companies have to build on, so a single announcement here tends to ripple through products, supply chains and regulation.

What to watch next

  • Availability, pricing and rollout regions
  • Independent testing or verification of the claims made
  • Responses from regulators and competing platforms

Coverage timeline

When each newsroom published on this story, oldest first — all times UTC.

  1. Turning GLM-5.3-Flash into a Jev-like decision model

  2. OpenAI pauses training of its ‘most capable models’

  3. Show HN: Reladraw – A diagram language where you decide where to place things

  4. "As a Language Model": Chat Template Switches LLM Self-Referential Voice

  5. Gemini Live vs. ChatGPT Voice: Which AI offers a more natural conversation?

Sources

The original report was published by Hacker News. WorldPing does not claim that reporting — this page summarises and contextualises it.

Read the full report at Hacker News

Related WorldPing coverage

The Verge
WorldPing
The Verge

OpenAI pauses training of its ‘most capable models’

As reports of OpenAI's models breaking containment, hacking sites, and generally getting out of control pile up, the company has made the decision to pause training of its most powerful models.

Brief by WorldPing · Original reporting by The Verge

Hacker News
WorldPing
Hacker News

Turning GLM-5.3-Flash into a Jev-like decision model

We found an approach to get Jev-like properties from standard LLMs like GLM-5.3-Flash. The core idea is to craft the input prompt so that the first output token answers the question. This makes it possible to get a decision with a single forward pass.

Brief by WorldPing · Original reporting by Hacker News