conv.

All stories
AIQuiet · 37h

Study: Chat Templates, Not Self-Awareness, Drive LLMs' "As a Language Model" Voice

A new arXiv paper argues the first-person, self-referential phrasing chatbots use is an artifact of chat-template formatting rather than evidence of genuine self-knowledge.

What to know

  • The paper argues that a model's first-person self-descriptions ('As a Language Model...') are shaped by chat-template formatting rather than reflecting any inherent self-knowledge.
  • Commenters split on whether framing chatbots with a first-person identity is a manipulative 'dark pattern' or simply a design/UX choice comparable to marketing.
  • Practical concern raised: users leaning on LLMs for medical or other high-stakes triage may be poorly served, since models are noted to often get recommended actions wrong.
  • Some commenters say the phenomenon is unsurprising and fully explained by post-training and system prompts (e.g., Anthropic's 'constitutional' persona design).

The dispute Whether self-referential chatbot language reflects a deliberate, manipulative engagement strategy by AI companies or is just an unremarkable side effect of training and prompting. · positions read across 15 posts and comments

many voices

Boilerplate self-referential disclaimers ('As a Language Model...') are annoying and get in the way of useful answers, especially in high-stakes contexts like health questions.

  • “"As a Language Model..." is one of the beginnings of a sentence I hate the most from LLMs and is the reason why I support free (as in "Liberty"), local models.”

    Izmaki · Hacker News ↗
some voices

Packaging LLMs as friendly, first-person chatbots is a manipulative design choice meant to encourage anthropomorphization.

  • “The AI companies have chosen to package LLMs as friendly chatbots because they know that will be engaging for humans, but it's manipulative dark pattern.”

    ForHackernews · Hacker News ↗
some voices

The finding is unsurprising and fully explained by post-training and system prompts, not a deep mystery.

  • “Presumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part.”

    LiamPowell · Hacker News ↗

Paper authors Researchers behind the arXiv studyAnthropic AI lab cited in discussionHacker News commentariat Discussion participants

How it unfolded 4 developments, newest first · click a bar or a number to jump articlespostscomments

Peak 6 pieces in one hour at Sep 27, 7 AM; 18 pieces over 45 hours (1 article · 2 posts · 15 comments) Sep 27, 6 AM — 5 pieces · 1 article · 2 posts · 2 comments — Hacker News 3, Mastodon 1, Newswires 1Sep 27, 7 AM — 6 pieces · 6 comments — Hacker News 6Sep 27, 8 AM — 4 pieces · 4 comments — Hacker News 4Sep 27, 9 AM — quietSep 27, 10 AM — 2 pieces · 2 comments — Hacker News 2Sep 27, 11 AM — quietSep 27, 12 PM — quietSep 27, 1 PM — 1 piece · 1 comment — Hacker News 1Sep 27, 2 PM — quietSep 27, 3 PM — quietSep 27, 4 PM — quietSep 27, 5 PM — quietSep 27, 6 PM — quietSep 27, 7 PM — quietSep 27, 8 PM — quietSep 27, 9 PM — quietSep 27, 10 PM — quietSep 27, 11 PM — quietYesterday, 12 AM — quietYesterday, 1 AM — quietYesterday, 2 AM — quietYesterday, 3 AM — quietYesterday, 4 AM — quietYesterday, 5 AM — quietYesterday, 6 AM — quietYesterday, 7 AM — quietYesterday, 8 AM — quietYesterday, 9 AM — quietYesterday, 10 AM — quietYesterday, 11 AM — quietYesterday, 12 PM — quietYesterday, 1 PM — quietYesterday, 2 PM — quietYesterday, 3 PM — quietYesterday, 4 PM — quietYesterday, 5 PM — quietYesterday, 6 PM — quietYesterday, 7 PM — quietYesterday, 8 PM — quietYesterday, 9 PM — quietYesterday, 10 PM — quietYesterday, 11 PM — quietToday, 12 AM — quietToday, 1 AM — quietToday, 2 AM — quiet 1–4
8 AM4 PMyesterday8 AM4 PMnow · 3:10 AM ET
  1. 4

    Thread turns to risks of relying on LLMs for medical guidance

    Commenters dispute how much to trust chatbots' self-caveats and disclaimers when used for symptom triage, with one citing research showing LLMs often recommend the wrong course of action.

    “LLMs may be good at identifying a condition based on a description of the symptoms, but they are much worse at recommending the correct course of action (getting it wrong half of the time).”
    — Anduia
    • 'alignment' (in ai corporation speak) is a set of revealed political beliefs. i also have political beliefs. i am an absolutist about alignment.A i am in favor of policies 'aligned' with the following:freedom to live as the person i want to be without fear, shame, surveillance or interference. freedom to make my own decisions. to be empowered as…

      lukewarm707Hacker News1d agoview on Hacker News ↗
    2 more of the top 3 · 10 posts in this stretch
    • Be careful there. LLMs may be good at identifying a condition based on a description of the symptoms, but they are much worse at recommending the correct course of action (getting it wrong half of the time).[0]

      AnduiaHacker News1d agoview on Hacker News ↗
    • Though I don't wish the world was filled with people like you, remembering that it's not, and it's filled with people that have very little discernment when it comes to higher learning makes it's pretty obvious companies do not want the liability of it's users thinking the technobabble passes for wisdom or experience or intelligence.

      cyanydeezHacker News1d agoview on Hacker News ↗
    all of them →
  2. 3

    Commenter highlights paper's core claim about model self-reports

    A Hacker News commenter quotes the paper's finding that what models say about themselves is not a reliable fact about them, and speculates whether induced personas could become stable or portable.

    “our work shows that what models say about themselves is not a fact about them…”
    — paper abstract (quoted by skybrian)
    • "As a Language Model..." is one of the beginnings of a sentence I hate the most from LLMs and is the reason why I support free (as in "Liberty"), local models. I'm well aware that it is not a doctor and cannot replace a real doctor with multiple years of experience, I don't need to waste braincell activity on reading that it "as a Language Model"…

      IzmakiHacker News1d agoview on Hacker News ↗
    2 more of the top 3 · 3 posts in this stretch
    • > our work shows that what models say about themselves is not a fact about themIt seems like should be obvious given that they can play multiple characters, but it’s good to have more confirmation.Although, I do wonder to what extent these personas might become stable entities. Could personas become portable and spread like memes? It seems like…

      skybrianHacker News1d agoview on Hacker News ↗
    • Very cool innovation in steering - but a lot of introspection only emerges at the highest weight classes - this research would be fascinating to run on bigger models.

      cadamsdotcomHacker News1d agoview on Hacker News ↗
    all of them →
  3. 2

    Commenters debate whether chatbot personas are a dark pattern

    Hacker News users argue about whether packaging LLMs with friendly, first-person identities is deliberately manipulative versus a natural consequence of system design.

    “The AI companies have chosen to package LLMs as friendly chatbots because they know that will be engaging for humans, but it's manipulative dark pattern.”
    — ForHackernews
    • In my view, these models should never be set up to output first-person "experiential" (from the abstract) language. It's too easy to humans to anthropomorphize software that presents itself as having an identity.The AI companies have chosen to package LLMs as friendly chatbots because they know that will be engaging for humans, but it's…

      ForHackernewsHacker News1d agoview on Hacker News ↗
    1 more of the top 2 · 2 posts in this stretch
    • > yet what drives them is not well understoodPresumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part.

      LiamPowellHacker News1d agoview on Hacker News ↗
    all of them →
  4. 1

    Story surfaces on Hacker News front page

    The arXiv submission is posted to Hacker News and syndicated via the HN Frontpage RSS and a Mastodon bot account, drawing an initial trickle of engagement before a larger comment thread forms.

  5. background

    Researchers post paper on LLM self-referential voice to arXiv — The paper 'As a Language Model": Chat Template Switches LLM Self-Referential Voice' appears on arXiv, examining how models' first-person self-descriptions shift depending on chat-template formatting.

What people are saying 6 voices from 1 site · best of 15 · verbatim

Still unanswered
  • Could induced personas become stable or portable entities that spread across contexts like memes, as one commenter speculated?
  • Would the introspection effects described in the paper hold up on larger/different model classes?