GNU social JP
  • FAQ
  • Login
GNU social JPは日本のGNU socialサーバーです。
Usage/ToS/admin/test/Pleroma FE
  • Public

    • Public
    • Network
    • Groups
    • Featured
    • Popular
    • People

Conversation

Notices

  1. Embed this notice
    Dare Obasanjo (carnage4life@mas.to)'s status on Monday, 08-Sep-2025 10:09:00 JST Dare Obasanjo Dare Obasanjo

    OpenAI published a paper showing that LLMs hallucinate for two primary reasons: first, due to statistical errors during training, and second, because LLM evaluations reward confident, even if incorrect, guesses but penalize “I don’t know”. So LLMs are trained to always guess when they aren’t sure.

    The paper proposes that to fix this, we must modify LLM evaluations to properly reward models that acknowledge when they don't know the answer.

    https://openai.com/index/why-language-models-hallucinate/

    In conversation about a year ago from mas.to permalink
    • Rich Felker repeated this.
    • Embed this notice
      draeath (draeath@infosec.exchange)'s status on Monday, 08-Sep-2025 10:08:59 JST draeath draeath
      in reply to

      @carnage4life "properly reward models that acknowledge when they don't know the answer"

      Wouldn't that mean rewarding every output? /hides

      In conversation about a year ago permalink
    • Embed this notice
      Rich Felker (dalias@hachyderm.io)'s status on Monday, 08-Sep-2025 10:11:23 JST Rich Felker Rich Felker
      in reply to
      • Bill, organizer of stuff

      @wcbdata @carnage4life I mean for answers making mathematical claims, you could evaluate them. This would fix some of the stupidest stuff like arguing about number of occurrences of a letter in a word or values of mathematical expressions. But it won't fix the larger fundamental problems.

      In conversation about a year ago permalink
    • Embed this notice
      Bill, organizer of stuff (wcbdata@vis.social)'s status on Monday, 08-Sep-2025 10:11:24 JST Bill, organizer of stuff Bill, organizer of stuff
      in reply to

      @carnage4life Interesting statement, but meaningless. LLMs fundamentally cannot work that way. To "know the truth," you need a deterministic, rules-based system - very different from (and wildly more complex and expensive than) an LLM. There was an excellent mathematical proof of this a few years ago, but I can't be bothered to go find it at this hour. To oversimplify a bit, it's like saying that to make fire safe we need it to stop when it's burning something.

      In conversation about a year ago permalink

Feeds

  • Activity Streams
  • RSS 2.0
  • Atom
  • Help
  • About
  • FAQ
  • TOS
  • Privacy
  • Source
  • Version
  • Contact

GNU social JP is a social network, courtesy of GNU social JP管理人. It runs on GNU social, version 2.0.2-dev, available under the GNU Affero General Public License.

Creative Commons Attribution 3.0 All GNU social JP content and data are available under the Creative Commons Attribution 3.0 license.