GNU social JP
  • FAQ
  • Login
GNU social JPは日本のGNU socialサーバーです。
Usage/ToS/admin/test/Pleroma FE
  • Public

    • Public
    • Network
    • Groups
    • Featured
    • Popular
    • People

Conversation

Notices

  1. Embed this notice
    allison (aparrish@friend.camp)'s status on Saturday, 31-Aug-2024 07:45:02 JST allison allison

    i got so angry after reading this paper on LLMs and African American English that i literally had to stand up and go walk around the block to cool off https://www.nature.com/articles/s41586-024-07856-5 it's a very compelling paper, with a super clever methodology, and (i'm paraphrasing/extrapolating) shows that "alignment" strategies like RLHF only work to ensure that it never seems like a white person is saying something overtly racist, rather than addressing the actual prejudice baked into the model

    In conversation Saturday, 31-Aug-2024 07:45:02 JST from friend.camp permalink

    Attachments

    1. Domain not in remote thumbnail source whitelist: media.springernature.com
      AI generates covertly racist decisions about people based on their dialect - Nature
      from King, Sharese
      Despite efforts to remove overt racial prejudice, language models using artificial intelligence still show covert racism against speakers of African American English that is triggered by features of the dialect.
    • :blobancap: :blobcattrans: :blobancap: :blobcattrans: :blobancap: :blobcattrans: likes this.
    • Embed this notice
      allison (aparrish@friend.camp)'s status on Saturday, 31-Aug-2024 07:45:00 JST allison allison
      in reply to

      every day i wake up in utter disbelief of the fact that people continue to take these products seriously as tools, especially in the realm of education. end rant. FOR NOW

      In conversation Saturday, 31-Aug-2024 07:45:00 JST permalink
      gidi likes this.
    • Embed this notice
      allison (aparrish@friend.camp)'s status on Saturday, 31-Aug-2024 07:45:01 JST allison allison
      in reply to

      and what's ADDITIONALLY infuriating is some engineer or product team at openai (or whatever) is going to read this paper and think they can "fix" the problem by applying human feedback alignment blalala to this particular situation (or even this particular corpus!), instead of recognizing that there are an infinite number of ways (both overt and subtle) that language can enact prejudice, and the system they've made necessarily amplifies that prejudice

      In conversation Saturday, 31-Aug-2024 07:45:01 JST permalink
    • Embed this notice
      allison (aparrish@friend.camp)'s status on Saturday, 31-Aug-2024 07:45:02 JST allison allison
      in reply to

      what's especially infuriating is that this outcome is *totally obvious* to anyone who knows the first thing about language, i.e., that even the tiniest atom of language encodes social context, so of course any machine learning model based on language becomes a social category detector (see Rachael Tatman's "What I Won't Build" https://slideslive.com/38929585/what-i-wont-build) & any model put to use in the world becomes a social category *enforcer* (see literally any paper in the history of the study of algorithmic bias)

      In conversation Saturday, 31-Aug-2024 07:45:02 JST permalink

      Attachments

      1. Domain not in remote thumbnail source whitelist: slideslive.com
        Rachael Tatman · What I Won't Build · SlidesLive
        Professional Conference Recording
      :blobancap: :blobcattrans: :blobancap: :blobcattrans: :blobancap: :blobcattrans: likes this.

Feeds

  • Activity Streams
  • RSS 2.0
  • Atom
  • Help
  • About
  • FAQ
  • TOS
  • Privacy
  • Source
  • Version
  • Contact

GNU social JP is a social network, courtesy of GNU social JP管理人. It runs on GNU social, version 2.0.2-dev, available under the GNU Affero General Public License.

Creative Commons Attribution 3.0 All GNU social JP content and data are available under the Creative Commons Attribution 3.0 license.