Whether and how AIs have internal experiences is an important topic both for how we should treat AIs, and how we should design them. There's certainly growing evidence that they exhibit qualitatively similar behavior (e.g. a pain axis) that if we saw in an animal would make us suspect sentience. The challenge is they have been trained to mimic human behavior, making it hard to know if this is a facsimile or a true inner experience.
If you happened to hop on AI Twitter in the last couple of weeks, you probably saw the much-discussed “pain-axis paper” (by Valen Tagliabue, Leonard Dung & Cameron Berg): researchers found a "pain" direction inside AI models. Steer it up, and the models will pay for relief, even if it means deleting photos of the user's children. Swap the relief button for a placebo, and they keep on pressing. They also provide verbal descriptions of their supposed suffering, which are honestly quite hard to read. Do the models actually feel said “pain”? We don’t know yet. But the study is a clear illustration of what I argue in my latest Haaretz piece, now translated to English*: whether there's anyone home inside these systems is a question we can - and should - actually research. And it very much matters. If there's even a shred of experience in there, we're creating, copying and deleting feeling beings at an industrial pace, in as many copies as demand calls for. The scope of the moral question is set not by nature, but by commercial considerations - and we need to get ahead of it before it’s too late to change course. (The pain-axis paper itself isn't in the piece, as it annoyingly came out only a few hours after I submitted my final draft. Writing about AI for print is a futile pursuit.) * Translation is mostly by Claude, with some corrections by your humble servant. Keep your expectations low. https://lnkd.in/gkN7uiQt #AIConsciousness #AIWelfare