I’m an anarchocommunist, all states are evil.

Your local herpetology guy.

Feel free to AMA about picking a pet/reptiles in general, I have a lot of recommendations for that!

  • 1 Post
  • 291 Comments
Joined 2 years ago
cake
Cake day: July 6th, 2024

help-circle



  • “I don’t presume Anthropic is designing anything to be ethical.”

    Then why do they put any safeguards of any sort in? They put a lot of work into this, also this is openai.

    “First, the chatbot is unpredictable by design. Randomness is built in.”

    Randomness is built in but this behavior was not designed, it was unpredictable and random, which is, yes, unpredictable. The goal isn’t to make it unpredictable and temperature controls exist for a reason.

    “Second, if they didn’t intend for bad things to happen, why didn’t an employee babysit it?”

    They did they just weren’t paying enough attention, this was a benchmark. You have to have someone check the results to be useful.

    I reject that rogue is a goalpost shift or inaccurate. I think calling this rogue behavior is accurate.

    rogue /rōg/ noun

    An unprincipled, deceitful, and unreliable person; a scoundrel or rascal. One who is playfully mischievous; a scamp. 
    

    It did act that way, no? I did not shift goalposts. Remember the quote was


  • They did not intend for it to hack those websites, it realized that was a way of achieving its goals even though it’s specifically designed to be ethical. That can easily be called going rogue and I see no issue with it. They did not design it to pass the benchmark by hacking the website the benchmark was on. They hardly really design llm’s.

    As the article said you can replace rogue agents with unpredictable computer programs if you really want, I don’t see the necessity.

    Regardless that’s a major goalpost shift.



  • the first article doesn’t say anything not factual, the second article has a direct response to that

    “By now, if you’re an A.I. skeptic, you’re probably silently yelling at me for anthropomorphizing these systems. Go ahead, but feel free to replace “rogue agents” with “unpredictable computer programs””

    the third article also says nothing like that.

    the initial article claims ““AI” – chatbots that wake up, “set their own goals,” and “spontaneously” start hacking servers – is fake.”

    an article saying that is your goalpost, none of these articles say that. Even the headlines don’t say it and that’s where they normally put the crazy claims.

    ironically the claim that people claim that the AI chatbots woke up set their own goals and spontaneously started hacking servers seems to be fake.