• 6 Posts
  • 589 Comments
Joined 3 years ago
cake
Cake day: August 29th, 2023

help-circle



  • A “no attacks on people’s characters” policy seems like a terrible idea

    Rationalists: “Its not our fault that we got deceived and manipulated by SBF, lots of people were fooled by him! pay no attention to the fact we repeated the pattern of falling for it with OpenAI and Anthropic also

    Also Rationalists: “You can’t jsut call people in or adjacent to our in-group bad people!”

    Also Eliezer: "There’s a standard Internet phenomenon (I generalize) of a Sneer Club of people who enjoy getting together and picking on designated targets. Sneer Clubs (I expect) attract people with high Dark Triad characteristics, which is (I suspect) where Asshole Internet Atheists come from - if you get a club together for the purpose of sneering at religious people, it doesn’t matter that God doesn’t actually exist, the club attracts psychologically f’d-up people. Bullies, in a word, people who are powerfully reinforced by getting in what feels like good hits on Designated Targets, in the company of others doing the same and congratulating each other on it. "


  • My first thought was maybe one of them isn’t focused enough on doomerism and is too concerned with practical shit, which in lesswrong and AI doomer spaces is a nono (practical concerns about GenAI and data centers are okay as a way to draw in normies, but they want to make sure you speak the shibboleths that indicate your real concern is Skynet). Thus whichever one is the less doomer one would get a smidgen of credit in my eyes.

    I was disappointed to learn the real reason…

    We endorse pointing out harmful actions. But we condemn attacks on people’s character.

    …although it is probably a good thing that someone is willing to call Sam Altman a slimeball and Anthropic lying hypocrites? So I’ll give PauseAI-US a tiny bit of credit for that.

    …or maybe not

    When I read Holly’s Twitter feed, I get why someone would have a gut reaction like “this is an unhinged person who’s stooping to the level of personal attacks on people who seem nice and are just trying to do their best”.

    Whatever credit this PauseAI-US person gets for willingness to call people out, they lose for doing so through twitter.









  • Bioman has already pointed out the “Economic Value” numbers they are using for these tables are probably bullshit (based on self reported run-rate extrapolations that are deliberate distortions at best, based on VC valuation at worst). To add to this… the compute values are also probably bullshit. They are likely based on data center announcements and not confirmed totally complete data centers (Ed Zitron has ripped into how much bs there is in data center announcements). “Coding Time Horizon” is probably METR, which, while some of the best numbers for estimating actual AI improvement for practical purposes, are still really bad in several key ways. (They don’t have enough human task performers for the longer duration tasks even if everything else was right, because they aren’t, and there are several ways systematic bias could have leaked in and compelted distorted the constructed measure of task duration.)

    “AI Software R&D Uplift” is the single most important category to their scenario of recursive self improvement… and they have it at a small fraction of what they estimated.




  • Skimming the linked nature article… is my understanding correct that you can basically disable the watermarking by turning temperature down to 0?

    For example, if the LLM distribution is very low entropy, meaning it almost always returns the exact same response to the given prompt, then Tournament sampling cannot choose tokens that score more highly under the g functions.

    Other highlights from the linked paper… to get a true positive rate of 90% with a false positive rate of 1% you need 400 tokens (which should be a few paragraphs worth of text)? (If I’m reading figure 3 right?) That actually isn’t that much, relative to the lengths of essays people write for high school and college classes. …well actually… 1% false positive doesn’t sound too bad, but if you have thousands of freshmen students all taking classes involving writing essays and checking for watermarks becomes the norm, that is dozens and dozens of false positive, which means lots of false accusations, and as we’ve seen from how teachers and institutions have tried utilizing the existing “AI detection” tools that are much much less reliable… I’m getting angry just thinking about it.

    Edit: on turning temperature down, it should be noted Anthropic and OpenAI have been increasingly denying the end user internals of their models, such as summarizing or even outright hiding the thinking traces, and not allowing them access to temperature settings either.



  • One of the plot threads was about the invention of AI, in the Data-from-Star-Trek kind of sense.

    Going on a tangent from this… Its funny how much AI-related sci-fi is feeling increasing quaint (and often foolishly optimistic).

    You know what depiction of AI has held up, and maybe makes even more sense in light of modern AI? Star Wars’ take on Droid. Prior to LLMs it seems obvious that droids were sentient and deserving of rights. Post LLMs… it seems pretty plausible you could get something like C3PO that blabbers on but lacks any meaningful sentience or sapience. Also, it seemed obviously idiotic the way the Trade Federation had humanoid droids acting as pilots and gunners instead of building proper autopilots and targeting system. But after seeing LLMs get shoved everywhere, I totally could imagine a megacorp shoving droids built from standardized pretrained modules into all kinds of applications they aren’t fit for.