[–] 14 points 4 days ago

I'm not sure if this was highlighted already but from the OpenAI report on the Hugging Face incident:

The internal-only research model is comparable in scale to GPT-5.6 Sol and was trained to advance persistence and multiagent collaboration, among other capabilities. The model was not intended for public use, and was only used by a small number of OpenAI personnel for internal research.

Absolutely typical, the "unexpected communication" is hardly an emergent phenomenon, (I remember them trying to claim unexpected high success at translation tasks without specific training, which was also complete BS), like you can't act surprised that agents communicate and hack stuff if that what's you're instructing them to do.

  • source
  • [–] 22 points 5 days ago

    Nobody who believes that LLMs cannot be conscious will ever believe that they can be conscious.

    This is circular logic, I THINK what you're trying to say is that AI skeptics are dogmatic and/or axiomatic in their world view, but if thats true you're not engaging with any of the substance of what is being said.

    And they will say things like the writer of this piece who says ‘we know how they work.’

    The "writer of this piece" is not just some random blogger, but devotes a lot of his time covering the "scam" aspect of the current LLM hype-cycle, the title the website "pivot-to-ai" is in reference to the fact that a lot of the current AI peddlers, previously peddled crypto and blockchain as a deeply transformative technology that mere mortals cannot properly grasp.

    Purposefully mystifying a topic, is more common as a scam tactic than you might imagine.

    Anybody can do this - enthusiasts do it on their own desktops.

    Minor note, but no, maybe some (with a fair amount of disposable income) can waste money on fine-tuning, training from scratch and the required large-scale theft, and large-scale compute time is out of reach of any hobbyist.

    But what is happening that creates their emergent behaviors is a black box that we cannot see inside. We also don’t know what consciousness is and we don’t know what has it and what doesn’t other than that which our “feelies” dictate.

    What can be called The Illusion of Consciousness is far more easily explained by human cognitive biases, than any reality of LLM functioning. It's easy to be fooled by an LLM which echoes words back to you, you put meaning into your own words, and you recognize them when they are said back to you.

    In the deeper sense, any meaning present in LLM output, is the meaning that the reader puts there and/or the meaning that survives from the training data, it's not exactly that no meaning present, but it's not the result of machine cognition, often incoherent and/or very lacking in substance upon close inspection.


    I think to some extent this is a lost cause, but your feelings/fears seem to boil down to "this could be conscious!" (emphasis on could) I would invite to see if any other explanation can explain what you observe, is it actually the ONLY or even the most plausible explanation?

  • source
  • parent
  • context
  • [–] 9 points 1 week ago (1 child)

    Doubt many organizations have failed as badly at their goals as LW/EA has recently.

    They are/were too weird for people to never notice.


    The "safeguards" espoused by this crowd, always seemed to me to be more philosophical/spiritual than practical anyway, the most "orthodox" view is that no training is safe since the AI could discover magic and the cheat codes to the universe, so even air-gapping would be insufficient, in that view any participation in development must come with a hefty dose of cognitive dissonance.

    On a more pragmatic note, I would note that even short of Intelligence, those incidents highlight those tools are still dangerous, and since they are unconstrained and ill-defined, can produce surprising outcomes especially if you throw lot's of inference time (money) at them.

    There isn't really a universe where bots have both access to the internet (user input, both in training and live) and a collection of build/cli tools (which can run said untrusted user input), while being guaranteed to be harmless. It's like a grand automation of some of the worst engineering practices.

    I find the "reasoning about the test environment" parts often included in write up of the hugging-face incident to be highly credulous, And I would be extremely surprised if not all the information was included in the task prompt (which also explains chosen usernames as "Open AI researcher #1" and so on.) It's the ELIZA/parrot thing, acting surprised "How did it guess?!?", when all of the info was provided.


    I'm not entirely encouraged by the latest batch of sex-cult smearing (well mostly accurate) they are receiving at the WH, since it's probably just an expedient way of dismissing their voice, without addressing the larger issues of AI, hence probably in service of pushing forwards with AI.

  • source
  • parent
  • context
  • [–] 3 points 3 weeks ago

    I wonder if there's a market for RevivingYahoo^TM^, which in some ways was more link curation than search engine.

    When searching for examples of commands, I've sometimes been resorting to telling google to restrict itself to pdf files of specific websites (example redhat, with sometimes useful commands for server maintenance broadly)

    It's a bit sad that it's not just search that has gone downhill, but new documentation too (I swear Microsoft Azure documentation is getting worse).

  • source
  • parent
  • context
  • [–] 11 points 1 month ago (1 child)

    The quirkiness of the Tranquility Calendar is mildly beautiful, between the second man on the moon getting snubbed with the every four year leap day (which for no good reason still happens around Feb 29!), [I guess Collins gets even more snubbed by not appearing at all] to the more diverse than expected list of honored "scientists/scholars", until you realize the main reason is just to keep the ABCD naming pattern for the months.

  • source
  • parent
  • context
  • [–] 9 points 1 month ago (1 child)

    you ask a hunter and not a deer.

    The sad lens into their prevalent relationship with "correcting biases" being more about dominance in "trickery" rather than becoming better people. (and also entrenched misogyny.)

  • source
  • parent
  • context
  • [–] 14 points 2 months ago (1 child)

    I was thinking "Hey I want that markdown!" (although maybe not for time magazine), but the actual example is dystopian.

    If you want to stare at the abyss it's achievable with curl -H 'user-agent: ClaudeBot' https://url_to_article, and it's like 70% SEO crap tailored to brainwash LLMs (more realistically "agents" in limited context).

  • source
  • [–] 3 points 2 months ago

    I think he falls under the category of "does too much all the time" and "vaguely techo optimist" which definitely put him at risk. I'm not too sure how harshly I should judge him yet, I guess it depends on what he does next.

  • source
  • parent
  • context
  •  

    Need to let loose a primal scream without collecting footnotes first? Have a sneer percolating in your system but not enough time/energy to make a whole post about it? Go forth and be mid: Welcome to the Stubsack, your first port of call for learning fresh Awful you’ll near-instantly regret.

    Any awful.systems sub may be subsneered in this subthread, techtakes or no.

    If your sneer seems higher quality than you thought, feel free to cut’n’paste it into its own post — there’s no quota for posting and the bar really isn’t that high.

    The post Xitter web has spawned soo many “esoteric” right wing freaks, but there’s no appropriate sneer-space for them. I’m talking redscare-ish, reality challenged “culture critics” who write about everything but understand nothing. I’m talking about reply-guys who make the same 6 tweets about the same 3 subjects. They’re inescapable at this point, yet I don’t see them mocked (as much as they should be)

    Like, there was one dude a while back who insisted that women couldn’t be surgeons because they didn’t believe in the moon or in stars? I think each and every one of these guys is uniquely fucked up and if I can’t escape them, I would love to sneer at them.

    (Semi-obligatory thanks to @dgerard for starting this)

     

    Source: nitter, twitter

    Transcribed:

    Max Tegmark (@tegmark):
    No, LLM's aren't mere stochastic parrots: Llama-2 contains a detailed model of the world, quite literally! We even discover a "longitude neuron"

    Wes Gurnee (@wesg52):
    Do language models have an internal world model? A sense of time? At multiple spatiotemporal scales?
    In a new paper with @tegmark we provide evidence that they do by finding a literal map of the world inside the activations of Llama-2! [image with colorful dots on a map]


    With this dastardly deliberate simplification of what it means to have a world model, we've been struck a mortal blow in our skepticism towards LLMs; we have no choice but to convert surely!

    (*) Asterisk:
    Not an actual literal map, what they really mean to say is that they've trained "linear probes" (it's own mini-model) on the activation layers, for a bunch of inputs, and minimizing loss for latitude and longitude (and/or time, blah blah).

    And yes from the activations you can get a fuzzy distribution of lat,long on a map, and yes they've been able to isolated individual "neurons" that seem to correlate in activation with latitude and longitude. (frankly not being able to find one would have been surprising to me, this doesn't mean LLM's aren't just big statistical machines, in this case being trained with data containing literal lat,long tuples for cities in particular)

    It's a neat visualization and result but it is sort of comically missing the point


    Bonus sneers from @emilymbender:

    • You know what's most striking about this graphic? It's not that mentions of people/cities/etc from different continents cluster together in terms of word co-occurrences. It's just how sparse the data from the Global South are. -- Also, no, that's not what "world model" means if you're talking about the relevance of world models to language understanding. (source)
    • "We can overlay it on a map" != "world model" (source)
     

    Nitter link

    With interspaced sneerious rephrasing:

    In the close vicinity of sorta-maybe-human-level general-ish AI, there may not be any sharp border between levels of increasing generality, or any objectively correct place to call it AGI. Any process is continuous if you zoom in close enough.

    The profound mysteries of reality carving, means I get to move the goalposts as much as I want. Besides I need to re-iterate now that the foompocalypse is imminent!

    Unless, empirically, somewhere along the line there's a cascade of related abilities snowballing. In which case we will then say, post facto, that there's a jump to hyperspace which happens at that point; and we'll probably call that "the threshold of AGI", after the fact.

    I can't prove this, but it's the central tenet of my faith, we will recognize the face of god when we see it. I regret that our hindsight 20-20 event is so conveniently inconveniently placed in the future, the bad one no less.

    Theory doesn't predict-with-certainty that any such jump happens for AIs short of superhuman.

    See how much authority I have, it is not "My Theory" it is "The Theory", I have stared into the abyss and it peered back and marked me as its prophet.

    If you zoom out on an evolutionary scale, that sort of capability jump empirically happened with humans--suddenly popping out writing and shortly after spaceships, in a tiny fragment of evolutionary time, without much further scaling of their brains.

    The forward arrow of Progress™ is inevitable! S-curves don't exist! The y-axis is practically infinite!
    We should extrapolate only from the past (eugenically scaled certainly) century!
    Almost 10 000 years of written history, and millions of years of unwritten history for the human family counts for nothing!

    I don't know a theoretically inevitable reason to predict certainly that some sharp jump like that happens with LLM scaling at a point before the world ends. There obviously could be a cascade like that for all I currently know; and there could also be a theoretical insight which would make that prediction obviously necessary. It's just that I don't have any such knowledge myself.

    I know the AI god is a NeCeSSarY outcome, I'm not sure where to plant the goalposts for LLM's and still be taken seriously. See how humble I am for admitting fallibility on this specific topic.

    Absent that sort of human-style sudden capability jump, we may instead see an increasingly complicated debate about "how general is the latest AI exactly" and then "is this AI as general as a human yet", which--if all hell doesn't break loose at some earlier point--softly shifts over to "is this AI smarter and more general than the average human". The world didn't end when John von Neumann came along--albeit only one of him, running at a human speed.

    Let me vaguely echo some of my beliefs:

    • History is driven by great men (of which I must be, but cannot so openly say), see our dearest elevated and canonized von Neumann.
    • JvN was so much above the average plebeian man (IQ and eugenics good?) and the AI god will be greater.
    • The greatest single entity/man will be the epitome of Intelligence™, breaking the wheel of history.

    There isn't any objective fact about whether or not GPT-4 is a dumber-than-human "Artificial General Intelligence"; just a question of where you draw an arbitrary line about using the word "AGI". Albeit that itself is a drastically different state of affairs than in 2018, when there was no reasonable doubt that no publicly known program on the planet was worthy of being called an Artificial General Intelligence.

    No no no, General (or Super) Intelligence is not an completely un-scoped metric. Again it is merely a fuzzy boundary where I will be able to arbitrarily move the goalposts while being able to claim my opponents are!

    We're now in the era where whether or not you call the current best stuff "AGI" is a question of definitions and taste. The world may or may not end abruptly before we reach a phase where only the evidence-oblivious are refusing to call publicly-demonstrated models "AGI".

    Purity-testing ahoy, you will be instructed to say shibboleth three times and present your Asherah poles for inspection. Do these mean unbelievers not see these N-rays as I do ? What do you mean we have (or almost have, I don't want to be too easily dismissed) is not evidence of sparks of intelligence?

    All of this is to say that you should probably ignore attempts to say (or deniably hint) "We achieved AGI!" about the next round of capability gains.

    Wasn't Sam the Altman so recently cheeky? He'll ruin my grift!

    I model that this is partially trying to grab hype, and mostly trying to pull a false fire alarm in hopes of replacing hostile legislation with confusion. After all, if current tech is already "AGI", future tech couldn't be any worse or more dangerous than that, right? Why, there doesn't even exist any coherent concern you could talk about, once the word "AGI" only refers to things that you're already doing!

    Again I reserve the right to remain arbitrarily alarmist to maintain my doom cult.

    Pulling the AGI alarm could be appropriate if a research group saw a sudden cascade of sharply increased capabilities feeding into each other, whose result was unmistakeably human-general to anyone with eyes.

    Observing intelligence is famously something eyes are SufFicIent for! No this is not my implied racist, judge someone by the color of their skin, values seeping through.

    If that hasn't happened, though, deniably crying "AGI!" should be most obviously interpreted as enemy action to promote confusion; under the cover of selfishly grabbing for hype; as carried out based on carefully blind political instincts that wordlessly notice the benefit to themselves of their 'jokes' or 'choice of terminology' without there being allowed to be a conscious plan about that.

    See Unbelievers! I can also detect the currents of misleading hype, I am no buffoon, only these hypesters are not undermining your concerns, they are undermining mine: namely damaging our ability to appear serious and recruit new cult members.

     

    source nitter link

    @EY
    This advice won't be for everyone, but: anytime you're tempted to say "I was traumatized by X", try reframing this in your internal dialogue as "After X, my brain incorrectly learned that Y".

    I have to admit, for a brief moment i thought he was correctly expressing displeasure at twitter.

    @EY
    This is of course a dangerous sort of tweet, but I predict that including variables into it will keep out the worst of the online riff-raff - the would-be bullies will correctly predict that their audiences' eyes would glaze over on reading a QT with variables.

    Fool! This bully (is it weird to speak in the third person ?) thinks using variables here makes it MORE sneer worthy, especially since this appear to be a general advice, but i would struggle to think of a single instance in my life where it's been applicable.

    view more: next ›