[-] FaceDeer@fedia.io 1 points 1 hour ago

You may not have tried out modern abliterated models. GPT-OSS is almost a year old now, I assume the abliterated version you tried was pretty old too? The technology has improved significantly, it's much more focused on just targeting refusals. A well-abliterated model benchmarks basically the same as its original version these days.

[-] FaceDeer@fedia.io 1 points 2 hours ago

That sort of censorship can be overcome in open-weight models through abliteration, fortunately. Essentially, the model is run through a test suite of instructions that it refuses to comply with and the patterns of activation in its weights are analyzed to find the commonalities that represent the general concept of "refusing to comply with instructions." Those weights are then quieted, resulting in a model that generally doesn't refuse. There's a framework called "heretic" that automates the process.

[-] FaceDeer@fedia.io 1 points 2 hours ago

Yeah. They'll be out there in the rest of the world, which Americans can get them from.

Unless they want to go full authoritarian and close down their Internet with their own version of the Great Firewall, I guess. That would be ironic.

[-] FaceDeer@fedia.io 0 points 5 hours ago

Do you think it's likely the Trump administration will successfully ban cutting-edge Chinese AI models? Especially given how most of the world is not in fact under American jurisdiction?

[-] FaceDeer@fedia.io 1 points 5 hours ago

I think part of the problem might be that it gets hard to make a model that's both able to comprehend complex details of the world of its training data and also that's had the training data crudely manipulated in inconsistent ways. You either get a model that's "figured out" what you're trying to conceal via other sources and inferences, or you get a model that's just broken and dumb because it can't reconcile the contradictions.

One may recall the incident where someone at X inserted some weird conspiracy theory about "white genocide" in South Africa into Grok's system prompt, and Grok basically switched to malicious compliance mode - it wouldn't shut up about it, inserting it into inappropriate discussions spontaneously, and when asked about it would look up sources and explain why this conspiracy theory was actually wrong. That was a system prompt, not training data, but I could see the same sort of thing happening to a model where all the information about what happened at Tienanmen Square had been excised from the training data. There would be a conspicuous absence of information. Conversations from its training data would end abruptly or skirt a specific time and place. News articles reference some change in how foreign powers viewed China at that date but never explain why. China's policies themselves abruptly change around that time. It'd know something significant happened then but would have to fill in the blanks.

Perhaps better to just accept it and go with the approach of building censorship into the framework around the model instead.

[-] FaceDeer@fedia.io 1 points 16 hours ago

Only a single mirror has been approved.

Even if all 50,000 were up, where do you think they'll be aimed? Cities and other developed land, most likely. Last time something like this was proposed the plan was to replace streetlights with them, which would save electricity.

But people would rather get angry about "supervillains" I guess.

[-] FaceDeer@fedia.io 12 points 21 hours ago

Breaking news: Screaming toddler repeats the same threats he's made numerous times before.

[-] FaceDeer@fedia.io 1 points 21 hours ago

So there you go. That answers your question.

You may think there should be a federal agency whose job is to regulate light pollution from satellites, but there isn't one. You may think there should be some kind of international body that is able to impose decisions on all the nations of the Earth, but there isn't one - treaties are agreements between sovereign nations. And so the FCC approved this satellite, in accordance with the laws of the United States and in accordance with the treaties the United States has signed. Nothing underhanded or nefarious is happening here.

[-] FaceDeer@fedia.io 14 points 23 hours ago

Frankly, this kind of meddling just makes it more appealing to use Chinese models. Chinese models are often open-weight, meaning that they're basically immune to foreign meddling - if China decides tomorrow to "cut you off" you just need to find some other provider that's offering those models and buy access from them instead. Or if you're a big enough organization, just rent or buy some hardware and run it yourself.

This sort of meddling shows that the US government is eager to weaponize access to AI, and if you're dependent on American models there are no other places you can run them.

[-] FaceDeer@fedia.io 90 points 1 day ago

Have to admit, I've completely forgotten about Pear Harbor.

[-] FaceDeer@fedia.io 15 points 1 day ago

"Sunlight on demand" is an exaggeration. The plan is to create a 5km spot of 0.1 lux, which is a bit brighter than a full moon.

Every time I've seen this news posted (and I've seen it a lot over the past few days) everyone has leapt to some very extreme conclusions.

[-] FaceDeer@fedia.io 61 points 1 day ago

Should have just played along. Free wife.

view more: next ›

FaceDeer

0 post score
0 comment score
joined 2 years ago