top 50 comments

sorted by: hot top controversial new old
[–] 110 points 2 years ago

DeepSeek about to get sent in for "maintenance" and docked 10K in social credit.

  • source
  • [–] 77 points 2 years ago* (4 children)

    Yet unlike American led LLM companies Chinese researchers open sourced their model leading to government investment

    So the government invests in a model that you can use, including theoretically removing these guardrails. And these models can be used by anyone and the technology within can be built off of, though they do have to be licensed for commercial use

    Whereas America pumps 500 billion into the AI industry for closed proprietary models that will serve only the capitalists creating them. If we are investing taxpayer money into concerns like this we should take a note from China and demand the same standards that they are seeing from deepseek. Deepseek is still profit motivated; it is not inherently bad for such a thing. But if you expect a great deal of taxpayer money then your work needs to open and shared with the people, as deepseeks was.

    Americans are getting tragically fleeced on this so a handful of people can get loaded. This happens all the time but this time there’s a literal example of what should be occurring happening right alongside. And yet what people end up concerning themselves with is Sinophobia rather than the fact that their government is robbing them blind

    Additionally American models still deliver pro capitalist propaganda, just less transparently: ask them about this issue and they will talk about the complexity of “trade secrets” and “proprietary knowledge” needed to justify investment and discouraging the idea of open source models, even though deepseeks existence proves it can be done collaboratively with financial success.

    The difference is that deepseeks censorship is clear: “I will not speak about this” can be frustrating but at least it is obvious where the lines are. The former is far more subversive (though to be fair it is also potentially a byproduct of content consumed and not necessarily direction from openai/google/whoever)

  • source
  • hideshow 8 child comments
  • [–] 41 points 2 years ago

    Closed AI sucks, but there are definitely open models from American companies like meta, you make great points though. Can't wait for more open models and hopefully, eventually, actually open source models that include training data which neither deepseek nor meta do currently.

  • source
  • parent
  • [–] 11 points 2 years ago

    But Deepseek isn't Open Source by any definition of that word that I'm familiar with. Sure, they release more components than ProprietaryAI (which is a low bar,) but what you're left with is still a blob with a lot of the source code not released and no data set published as far as I can tell. Also, if I wanted to train my own model with the tools released, I'd still need millions of GPU hours. As I said, they are more transparent than others, but let's not warp the definitions of words just to give a "win" to another company that is just making another hallucination machine.

  • source
  • parent
  • [–] 43 points 2 years ago (55 children)

    If your system relies on censoring opposition to it then its probably not very good.

  • source
  • hideshow 57 child comments
  • load more comments (53 replies)
    [–] 38 points 2 years ago (12 children)

    Just run the LLM locally with open-webui and you can tweak the system prompt to ignore all the censorship

  • source
  • hideshow 15 child comments
  • load more comments (9 replies)
    [–] 35 points 2 years ago (1 child)

    Yeah, it’s pretty blatant. A bit after it hit the scene I got curious and started asking it about how many people various governments have killed. The answer for my own US of A was as long as it was horrifying.

    Then I get to China and it starts laying out a detailed description for a few seconds, then the answer disappears and is replaced by the “out of scope” or “can’t do that right now” or whatever it was at the time.

    It makes me think their model might be fine, but then they have some kind of watchdog layered on top of it to detect the verboten subjects and interfere. I guess that feels better from a technical standpoint, even if it is equally bad from a personal/political one.

  • source
  • hideshow 2 child comments
  • [–] 20 points 2 years ago* (2 children)

    DeepSeek isn't the only AI to censor itself after it generates text.

    I once asked Copilot for the origin of the "those just my little ladybugs" meme, and once it generated the text "perineum and anus" it wiped the answer it had written thus far and said that it couldn't look for that right now. I checked again today and it had since sanitized the answer so it generates in full.

  • source
  • parent
  • hideshow 3 child comments
  • load more comments (1 reply)
  • [–] 31 points 2 years ago (3 children)

    i mean, just ask DeepSeek on a clean slate to tell about Beijin.

  • source
  • hideshow 4 child comments
  • [–] 8 points 2 years ago (1 child)
  • [–] 13 points 2 years ago (2 children)

    its the capital city of China :D

    you know, where something happend on a specific square in the specific year of 1984.

  • source
  • parent
  • hideshow 4 child comments
  • load more comments (2 replies)
    [+] 29 points 2 years ago* (last edited 1 year ago) (1 child)
  • [–] 19 points 2 years ago (2 children)

    Is this real? On account of how LLMs tokenize their input, this can actually be a pretty tricky task for them to accomplish. This is also the reason why it's hard for them to count the amount of 'R's in the word 'Strawberry'.

  • source
  • hideshow 4 child comments
  • [–] 6 points 2 years ago (9 children)

    It’s probably deepseek r1, which is a “reasoning” model so basically it has sub-models doing things like running computation while the “supervisor” part of the model “talks to them” and relays back the approach. Trying to imitate the way humans think. That being said, models are getting “agentic” meaning they have the ability to run software tools against what you send them, and while it’s obviously being super hyped up by all the tech bro accellerationists, it is likely where LLMs and the like are headed, for better or for worse.

  • source
  • parent
  • hideshow 9 child comments
  • load more comments (9 replies)
  • [–] 3 points 2 years ago

    The LLM doesn't have to innately implement filtering. You can use a more traditional and concrete filtering strategy on top. So you sneak something problematic by in the prompt and it's too clever to be caught by the input filter, but then on the output the filter can catch that the prompt tricked the LLM into generating something undesired. Another comment specified they tried this and it started to work but then suddenly it seemingly shut out the reply in the middle, presumably the minute the LLM spit something at a more traditional filter and that shut it down.

    I think I've seen this sort of approach has been applied to largely mask embarassing answers that become memes, or to detect input known not to work, and to shut it down or redirect it to a better facility (e.g. redirecting math to wolfram alpha).

  • source
  • parent
  • [–] 14 points 1 year ago (14 children)

    The censorship surrounding taiwan is super lame. You can't even ask about stuff that happened 50+ years ago. Even if you ask about something else, if somehow the answer includes the world Taiwan, it deletes it.

  • source
  • hideshow 14 child comments
  • load more comments (14 replies)
    [–] 11 points 1 year ago (5 children)

    It's open source, although not 'cause it wants to be but 'cause it's the best way to compete with mainstream non-chinese software internationally, you can easily remove any censorship included by default.

  • source
  • hideshow 5 child comments
  • load more comments (5 replies)
    [–] 8 points 2 years ago

    HAHAHA! When I tried it, it started answering it, but quit and showed me the OOS message instead...

  • source
  • [–] 8 points 2 years ago

    I was told there would be no math

  • source
  • [–] 7 points 2 years ago

    Try an uncensored version, because everyone knows Communists hate Hexadecimal /s

  • source
  • [+] 4 points 2 years ago* (last edited 1 year ago) (3 children)
    load more comments (3 replies)
    [–] 4 points 1 year ago
    load more comments
    view more: next ›