[–] 18 points 5 hours ago (3 children)

Joke's on them, I'm almost done training FaceBot 9000.

Granted, it's a 1.2B model and the training data is entirely My Little Pony fanfics. It's insane and thinks it's Princess Luna. But it's going to dominate this benchmark.

  • source
  • [–] 10 points 17 hours ago (3 children)

    That's not how supply and demand works. When a market is out of economic equilibrium like this, with demand sharply rising, that is literally what drives competitors to arise and to try to increase supply to match. This is fertile ground for exactly that outcome to happen.

  • source
  • parent
  • context
  • [–] 1 point 1 day ago

    Heh, I downloaded that one just recently, I read that it was good at natural prose and I've been working on a little pet project to make a framework for auto-writing short stories based on a simple premise. Haven't tested it extensively yet though.

  • source
  • parent
  • context
  • [–] 1 point 1 day ago (2 children)

    Sadly they don't seem to have 36B-A3B models on their roadmap any more, they didn't do one for 3.8 either. I agree it was a nice sweet spot between speed and capability, I still use the 3.6 version of 36B-A3B for larger-scale local work. Maybe someone else will aim for that. Or Qwen4 will do something new with the architecture that makes it unnecessary.

  • source
  • parent
  • context
  • view more: next ›