My impression is the general consensus is we don’t want huge corporations stealing data to train their AI models only to turn around and cram it down our throats anywhere they can with increasingly negative experiences.
That being said, while I would generally agree with that, I still find it interesting and especially if I can host it myself.
128k token context is pretty sweet. Mistral nemo also just launched with a similar context. Good times.
How does the Nemo 12B compare to the Llama 3.1 8B?
At long context (close to the full 128K), Nemo is way better than llama 8B in my testing.
Turns out they are both very sensitive to quantization though.
TBH I didn’t know people here were running LLMs. Seems like most of Lemmy is very broadly anti AI?
My impression is the general consensus is we don’t want huge corporations stealing data to train their AI models only to turn around and cram it down our throats anywhere they can with increasingly negative experiences. That being said, while I would generally agree with that, I still find it interesting and especially if I can host it myself.