good to have a quick overview, though I am missing some kind of depth, like
>There are others — MLX-native runtimes, llama.cpp directly — but Ollama and LM Studio cover 95% of people reading this.
Sure, but... MLX native runtime have a lot of advantages right? and they are not necessarily more complicated like eg ollama. It's worth to mention them, too.
Coming to the next missing part: use cases.
Running it locally competes with dozens of freely available chat bots. And for most simple parts, seriously, you don't need that sophisticated payed AI service. But then: how do local model's perform exactly here: let's start with coding, image generation and more? Like stuff that I probably really want to run locally.
>There are others — MLX-native runtimes, llama.cpp directly — but Ollama and LM Studio cover 95% of people reading this.
Sure, but... MLX native runtime have a lot of advantages right? and they are not necessarily more complicated like eg ollama. It's worth to mention them, too.
Coming to the next missing part: use cases.
Running it locally competes with dozens of freely available chat bots. And for most simple parts, seriously, you don't need that sophisticated payed AI service. But then: how do local model's perform exactly here: let's start with coding, image generation and more? Like stuff that I probably really want to run locally.
* nativ
* omlx
* mtplx