It’s also largely good enough now that you can do coding (which is a big driver of usage) on a local model. It’s slower, but so what. Give it a prompt and go make a cup of coffee, vs give it a prompt and instantly get your code. Still faster than doing it yourself.
This guy codes, and as soon as I can run a decent model with 8 gpu (the max a commercial motherboard can take), I am 100% getting some uncensored model and using it for my main LLM, MAYBE with the occasion claude request ($20 per month) for ultra-complex stuff). Bye bye chatgpt and gemini, I wont miss you.
I haven’t played with open router yet but apparently that’s the killer feature: use a local model most of the time, automatically switch to a paid one when needed.
It’s also largely good enough now that you can do coding (which is a big driver of usage) on a local model. It’s slower, but so what. Give it a prompt and go make a cup of coffee, vs give it a prompt and instantly get your code. Still faster than doing it yourself.
This guy codes, and as soon as I can run a decent model with 8 gpu (the max a commercial motherboard can take), I am 100% getting some uncensored model and using it for my main LLM, MAYBE with the occasion claude request ($20 per month) for ultra-complex stuff). Bye bye chatgpt and gemini, I wont miss you.
I haven’t played with open router yet but apparently that’s the killer feature: use a local model most of the time, automatically switch to a paid one when needed.
ive not checked in in a while, what are the best local models right now?