This guy codes, and as soon as I can run a decent model with 8 gpu (the max a commercial motherboard can take), I am 100% getting some uncensored model and using it for my main LLM, MAYBE with the occasion claude request ($20 per month) for ultra-complex stuff). Bye bye chatgpt and gemini, I wont miss you.
I haven’t played with open router yet but apparently that’s the killer feature: use a local model most of the time, automatically switch to a paid one when needed.
This guy codes, and as soon as I can run a decent model with 8 gpu (the max a commercial motherboard can take), I am 100% getting some uncensored model and using it for my main LLM, MAYBE with the occasion claude request ($20 per month) for ultra-complex stuff). Bye bye chatgpt and gemini, I wont miss you.
I haven’t played with open router yet but apparently that’s the killer feature: use a local model most of the time, automatically switch to a paid one when needed.