>>109311709
okay smartass go ask
What are the best current local models useful for productive workloads on an RTX 5070 Ti? And for what usecase?
Should I give up on trying to run a functional agent like Hermes? Or is it possible with say, Gemma 4 26B Q4? I'm guessing Qwen3.6 35B is a bit too tough.
What about other usecases? Like code autocomplete? I heard this is a legit option for programming workflows because of less latency than cloud models, but would I need to spend time finetuning it for my project or a specific language?
chatgpt then and fuck off