ILoveDotUK
Hosting Provider
- Joined
- Jul 11, 2026
- Posts
- 37
- Reaction score
- 25
TinyLlama 1.1b could be a good fit because its only 1.1b parameters
I've just popped Helmuts a DM with a potential solution to this, will see what he comes back with![]()
For us to self-host a simple LLM ("simpleton") so that everyone can use it
I've tried literally every local model and they are all either incredibly slow or just can't hack what I throw at them. I'm sure in a few years we will have fable5 like local capabilities but until then I'm more than willing to shell out literally hundreds if not thousands of dollars per month to anthropic and openai for their flagship models. Why? Because they are phenomenal, save me thousands of FTE hours and are well worth the money in my opinion. Local models are still beaten by free versions of claude/chatgpt (or the lowest subscription) in my opinion.Guys, instead of paid models why not just implement free models which could be ran locally and therefore wouldn't require any purchases to be made ??
I run an instance of Qwen3-Coder locally and its extremely effective for complex tasks.
Its benchmarks are on par with Claude Sonnet 4.5 ...
If you need any help etc feel free to reach out. I run a variety of models nice and easily via llama
We use essential cookies to make this site work, and optional cookies to enhance your experience.