r/LocalLLM • u/yasintoy • 7h ago
Question Engineers running open-source LLMs in production: what is the hardest part today?
/r/mlops/comments/1w4rm61/engineers_running_opensource_llms_in_production/
0
Upvotes
r/LocalLLM • u/yasintoy • 7h ago
1
u/Atretador ArchLinux Xeon E5 2673 V4 20C/40T 4x16Gb DDR4 2133 2xMI50 16Gb 7h ago
Qwen 3.6 35B A3B - general backend and frontend work, debugging with real credentials
Its cheap and fast
nothing really
long sessions with big context slows down on shitty hardware
control, privacy and security
if a runtime is faster I run it
hardware prices
my power cost is bout the same as a cheap sub like a gpt plus
its on my table
this doesnt seem to be a question for local models
for cloud providers? I mostly run Opencode go with MiMo, price is just insane
how would I have more control than: its on my table