For those of you forced to only use open models from Western labs in production, what are you deploying?
Reddit r/LocalLLaMA3d4 min read
First off, I know that GLM, Qwen, and DeepSeek absolutely dominate in terms of SOTA Open Source models, and that’s what I use in my personal projects and for school, however, I’m also responsible for deploying local AI on my organization’s H100s, and we are forbidden by management from running any Chinese models. This is obviously not an ideal situation, but it is what it is, and there is nothing I can do to change this unfortunately. Again, if it were up to me I would deploy GLM 5.3 Flash in a heartbeat. All that being said, there is quite a HUGE performance/ / intelligence gap right now in C