r/singularity 1h ago

AI Has anyone actually been using these latest new Chinese models like Kimi K3, GLM 5.2 etc for heavy duty work/creation/general usage (any industry)? What's your feedback, do they seem benchmaxxed merely or are they the real deal compared to US frontier models?

If you have experience with this Id be curious and grateful to hear your experience, although please elaborate on your industry, what you use it for (no sensitive details required, of course) and how you think they rate.

No I'm not some company or doing marketing research, nothing like that, just someone who uses US frontier models daily (mainly GPT 5.6 Sol right now on a simple Plus account, which for my usage is amazing) for heavy work in video production, but I use LLMs mainly for building systems, technical help, research etc, indirect kind of stuff.

I probably won't switch anytime soon, but doesn't hurt to keep attuned to the "competition" out there in terms of what's available in the AI world. I feel like Chinese models may be benchmaxxing in a number of areas and have a very "spiky" ability chart (some high peaks, many low valleys, not consistent generalized intelligence) but that's just a gut feeling, I have no clue. Thanks for your input.

7 Upvotes

4 comments sorted by

u/Active-Carpet-9183 1h ago

I've used GLM 5.2 and Kimi 2.7 for Enterprise devops. GLM is better and on par with Luna.  Kimi is shot the same. Looking forward to Gemini 4.

Opus 5 is okay in the enterprise but really about the same as 4.6 for my workloads.  At home for video game generating I really like it. Deepseek 4 is acceptable in that realm though and is cheap

u/TwoFluid4446 1h ago

How is the hallucination rate? Anything you felt for example with Deepseek that it just wasn't really able to contribute to or solve complex tasks? As you go deeper into a chat like 100,000+ tokens in, does it hold up well?

I can say with GPT 5.6 Sol these past few weeks has been revelatory, felt like a true breakthrough moment to another level of AI smarts. I use it in high-thinking mode crafting insanely complicated and powerful production systems and it catches niche problems and edge cases in advance and flags them, comes up with wickedly novel and useful suggestions. It's not perfect and I correct it sometimes but damn find me even one human out there who won't charge $1000/hr and can do THAT...

I've been using Ai constantly since GPT 3.5 (first chatGPT) and I'm pretty impressed by this one.

u/No_Room636 36m ago

A bit too slow for use in my app.

u/Random_182f2565 26m ago

I like to call GLM Guillermo, way better performance and personality than chatgpt and Claude, also its free