3 ms·
GLM 5.3-flash fits the bill
by ThouYS 1mo ago
GLM 5.3-flash fits the bill
- andai 1mo agoWhat is it equivalent to? What kind of things are you using it for? I haven't tested it yet but on all the benchmarks it looks like it's 5-7x slower for agentic tasks.
- ThouYS 1mo agoI made some webapps with it, and have it running my hermes agent (which also does a lot of coding, but not webapps). Not sure what it's equivalent to, but it's super cheap and I am happy with the results
- glub 1mo agoIt's a mix of slightly worse kimi k3 for UI work and slightly smarter than luna for everything else. But yeah, it's very slow. I've put it to work as an LLM-as-RAG agent.
- andai 1mo agoI was wondering that, when DeepSeek became so cheap a while back, if it would be suitable as a superior embedding model. Although, RAG means search and search means latency?
- mark_l_watson 1mo agoand $0.15 1M input, $0.50 1M output I am a huge enthusiast of running local models, but when multiple quality USA vendors provide models like GLM 5.3-flash, I run locally just for the fun of it. For the purposes of comparing to Fable 5.1, I would mention GLM 5.3 that is about 1/12 the cost.