四大模型对比
- Dominant language
- JavaScript
- Stars
- 1
- Forks
- 0
- PR merge metrics
- No merged PRs in 30d
Description
### Gemini3
Gemini3 pro 写代码准确性和推理能力最强,棘手的问题让 gemini3pro 处理,常规问题用 gemini3 flash 即可。
gemini3 flash 这一档是 ai coding 的门槛,这个门槛以下的模型都不好用。
### GLM5
GLM 4.7 很糟糕,GLM5 可用,编程体验和 gemini3 flash 在同一档次,但处于下位,无法超越。最大的问题是速度慢且不支持多模态。纯写代码是没问题。包含在百炼的编程套餐里,使用性价比高。可作为编程主力使用。
### QWen 3.5
Qwen3 代码能力很糟,升级 Qwen3.5 之后提升幅度巨大,编程水平略小于GLM5,但优势是速度快、且支持多模态,这两点优势巨大。和GLM5混搭使用。跟 GLM5 一样都跟 gemini3 flash 处于同水平。
### KIMI k2.5
推理能力很强,但长思考很啰嗦,有失忆的情况,必须要配备一个精心编写的 CLAUDE.md 才行,这个需要慢慢调。跟 claude code 做工程时的体验不如 QWen 和 GLM,需要沉淀一下最佳实践。
### deepseek v4
坐等v4发布,很期待
--------------------
- 一梯队:claude opus 4.5 >= gemini3pro
- 二梯队:gemini3 flash > GLM5/Qwen3.5 >= Kimi k2.5 >= claude sonnet 4.5
- 三梯队:MiniMax/GLM4.7 > DeepSeek v3.2 > QWen3-coder
Contributor guide
No contributing guide indexed for this repository
Research direction
No file, test, or entry point is mentioned. Review the repository's existing content structure and determine where this model comparison belongs; done should include a clearly defined scope for the comparison and an agreed location for maintaining it.
Written by the indexing model from the issue text.
Assessment
- Domain
- content
- Issue type
- Documentation
- Difficulty
- 2/5
- Estimated time
- 1-3 hours
- Activity status
- Stale
- Clarity
- Needs clarification
- Newbie friendliness
- 25/100