DEV Community

ninghonggang
ninghonggang

Posted on

Two Juejin reviews of AI coding tools, two top-eights that share one name

I went down a rabbit hole this evening reading two Juejin roundups from the same two weeks of December 2025 that both purport to rank the top AI coding tools for 2026, and what crystallized for me is that they are not even close to ranking the same universe. One prints Cursor at S档 and Claude Code and Codex CLI at A档 on a tier letter, with v0 and Lovable above Rork and VibeCode App. The other prints Tencent CodeBuddy at 9.6 out of 10 on a five-axis decimal scorecard, with Cody at 8.2, Ghostwriter at 8.0, Codeium at 7.8, Tabnine at 7.6, CodeWhisperer at 7.5, JetBrains AI Assistant at 7.4, and Blackbox at 7.2. Same reader-question, same search results page, two posts, two completely disjoint top-eight rosters, and not a single shared reference across the two columns.

The angle I want to put down is that the convergent-consensus I have been tracking in prior pieces has now hit its opposite case, and the opposite case is what the convergent-consensus has been hiding. The tier-letter piece sorts on chat-plus-tool-call capability plus terminal evidence plus "whether a non-technical user can ship with it", which is why Cursor and Lovable and v0 crowd the top and Tencent CodeBuddy never appears. The decimal piece sorts on code completion plus code generation plus 自主代理 plus 多模态 plus team collaboration plus 安全合规, which is why Tencent CodeBuddy's enterprise tier plus 等保三级 row lands it at 9.6 and Cursor never appears. To be fair I would take the exact 9.6 decimal and the exact tier letters with a grain of salt because the test corpus behind each is never disclosed, but the structural tell is that the two formats have not just used different scorecards — they have used different scoring apparatuses that produce different verdicts on different tools.

The meta-pattern I want to write down is that the December 2025 search results page for "AI coding tools 2026" is now producing, in the same two-week window, two posts whose top-eight rosters share exactly one name — Replit. The tier-letter piece names Cursor, Claude Code, Codex CLI, Lovable, v0, Rork, VibeCode App, Anything, Chef, and Replit in roughly that order; the decimal piece names Tencent CodeBuddy, Cody, Ghostwriter, Codeium, Tabnine, CodeWhisperer, JetBrains AI Assistant, and Blackbox with no overlap except Replit on both. The convergent-consensus I have been writing about is visibly not converging — one post says the verdict is Cursor plus Claude Code, the other says the verdict is Tencent CodeBuddy plus Cody, and neither knows about the other. Honestly I am a little skeptical of any 2026 picking workflow that treats one of these pieces as the answer, because each was written to land on its own winner by sorting on axes that ignore the other piece's top picks entirely.

The practical takeaway I want to put down is that I now read the December two-format standoff the way I used to read the four-camp pieces — as two separate ranking dialects sorting on two non-overlapping definitions of what counts as a coding tool, and the cross-format translation is the engineering work the reader has to do by hand. The tier-letter piece answers the within-capability question including non-coder use cases like Lovable and v0 and Rork, and the decimal piece answers the within-IDE question for an enterprise developer inside VS Code or JetBrains with compliance requirements. My gut says the reader has to combine both pieces to assemble a 2026 stack and not pick one as the answer, because Cursor and Claude Code are at the top of one column and absent from the other, and Tencent CodeBuddy is at the top of the other column and absent from the first, and the two recommendations are not in conflict — they are answering different reader-questions.

I will reassess in three months. For now I am still mostly on Cursor Pro plus Claude Code for coding and ChatGPT Plus for general chat, and I am keeping a per-piece log of which December 2025 roundup named which tool in which tier, because the verdict apparatus has now visibly bifurcated inside a single search week and the search results page is not going to surface this bifurcation by itself. Give it six months and I expect either one of the two formats to add a row for the other's top picks, or a third camp to spin up that publishes a cross-format compatibility column listing which tools appear in which tier and which scorecard, and whichever moves first will tell me whether the December two-format standoff has finally collapsed into a cross-format verdict or whether the picking workflow has permanently turned into a multi-format integration job the engineer has to do by hand.

Top comments (0)