DEV Community

#llm

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
LLM中如果一个问题容易验证 那么AI就容易学会解决!说说这个特性与P与NP问题的关联性

LLM中如果一个问题容易验证 那么AI就容易学会解决!说说这个特性与P与NP问题的关联性

Comments
1 min read
The Open-Weight Inflection Point: Kimi K3, Claude Opus 5, and Microsoft MAI Signal a Market Shift

The Open-Weight Inflection Point: Kimi K3, Claude Opus 5, and Microsoft MAI Signal a Market Shift

Comments
3 min read
Micro-compaction: amortizing context compression in agent loops

Micro-compaction: amortizing context compression in agent loops

1
Comments 3
5 min read
Kimi K3 is the largest open-weight model ever released — and you probably still can't run it

Kimi K3 is the largest open-weight model ever released — and you probably still can't run it

7
Comments
2 min read
Benchmarking GPT-4o, Claude 3.5 Sonnet, and Llama 3 for Automated Code Auditing & Vulnerability Detection

Benchmarking GPT-4o, Claude 3.5 Sonnet, and Llama 3 for Automated Code Auditing & Vulnerability Detection

Comments
2 min read
Same DeepSeek V4 Flash, Different Agent: Why the Runtime Changes the Result

Same DeepSeek V4 Flash, Different Agent: Why the Runtime Changes the Result

Comments
2 min read
How I Use DeepSeek V4 Flash: Reserve the Strongest Model for Uncertainty

How I Use DeepSeek V4 Flash: Reserve the Strongest Model for Uncertainty

Comments
2 min read
Building a podcast summarizer in 20 lines of Python

Building a podcast summarizer in 20 lines of Python

Comments
3 min read
Mind Maps in MCP: A Powerful Feature or Just Eye Candy?

Mind Maps in MCP: A Powerful Feature or Just Eye Candy?

1
Comments
5 min read
Stop Your AI Coding CLI From Wasting Tokens on "Hi" and "Thanks"

Stop Your AI Coding CLI From Wasting Tokens on "Hi" and "Thanks"

5
Comments 5
6 min read
A noisy judge does not just add error bars. It shrinks the effect you are trying to measure.

A noisy judge does not just add error bars. It shrinks the effect you are trying to measure.

1
Comments
11 min read
Stop building 'Chatbots' and start building Agentic Workflows

Stop building 'Chatbots' and start building Agentic Workflows

Comments
3 min read
How I Added Low-Token Vision to DeepSeek V4 Flash

How I Added Low-Token Vision to DeepSeek V4 Flash

Comments
6 min read
Why JSON Schema Field Order Breaks Structured Output Accuracy

Why JSON Schema Field Order Breaks Structured Output Accuracy

Comments
7 min read
Evolution of Language Models - Every LLM Breakthrough Was Just a Bug Fix

Evolution of Language Models - Every LLM Breakthrough Was Just a Bug Fix

1
Comments 1
8 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.