DEV Community

#llm

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Plain Japanese in, ComfyUI workflow out

Plain Japanese in, ComfyUI workflow out

Comments
6 min read
Cache hits, claude.md, and angry illustrators

Cache hits, claude.md, and angry illustrators

Comments
6 min read
Copy the template, don't generate from zero

Copy the template, don't generate from zero

Comments
6 min read
Xiaomi AI training, live: what the MiMo 2.6 RL dashboard shows

Xiaomi AI training, live: what the MiMo 2.6 RL dashboard shows

1
Comments 1
9 min read
Japan's game studios: 80% already use gen AI

Japan's game studios: 80% already use gen AI

Comments
6 min read
My LLM Eval Reused an Old Result After the Input Changed

My LLM Eval Reused an Old Result After the Input Changed

5
Comments 1
4 min read
Darwin-180B-RSI Tops Global Legal Reasoning Benchmarks — Without Ever Training on Legal Data

Darwin-180B-RSI Tops Global Legal Reasoning Benchmarks — Without Ever Training on Legal Data

1
Comments
4 min read
Ollama Connection Refused? The 60-Second Triage

Ollama Connection Refused? The 60-Second Triage

5
Comments 1
4 min read
Diseñar agentes LLM con contratos verificables

Diseñar agentes LLM con contratos verificables

Comments
5 min read
yoDEV Decisions: compara Jev con GPT, Gemini, Claude o cualquier LLM con tus propios datos

yoDEV Decisions: compara Jev con GPT, Gemini, Claude o cualquier LLM con tus propios datos

Comments
4 min read
PhantomEnvironments: How Fictional Worlds Solve the Agent Training Bottleneck

PhantomEnvironments: How Fictional Worlds Solve the Agent Training Bottleneck

3
Comments
5 min read
What agent frameworks cost on the wire: measurements from agentic-arena

What agent frameworks cost on the wire: measurements from agentic-arena

Comments 1
2 min read
When Both Answers Are Defensible, the Right Answer Is That It Is a Coin Flip

When Both Answers Are Defensible, the Right Answer Is That It Is a Coin Flip

Comments
7 min read
Judge cheap, audit confidence: a CI gate for LLM evals (open source and measured)

Judge cheap, audit confidence: a CI gate for LLM evals (open source and measured)

Comments
3 min read
Lost in the Middle: My RAG Found the Right Chunk and the LLM Ignored It

Lost in the Middle: My RAG Found the Right Chunk and the LLM Ignored It

Comments
7 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.