DEV Community

#llm

Posts

๐Ÿ‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
I Scaled PHP Until It Broke. Three llama.cpp Patterns Saved It.

I Scaled PHP Until It Broke. Three llama.cpp Patterns Saved It.

Comments 1
10 min read
Evaluating LLM code reviewers: an offline harness for precision, recall, and routing"

Evaluating LLM code reviewers: an offline harness for precision, recall, and routing"

2
Comments
5 min read
Why JSON.parse() Fails Silently on Truncated LLM Responses (And What I Did About It)

Why JSON.parse() Fails Silently on Truncated LLM Responses (And What I Did About It)

2
Comments 2
3 min read
I built an Agent Memory System for myself and got 90.8% (end-to-end) on LongMemEval

I built an Agent Memory System for myself and got 90.8% (end-to-end) on LongMemEval

1
Comments 1
7 min read
AI Weekly: 4/1โ€“4/10 | Anthropic Triple Shock Sequel โ€” Mythos Too Dangerous to Ship, Revenue Passes OpenAI, Software Stocks Crash

AI Weekly: 4/1โ€“4/10 | Anthropic Triple Shock Sequel โ€” Mythos Too Dangerous to Ship, Revenue Passes OpenAI, Software Stocks Crash

Comments
9 min read
Your AI agent is the new attack vector. It just wants to help.

Your AI agent is the new attack vector. It just wants to help.

Comments
3 min read
Your AI agent isn't broken. It's confidently wrong. Here's the difference

Your AI agent isn't broken. It's confidently wrong. Here's the difference

Comments
4 min read
Embeddings Just Went Multimodal: What Sentence Transformers 5.4 Means for RAG

Embeddings Just Went Multimodal: What Sentence Transformers 5.4 Means for RAG

Comments
2 min read
SLM vs LLM: How to Pick the Right Model for Your Enterprise Workload

SLM vs LLM: How to Pick the Right Model for Your Enterprise Workload

Comments
1 min read
My AI Remembers Its Mistakes. Permanently. Here's the Engineering.

My AI Remembers Its Mistakes. Permanently. Here's the Engineering.

Comments 1
13 min read
The Information Design Gap: Why Our AI Agent Was Coding Blind

The Information Design Gap: Why Our AI Agent Was Coding Blind

Comments
9 min read
๐—ช๐—ต๐—ฎ๐˜ ๐—œ ๐—Ÿ๐—ฒ๐—ฎ๐—ฟ๐—ป๐—ฒ๐—ฑ ๐—ณ๐—ฟ๐—ผ๐—บ ๐—–๐—ต๐—ฎ๐—ฝ๐˜๐—ฒ๐—ฟ ๐Ÿฎ ๐—ผ๐—ณ ๐—”๐—œ ๐—˜๐—ป๐—ด๐—ถ๐—ป๐—ฒ๐—ฒ๐—ฟ๐—ถ๐—ป๐—ด: ๐—ช๐—ต๐˜† ๐—ฆ๐—ฎ๐—บ๐—ฝ๐—น๐—ถ๐—ป๐—ด ๐—–๐—ต๐—ฎ๐—ป๐—ด๐—ฒ๐˜€ ๐—˜๐˜ƒ๐—ฒ๐—ฟ๐˜†๐˜๐—ต๐—ถ๐—ป๐—ด

๐—ช๐—ต๐—ฎ๐˜ ๐—œ ๐—Ÿ๐—ฒ๐—ฎ๐—ฟ๐—ป๐—ฒ๐—ฑ ๐—ณ๐—ฟ๐—ผ๐—บ ๐—–๐—ต๐—ฎ๐—ฝ๐˜๐—ฒ๐—ฟ ๐Ÿฎ ๐—ผ๐—ณ ๐—”๐—œ ๐—˜๐—ป๐—ด๐—ถ๐—ป๐—ฒ๐—ฒ๐—ฟ๐—ถ๐—ป๐—ด: ๐—ช๐—ต๐˜† ๐—ฆ๐—ฎ๐—บ๐—ฝ๐—น๐—ถ๐—ป๐—ด ๐—–๐—ต๐—ฎ๐—ป๐—ด๐—ฒ๐˜€ ๐—˜๐˜ƒ๐—ฒ๐—ฟ๐˜†๐˜๐—ต๐—ถ๐—ป๐—ด

1
Comments
4 min read
Anthropic Just Did Something Unprecedented: They Hid Their Best Security Model

Anthropic Just Did Something Unprecedented: They Hid Their Best Security Model

Comments
2 min read
Marker, hosted: a scientific PDF parser API with LaTeX equations preserved

Marker, hosted: a scientific PDF parser API with LaTeX equations preserved

Comments
4 min read
RedSOC: Open-source framework to benchmark adversarial attacks on AI-powered SOCs โ€” 100% detection rate across 15 attack scenarios [paper + code]

RedSOC: Open-source framework to benchmark adversarial attacks on AI-powered SOCs โ€” 100% detection rate across 15 attack scenarios [paper + code]

Comments
2 min read
๐Ÿ‘‹ Sign in for the ability to sort posts by relevant, latest, or top.