Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
llm
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
All Data and AI Weekly #224-12 Jan 2026
Timothy Spann
Timothy Spann
Timothy Spann
Follow
Jan 12
All Data and AI Weekly #224-12 Jan 2026
#
mcp
#
llm
#
snowflake
#
genai
5
 reactions
Comments
Add Comment
4 min read
The LLM Control Stack: From Words to Weights
Sai Srinivas
Sai Srinivas
Sai Srinivas
Follow
Jan 1
The LLM Control Stack: From Words to Weights
#
deeplearning
#
llm
#
promptengineering
#
ai
Comments
Add Comment
4 min read
LLMs Can Now Write GPU Kernels That Beat torch.compile
Jaber Jaber
Jaber Jaber
Jaber Jaber
Follow
Jan 23
LLMs Can Now Write GPU Kernels That Beat torch.compile
#
gpu
#
cuda
#
triton
#
llm
1
 reaction
Comments
Add Comment
7 min read
The Squeezing Effect: Why Your Aligned AI Model Gets Worse
Bilal Saeed
Bilal Saeed
Bilal Saeed
Follow
Dec 18 '25
The Squeezing Effect: Why Your Aligned AI Model Gets Worse
#
ai
#
llm
#
machinelearning
#
aiimplementation
Comments
Add Comment
3 min read
Developers Love Tools. AI Needs Better Instructions.
Mvtsahil (Sahil Khan)
Mvtsahil (Sahil Khan)
Mvtsahil (Sahil Khan)
Follow
Jan 23
Developers Love Tools. AI Needs Better Instructions.
#
ai
#
llm
#
productivity
#
softwaredevelopment
8
 reactions
Comments
Add Comment
3 min read
Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem
vishalmysore
vishalmysore
vishalmysore
Follow
Dec 18 '25
Optimal Chunking for Ontology RAG: Empirical Analysis & Orphan Axiom Problem
#
algorithms
#
rag
#
llm
#
ai
Comments
Add Comment
12 min read
How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications
Kuldeep Paul
Kuldeep Paul
Kuldeep Paul
Follow
Dec 19 '25
How to Build Multi-Provider Failover Strategies with Bifrost for Ultra‑Reliable AI Applications
#
ai
#
architecture
#
llm
5
 reactions
Comments
Add Comment
8 min read
Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%
Kuldeep Paul
Kuldeep Paul
Kuldeep Paul
Follow
Dec 19 '25
Semantic Caching with Bifrost: Reduce LLM Costs and Latency by Up to 70%
#
rag
#
performance
#
llm
#
ai
Comments
Add Comment
7 min read
The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember
osman uygar köse
osman uygar köse
osman uygar köse
Follow
Dec 30 '25
The Art of Context Windows: Our AI Had Alzheimer's: Here's How We Taught It To Remember
#
ai
#
llm
#
architecture
#
python
3
 reactions
Comments
Add Comment
9 min read
Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab
Tongyi Lab
Tongyi Lab
Tongyi Lab
Follow
Dec 19 '25
Dec 19, 2025 | The Tongyi Weekly: Your weekly dose of cutting-edge AI from Tongyi Lab
#
ai
#
opensource
#
llm
#
genai
Comments
Add Comment
5 min read
📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌
Prashant Lakhera
Prashant Lakhera
Prashant Lakhera
Follow
Dec 19 '25
📌 Most models use Grouped Query Attention. That doesn’t mean yours should.📌
#
llm
#
chatgpt
#
ai
#
deepseek
1
 reaction
Comments
Add Comment
1 min read
Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents
Emma Wilson
Emma Wilson
Emma Wilson
Follow
Jan 23
Comparative Cost & ROI: Chatbots vs LLM Integrations vs Autonomous Agents
#
agents
#
ai
#
llm
#
productivity
Comments
Add Comment
5 min read
Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis
Sara_T
Sara_T
Sara_T
Follow
Dec 18 '25
Mooncake Memory Deep Dive: KVCache, Token Cost, DRAM Usage, and Saturation Analysis
#
performance
#
llm
#
backend
#
ai
Comments
Add Comment
5 min read
A Deep Dive into Deep Agent Architecture for AI Coding Assistants
Alexsandro Souza
Alexsandro Souza
Alexsandro Souza
Follow
Jan 22
A Deep Dive into Deep Agent Architecture for AI Coding Assistants
#
agents
#
ai
#
architecture
#
llm
4
 reactions
Comments
1
 comment
16 min read
Part 1: Why Transformers Still Forget
Pranava Kailash Subramaniam Prema
Pranava Kailash Subramaniam Prema
Pranava Kailash Subramaniam Prema
Follow
Dec 18 '25
Part 1: Why Transformers Still Forget
#
ai
#
llm
#
productivity
Comments
Add Comment
5 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account