Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
nvidia
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Why We're Stuck With GPUs This Long?
Rooted
Rooted
Rooted
Follow
Jul 5
Why We're Stuck With GPUs This Long?
#
ai
#
llm
#
nvidia
#
startup
1
 reaction
Comments
Add Comment
5 min read
Notes on Serving LLMs with TensorRT-LLM and Triton
member_2e5ba30f
member_2e5ba30f
member_2e5ba30f
Follow
May 31
Notes on Serving LLMs with TensorRT-LLM and Triton
#
llm
#
nvidia
#
machinelearning
#
performance
Comments
Add Comment
4 min read
The dozen layers under a GPU pod
Harshit Luthra
Harshit Luthra
Harshit Luthra
Follow
Jul 2
The dozen layers under a GPU pod
#
gpu
#
kubernetes
#
mlops
#
nvidia
Comments
Add Comment
14 min read
NVIDIA Nemotron 3 Ultra & GLM-5.2: The Open Model Flood Is Here (June 2026)
DoremonAI
DoremonAI
DoremonAI
Follow
Jun 30
NVIDIA Nemotron 3 Ultra & GLM-5.2: The Open Model Flood Is Here (June 2026)
#
ai
#
opensource
#
machinelearning
#
nvidia
1
 reaction
Comments
Add Comment
2 min read
Things I learned building my first multi-agent AI system on Azure + NVIDIA
Sachin Magon
Sachin Magon
Sachin Magon
Follow
Jun 29
Things I learned building my first multi-agent AI system on Azure + NVIDIA
#
ai
#
azure
#
nvidia
#
python
1
 reaction
Comments
2
 comments
4 min read
Diffusion Language Models Are Here: Deep Dive into NVIDIA's Nemotron-Labs DLM Architecture
Manoranjan Rajguru
Manoranjan Rajguru
Manoranjan Rajguru
Follow
May 24
Diffusion Language Models Are Here: Deep Dive into NVIDIA's Nemotron-Labs DLM Architecture
#
ai
#
machinelearning
#
llm
#
nvidia
Comments
Add Comment
15 min read
I Built a Prompt Forging Engine with NVIDIA Llama 3.1 70B for Google Prompt Wars 2026
Divyansh
Divyansh
Divyansh
Follow
Jun 24
I Built a Prompt Forging Engine with NVIDIA Llama 3.1 70B for Google Prompt Wars 2026
#
promptengineering
#
nvidia
#
python
#
webdev
3
 reactions
Comments
2
 comments
3 min read
Nvidia wants enterprises to run agents safely. NemoClaw is how.
Andrew Kew
Andrew Kew
Andrew Kew
Follow
Jun 22
Nvidia wants enterprises to run agents safely. NemoClaw is how.
#
ai
#
agents
#
nvidia
#
devops
Comments
Add Comment
3 min read
⚡️Self-Hosting Experience with Jetson Orin Nano and Ollama 🦙
Anna Villarreal
Anna Villarreal
Anna Villarreal
Follow
Jun 20
⚡️Self-Hosting Experience with Jetson Orin Nano and Ollama 🦙
#
ubuntu
#
nvidia
#
webdev
#
ai
11
 reactions
Comments
Add Comment
8 min read
'"An LLM and a harness": Nvidia''s simple thesis on what agents actually are'
Andrew Kew
Andrew Kew
Andrew Kew
Follow
Jun 22
'"An LLM and a harness": Nvidia''s simple thesis on what agents actually are'
#
ai
#
llm
#
agents
#
nvidia
Comments
1
 comment
2 min read
nvidia-peermem "Invalid argument" on Ubuntu — Fix GPUDirect RDMA with DMA-BUF
Fabricio
Fabricio
Fabricio
Follow
Jun 21
nvidia-peermem "Invalid argument" on Ubuntu — Fix GPUDirect RDMA with DMA-BUF
#
nvidia
#
nvidiapeermem
#
ai
#
hpc
3
 reactions
Comments
Add Comment
3 min read
Bypassing the OS to Run LLMs: What I Learned Building a Firmware-Centric Runtime
Rohan R
Rohan R
Rohan R
Follow
Jun 21
Bypassing the OS to Run LLMs: What I Learned Building a Firmware-Centric Runtime
#
llm
#
nvidia
#
cuda
#
firmware
Comments
Add Comment
5 min read
12B Gemma 4 QAT Deployment with GCE, NVIDIA L4, MCP, and Antigravity CLI
xbill
xbill
xbill
Follow
for
Google Developer Experts
Jun 16
12B Gemma 4 QAT Deployment with GCE, NVIDIA L4, MCP, and Antigravity CLI
#
mcps
#
qat
#
gemma4
#
nvidia
6
 reactions
Comments
Add Comment
16 min read
Two Qwen3 Models on One DGX Spark: The Residency Math for Local LLM Coding
Devashish
Devashish
Devashish
Follow
Jun 16
Two Qwen3 Models on One DGX Spark: The Residency Math for Local LLM Coding
#
localllm
#
vllm
#
ai
#
nvidia
1
 reaction
Comments
Add Comment
5 min read
Stop Paying $20/month. Use NVIDIA Build: 80+ Free AI Models
Lorenzo Zarantonello
Lorenzo Zarantonello
Lorenzo Zarantonello
Follow
Jun 17
Stop Paying $20/month. Use NVIDIA Build: 80+ Free AI Models
#
ai
#
programming
#
nvidia
#
llm
28
 reactions
Comments
5
 comments
2 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account