DEV Community

#nvidia

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
RTX 5090 Cooling, BeeLlama VRAM Opts, Resizable BAR Performance Gains

RTX 5090 Cooling, BeeLlama VRAM Opts, Resizable BAR Performance Gains

1
Comments
4 min read
I Built a Prompt Forging Engine with NVIDIA Llama 3.1 70B for Google Prompt Wars 2026

I Built a Prompt Forging Engine with NVIDIA Llama 3.1 70B for Google Prompt Wars 2026

3
Comments 2
3 min read
LLM Compilers, GGUF Quantization, & Radeon RX 9060 Benchmarks

LLM Compilers, GGUF Quantization, & Radeon RX 9060 Benchmarks

Comments
3 min read
Go+CUDA Optimization, LLM VRAM Benchmarks & NVIDIA G-SYNC Firmware 1.1.6

Go+CUDA Optimization, LLM VRAM Benchmarks & NVIDIA G-SYNC Firmware 1.1.6

2
Comments
3 min read
Intel Xe3P Leaks 160GB LPDDR5X; FlashAttention-2 in CuTe & Custom CUDA GPT-2 Engine

Intel Xe3P Leaks 160GB LPDDR5X; FlashAttention-2 in CuTe & Custom CUDA GPT-2 Engine

Comments
3 min read
Nvidia wants enterprises to run agents safely. NemoClaw is how.

Nvidia wants enterprises to run agents safely. NemoClaw is how.

Comments
3 min read
⚡️Self-Hosting Experience with Jetson Orin Nano and Ollama 🦙

⚡️Self-Hosting Experience with Jetson Orin Nano and Ollama 🦙

11
Comments
8 min read
'"An LLM and a harness": Nvidia''s simple thesis on what agents actually are'

'"An LLM and a harness": Nvidia''s simple thesis on what agents actually are'

Comments 1
2 min read
GPU Bottleneck Analyzer, NVIDIA Rubin VRAM Demands, and Qwen VRAM Optimization

GPU Bottleneck Analyzer, NVIDIA Rubin VRAM Demands, and Qwen VRAM Optimization

1
Comments
4 min read
nvidia-peermem "Invalid argument" on Ubuntu — Fix GPUDirect RDMA with DMA-BUF

nvidia-peermem "Invalid argument" on Ubuntu — Fix GPUDirect RDMA with DMA-BUF

3
Comments
3 min read
GPU Hardware & Driver Update: RTX 5090 Benchmarks, llama.cpp MTP, Windows 11 Fix

GPU Hardware & Driver Update: RTX 5090 Benchmarks, llama.cpp MTP, Windows 11 Fix

Comments
3 min read
Bypassing the OS to Run LLMs: What I Learned Building a Firmware-Centric Runtime

Bypassing the OS to Run LLMs: What I Learned Building a Firmware-Centric Runtime

Comments
5 min read
CUDA Cutile-rs Beta, AMD FSR 4.1 Release, & Forza Horizon 6 GPU Benchmarks

CUDA Cutile-rs Beta, AMD FSR 4.1 Release, & Forza Horizon 6 GPU Benchmarks

Comments
3 min read
Two Qwen3 Models on One DGX Spark: The Residency Math for Local LLM Coding

Two Qwen3 Models on One DGX Spark: The Residency Math for Local LLM Coding

1
Comments
5 min read
Stop Paying $20/month. Use NVIDIA Build: 80+ Free AI Models

Stop Paying $20/month. Use NVIDIA Build: 80+ Free AI Models

25
Comments 5
2 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.