Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
gpu
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
Stop Comparing GPU Clouds Only by $/hour
GridPort
GridPort
GridPort
Follow
for
Highreso Co., Ltd.
Aug 24
Stop Comparing GPU Clouds Only by $/hour
#
gpu
#
cloud
#
machinelearning
#
llm
Comments
Add Comment
10 min read
My CUDA/GPU Journey: From "What Even Is a GPU?" to Actually Fascinated
Karthik Unnikrishnan
Karthik Unnikrishnan
Karthik Unnikrishnan
Follow
Aug 23
My CUDA/GPU Journey: From "What Even Is a GPU?" to Actually Fascinated
#
ai
#
gpu
#
nvidia
#
amd
1
 reaction
Comments
Add Comment
4 min read
I Got 28 TPS Out of Free Kaggle GPUs. Here's What It Took.
Aditya Raut
Aditya Raut
Aditya Raut
Follow
Aug 23
I Got 28 TPS Out of Free Kaggle GPUs. Here's What It Took.
#
ai
#
gpu
#
inference
#
python
Comments
Add Comment
5 min read
Choosing the Right GPU for Your Model — A Sizing Method, Not a Guess
Josef Doornink
Josef Doornink
Josef Doornink
Follow
Aug 19
Choosing the Right GPU for Your Model — A Sizing Method, Not a Guess
#
ai
#
gpu
#
llm
Comments
1
 comment
9 min read
What really fits in 8GB VRAM
the kilted dev
the kilted dev
the kilted dev
Follow
Aug 18
What really fits in 8GB VRAM
#
localllm
#
vram
#
gpu
#
buildinpublic
Comments
Add Comment
7 min read
I ran a GPU inference app for a month on Azure serverless GPU. Here's the actual bill.
aco dog
aco dog
aco dog
Follow
Aug 18
I ran a GPU inference app for a month on Azure serverless GPU. Here's the actual bill.
#
azure
#
gpu
#
cloud
#
devops
1
 reaction
Comments
Add Comment
4 min read
From API to GPU, Week 5: Tensors, the Data Structure Behind Every Model
Dinesh Kumar Ramasamy
Dinesh Kumar Ramasamy
Dinesh Kumar Ramasamy
Follow
Aug 17
From API to GPU, Week 5: Tensors, the Data Structure Behind Every Model
#
ai
#
llm
#
gpu
#
machinelearning
Comments
Add Comment
12 min read
Mastering Low-Precision AI: FP8 and FP4 Support Across Frameworks in Mid-2026
Dmitry Noranovich
Dmitry Noranovich
Dmitry Noranovich
Follow
Aug 13
Mastering Low-Precision AI: FP8 and FP4 Support Across Frameworks in Mid-2026
#
nvidia
#
gpu
#
ai
#
deeplearning
Comments
Add Comment
2 min read
Rust Portable SIMD Now Runs on the GPU and It Changes Everything About Cross-Platform Parallelism
Charles
Charles
Charles
Follow
Aug 13
Rust Portable SIMD Now Runs on the GPU and It Changes Everything About Cross-Platform Parallelism
#
rust
#
gpu
#
programming
#
performance
Comments
Add Comment
4 min read
Understanding GPU Memory: VRAM, Bandwidth, and Why Your Model Won't Fit
Aarush Karak
Aarush Karak
Aarush Karak
Follow
Aug 13
Understanding GPU Memory: VRAM, Bandwidth, and Why Your Model Won't Fit
#
gpu
#
cuda
#
vram
#
memory
1
 reaction
Comments
Add Comment
2 min read
GPU_WORKLOAD_MISMATCH Part II: From Detection to Runtime Enforcement for AI Infrastructure
Carnell Smith
Carnell Smith
Carnell Smith
Follow
Aug 16
GPU_WORKLOAD_MISMATCH Part II: From Detection to Runtime Enforcement for AI Infrastructure
#
ai
#
gpu
#
cybersecurity
#
machinelearning
1
 reaction
Comments
1
 comment
9 min read
Rust SIMD Just Came to the GPU — and It Changes How We Think About Parallel Programming
Charles
Charles
Charles
Follow
Aug 11
Rust SIMD Just Came to the GPU — and It Changes How We Think About Parallel Programming
#
rust
#
gpu
#
programming
#
performance
2
 reactions
Comments
Add Comment
4 min read
Renting GPUs for AI? Start with VRAM, Not the GPU
Kavya
Kavya
Kavya
Follow
Aug 5
Renting GPUs for AI? Start with VRAM, Not the GPU
#
machinelearning
#
llm
#
gpu
#
cloud
Comments
Add Comment
2 min read
Why memory bandwidth matters more than TFLOPS for LLM inference
Kavya
Kavya
Kavya
Follow
Aug 3
Why memory bandwidth matters more than TFLOPS for LLM inference
#
llm
#
gpu
#
nvidia
#
machinelearning
Comments
Add Comment
3 min read
KV Cache Quantization: I Stretched Qwen 35B's Context 8 on 12GB VRAM
Ken Imoto
Ken Imoto
Ken Imoto
Follow
Jul 28
KV Cache Quantization: I Stretched Qwen 35B's Context 8 on 12GB VRAM
#
llm
#
ai
#
gpu
#
performance
1
 reaction
Comments
1
 comment
3 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account