Skip to content
Navigation menu
Search
Powered by Algolia
Search
Log in
Create account
DEV Community
Close
#
gpu
Follow
Hide
Posts
Left menu
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
Right menu
From API to GPU, Week 2: What Actually Happens Behind the API
Dinesh Kumar Ramasamy
Dinesh Kumar Ramasamy
Dinesh Kumar Ramasamy
Follow
Jul 19
From API to GPU, Week 2: What Actually Happens Behind the API
#
ai
#
llm
#
gpu
#
machinelearning
Comments
Add Comment
29 min read
local-llm: A Field Report on Running SOTA Models on Your Own Hardware
Reno Lu
Reno Lu
Reno Lu
Follow
Jul 20
local-llm: A Field Report on Running SOTA Models on Your Own Hardware
#
localllm
#
gpu
#
selfhosting
#
inference
1
 reaction
Comments
Add Comment
3 min read
One RTX 5090 vs a 12-GPU Cluster — Benchmarking a Decade of GPUs on the Same Go Proof
soy
soy
soy
Follow
Jul 18
One RTX 5090 vs a 12-GPU Cluster — Benchmarking a Decade of GPUs on the Same Go Proof
#
gpu
#
benchmark
#
machinelearning
#
cuda
Comments
Add Comment
4 min read
GPUs for AI in 2026: NVIDIA, AMD, Intel Compared
Rost
Rost
Rost
Follow
Jul 14
GPUs for AI in 2026: NVIDIA, AMD, Intel Compared
#
gpu
#
ai
#
nvidia
#
hardware
Comments
Add Comment
9 min read
WebGPU Explained: The Browser’s New Graphics and Compute Engine
Pratik sharma
Pratik sharma
Pratik sharma
Follow
Jul 17
WebGPU Explained: The Browser’s New Graphics and Compute Engine
#
webgpu
#
webdev
#
javascript
#
gpu
12
 reactions
Comments
2
 comments
11 min read
Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First
Abdollah Ebadi
Abdollah Ebadi
Abdollah Ebadi
Follow
Jul 11
Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First
#
comfyui
#
python
#
machinelearning
#
gpu
Comments
Add Comment
14 min read
Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali
soy
soy
soy
Follow
Jul 11
Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali
#
gpu
#
nvidia
#
hardware
Comments
Add Comment
3 min read
Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)
Arsen Apostolov
Arsen Apostolov
Arsen Apostolov
Follow
Jul 9
Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)
#
llm
#
ollama
#
vllm
#
gpu
Comments
Add Comment
3 min read
Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips
circuitrocks
circuitrocks
circuitrocks
Follow
Jul 8
Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips
#
riscv
#
microcontrollers
#
gpu
#
pcbdesign
Comments
Add Comment
2 min read
CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?
uttesh
uttesh
uttesh
Follow
Jul 8
CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?
#
ai
#
llm
#
gpu
#
beginners
1
 reaction
Comments
Add Comment
4 min read
GPUs keep falling off the PCIe bus, and standard node health does not notice
Leo
Leo
Leo
Follow
Jul 20
GPUs keep falling off the PCIe bus, and standard node health does not notice
#
kubernetes
#
eks
#
gpu
#
selfhealing
Comments
2
 comments
3 min read
The KV cache, why LLM inference is memory-bound, not compute-bound
I Want To Learn Programming
I Want To Learn Programming
I Want To Learn Programming
Follow
Jul 4
The KV cache, why LLM inference is memory-bound, not compute-bound
#
gpu
#
llm
#
inference
#
performance
Comments
Add Comment
4 min read
The Same RTX 5090, but the GPU Sat Idle — a CPU-Bound Go Solver and the Case for L2 Cache
soy
soy
soy
Follow
Jul 18
The Same RTX 5090, but the GPU Sat Idle — a CPU-Bound Go Solver and the Case for L2 Cache
#
cpu
#
gpu
#
benchmark
#
hardware
Comments
Add Comment
6 min read
The GPU Utilization Number That's Quietly Wrecking AI Team Budgets
Mike Smith
Mike Smith
Mike Smith
Follow
for
Hostrunway
Jul 3
The GPU Utilization Number That's Quietly Wrecking AI Team Budgets
#
gpu
#
ai
#
machinelearning
#
deeplearning
Comments
Add Comment
5 min read
DGX Spark hitting 83 C under sustained Ollama load — solved by clock-locking via nvidia-smi -lgc
Deal Estate
Deal Estate
Deal Estate
Follow
Jul 1
DGX Spark hitting 83 C under sustained Ollama load — solved by clock-locking via nvidia-smi -lgc
#
nvidia
#
gpu
#
llm
#
ai
Comments
Add Comment
3 min read
đź‘‹
Sign in
for the ability to sort posts by
relevant
,
latest
, or
top
.
We're a place where coders share, stay up-to-date and grow their careers.
Log in
Create account