DEV Community

#gpu

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
From API to GPU, Week 2: What Actually Happens Behind the API

From API to GPU, Week 2: What Actually Happens Behind the API

Comments
29 min read
local-llm: A Field Report on Running SOTA Models on Your Own Hardware

local-llm: A Field Report on Running SOTA Models on Your Own Hardware

1
Comments
3 min read
One RTX 5090 vs a 12-GPU Cluster — Benchmarking a Decade of GPUs on the Same Go Proof

One RTX 5090 vs a 12-GPU Cluster — Benchmarking a Decade of GPUs on the Same Go Proof

Comments
4 min read
GPUs for AI in 2026: NVIDIA, AMD, Intel Compared

GPUs for AI in 2026: NVIDIA, AMD, Intel Compared

Comments
9 min read
WebGPU Explained: The Browser’s New Graphics and Compute Engine

WebGPU Explained: The Browser’s New Graphics and Compute Engine

12
Comments 2
11 min read
Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First

Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First

Comments
14 min read
Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali

Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali

Comments
3 min read
Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Comments
3 min read
Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips

Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips

Comments
2 min read
CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

1
Comments
4 min read
GPUs keep falling off the PCIe bus, and standard node health does not notice

GPUs keep falling off the PCIe bus, and standard node health does not notice

Comments 2
3 min read
The KV cache, why LLM inference is memory-bound, not compute-bound

The KV cache, why LLM inference is memory-bound, not compute-bound

Comments
4 min read
The Same RTX 5090, but the GPU Sat Idle — a CPU-Bound Go Solver and the Case for L2 Cache

The Same RTX 5090, but the GPU Sat Idle — a CPU-Bound Go Solver and the Case for L2 Cache

Comments
6 min read
The GPU Utilization Number That's Quietly Wrecking AI Team Budgets

The GPU Utilization Number That's Quietly Wrecking AI Team Budgets

Comments
5 min read
DGX Spark hitting 83 C under sustained Ollama load — solved by clock-locking via nvidia-smi -lgc

DGX Spark hitting 83 C under sustained Ollama load — solved by clock-locking via nvidia-smi -lgc

Comments
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.