DEV Community

#gpu

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
A ledger that only bills the dead

A ledger that only bills the dead

Comments
11 min read
KV Cache Quantization: I Stretched Qwen 35B's Context 8 on 12GB VRAM

KV Cache Quantization: I Stretched Qwen 35B's Context 8 on 12GB VRAM

1
Comments 1
3 min read
GPU Monitoring & Metrics for MLOps

GPU Monitoring & Metrics for MLOps

Comments
1 min read
The 60% idle GPU that turned out to be a network policy

The 60% idle GPU that turned out to be a network policy

Comments
3 min read
Accidentally quadratic: buffer copies made MCTS in DeepMind's mctx 3 slower

Accidentally quadratic: buffer copies made MCTS in DeepMind's mctx 3 slower

1
Comments
6 min read
Building CI/CD Pipelines for GPU Validation

Building CI/CD Pipelines for GPU Validation

1
Comments
10 min read
From API to GPU, Week 2: What Actually Happens Behind the API

From API to GPU, Week 2: What Actually Happens Behind the API

Comments
29 min read
GPUs for AI in 2026: NVIDIA, AMD, Intel Compared

GPUs for AI in 2026: NVIDIA, AMD, Intel Compared

Comments
9 min read
WebGPU Explained: The Browser’s New Graphics and Compute Engine

WebGPU Explained: The Browser’s New Graphics and Compute Engine

12
Comments 2
11 min read
Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First

Running Multiple ComfyUI Instances in Parallel on a Single GPU — What Actually Breaks First

Comments
14 min read
Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali

Linux 7.2 Improves Multi-GPU Displays, M3 Support, Mesa Rusticl Defaults Arm Mali

Comments
3 min read
Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Does a Second GPU Increase Ollama's Context Window? (Quadro P2000 + RTX 3090 Tested)

Comments
3 min read
Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips

Bitluni's 8,192-Core DIY GPU Is Built From 13-Cent RISC-V Chips

Comments
2 min read
CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

CPU vs GPU: Why Large Language Models Need GPUs — What Really Happens After You Press Enter?

1
Comments
4 min read
local-llm: A Field Report on Running SOTA Models on Your Own Hardware

local-llm: A Field Report on Running SOTA Models on Your Own Hardware

1
Comments 1
3 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.