DEV Community

#gpu

Posts

👋 Sign in for the ability to sort posts by relevant, latest, or top.
PCIe lanes, not GPU VRAM, is the spec that kills your homelab GPU plan

PCIe lanes, not GPU VRAM, is the spec that kills your homelab GPU plan

Comments
5 min read
Kubernetes SR-IOV Multi-Rail GPU Networking: Missing Route Configuration for Cross-Rail Communication

Kubernetes SR-IOV Multi-Rail GPU Networking: Missing Route Configuration for Cross-Rail Communication

Comments
11 min read
How to read a Blender cloud render benchmark: GPU seconds are not the round trip

How to read a Blender cloud render benchmark: GPU seconds are not the round trip

Comments
4 min read
Vast.ai CLI 101: Finding a GPU Offer You Can Actually Use

Vast.ai CLI 101: Finding a GPU Offer You Can Actually Use

Comments
7 min read
From Naive CUDA to Performance Engineering: My First GPU Matmul Journey

From Naive CUDA to Performance Engineering: My First GPU Matmul Journey

Comments
3 min read
Your Apple Silicon GPU Loses to One CPU Core Until a Million Rows. I Measured 111 Operations, Then Rebuilt ArrowMetal 0.2.0 Around the Answer

Your Apple Silicon GPU Loses to One CPU Core Until a Million Rows. I Measured 111 Operations, Then Rebuilt ArrowMetal 0.2.0 Around the Answer

Comments
10 min read
Before You Rent a GPU Server, Check These 10 Things

Before You Rent a GPU Server, Check These 10 Things

Comments
3 min read
How to Use Your NVIDIA GPU for Local AI on Linux in 2026

How to Use Your NVIDIA GPU for Local AI on Linux in 2026

Comments
5 min read
DataFusion on the Apple Silicon GPU: One Optimizer Rule, the Same SQL, Sorts 6.9x to 28.8x Faster

DataFusion on the Apple Silicon GPU: One Optimizer Rule, the Same SQL, Sorts 6.9x to 28.8x Faster

1
Comments
9 min read
Same Source, Two Cost Curves: Local Preview vs GPU Final

Same Source, Two Cost Curves: Local Preview vs GPU Final

Comments
5 min read
Let your AI agent rent a GPU: llms.txt, --json and --budget

Let your AI agent rent a GPU: llms.txt, --json and --budget

1
Comments 1
4 min read
Evaluating Multi-Node LLM Orchestrators: Exo, GPUStack, and LocalAI

Evaluating Multi-Node LLM Orchestrators: Exo, GPUStack, and LocalAI

Comments
1 min read
VRAM for local LLMs: why memory bandwidth sets your tokens per second

VRAM for local LLMs: why memory bandwidth sets your tokens per second

2
Comments 2
8 min read
How Much VRAM Do You Really Need to Run a 70B LLM?

How Much VRAM Do You Really Need to Run a 70B LLM?

Comments
7 min read
Exploring Result Visibility of Fixed-Latency Instructions on the SM120 Architecture

Exploring Result Visibility of Fixed-Latency Instructions on the SM120 Architecture

1
Comments
10 min read
👋 Sign in for the ability to sort posts by relevant, latest, or top.