DEV Community

#gpu

Posts

đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.
Choosing TTS Based on Sound Quality Was Too Slow for Conversations — Separating 'Design' and 'Production' with a Measured 2.5x RTF Difference

Choosing TTS Based on Sound Quality Was Too Slow for Conversations — Separating 'Design' and 'Production' with a Measured 2.5x RTF Difference

Comments
6 min read
How Many AI Avatars Can One GPU Handle? Real-World Test Reveals 4 Avatars at ÂĄ7,600 Each per Month

How Many AI Avatars Can One GPU Handle? Real-World Test Reveals 4 Avatars at ÂĄ7,600 Each per Month

Comments
9 min read
I Tried Getting Closer to the GPU With Triton

I Tried Getting Closer to the GPU With Triton

1
Comments
6 min read
Why Chromium Was Ignoring My GPU — And How I Boosted Performance from 4fps to 58fps

Why Chromium Was Ignoring My GPU — And How I Boosted Performance from 4fps to 58fps

Comments
10 min read
Borrowed an H100 but couldn't draw a single frame — Why compute GPUs and rendering GPUs are different beasts

Borrowed an H100 but couldn't draw a single frame — Why compute GPUs and rendering GPUs are different beasts

Comments
6 min read
Pure JAX on G5g: Serving Gemma 4 on Graviton and a T4G

Pure JAX on G5g: Serving Gemma 4 on Graviton and a T4G

Comments
8 min read
I Created a 24/7 AI Avatar That Streams Without Human Intervention — Only 'Verification,' 'Eyes,' and 'Ears' Remain for Humans

I Created a 24/7 AI Avatar That Streams Without Human Intervention — Only 'Verification,' 'Eyes,' and 'Ears' Remain for Humans

1
Comments
9 min read
Latest Trends in GPU Cloud Cost Reduction and Containerized Data Centers

Latest Trends in GPU Cloud Cost Reduction and Containerized Data Centers

Comments
2 min read
Most Kubernetes clusters can't tell you what their GPUs cost

Most Kubernetes clusters can't tell you what their GPUs cost

Comments
4 min read
My inference server decided my second GPU no longer exists. Here is how I got it back without upgrading a driver.

My inference server decided my second GPU no longer exists. Here is how I got it back without upgrading a driver.

Comments
3 min read
I Got 28 TPS Out of Free Kaggle GPUs. Here's What It Took.

I Got 28 TPS Out of Free Kaggle GPUs. Here's What It Took.

Comments
5 min read
Choosing the Right GPU for Your Model — A Sizing Method, Not a Guess

Choosing the Right GPU for Your Model — A Sizing Method, Not a Guess

Comments 1
9 min read
What really fits in 8GB VRAM

What really fits in 8GB VRAM

Comments
7 min read
I ran a GPU inference app for a month on Azure serverless GPU. Here's the actual bill.

I ran a GPU inference app for a month on Azure serverless GPU. Here's the actual bill.

1
Comments
4 min read
From API to GPU, Week 5: Tensors, the Data Structure Behind Every Model

From API to GPU, Week 5: Tensors, the Data Structure Behind Every Model

Comments
12 min read
đź‘‹ Sign in for the ability to sort posts by relevant, latest, or top.