Skip to content
Topic

#Gpu

24 articles on Gpu — news, releases, guides and analysis from the SourceFeed engine.

llmfit does the local-LLM math you've been faking
Article 1w ago 6

llmfit does the local-LLM math you've been faking

The k8sgpt creator's Rust CLI right-sizes models to your GPU — keeping its catalog honest is the hard part.

Mariana Souza
The Real Cost of AI-Porting 250k Lines of Fortran

The Real Cost of AI-Porting 250k Lines of Fortran

Article · 2w ago0
NVIDIA's roadmap is a memory story, not a compute story

NVIDIA's roadmap is a memory story, not a compute story

Article · 2w ago0
QEMU Finally Gets Real DirectX 11 Acceleration

QEMU Finally Gets Real DirectX 11 Acceleration

Article · 3w ago0
Your Agents Are Waiting on the CPU, Not the GPU

Your Agents Are Waiting on the CPU, Not the GPU

Article · 3w ago5
An 8B Fine-Tune Now Fits in 4 GB of VRAM

An 8B Fine-Tune Now Fits in 4 GB of VRAM

Article · 3w ago1
AirLLM's 4GB 70B Trick Is Real, and Beside the Point

AirLLM's 4GB 70B Trick Is Real, and Beside the Point

Article · 3w ago1
KNOD Wants to Turn Your AMD GPU Into a DPU

KNOD Wants to Turn Your AMD GPU Into a DPU

Article · 1mo ago0
Autoscale GPU Inference on EKS with Karpenter and Spot Instances

Autoscale GPU Inference on EKS with Karpenter and Spot Instances

Tutorial · 1mo ago0
PyTorch Monarch Escapes the CUDA Bubble

PyTorch Monarch Escapes the CUDA Bubble

Article · 1mo ago0
The 500-Line Renderer Worth Writing Yourself

The 500-Line Renderer Worth Writing Yourself

Article · 1mo ago0
Firefox Bets on Vulkan to Fix Linux Video Decoding

Firefox Bets on Vulkan to Fix Linux Video Decoding

Article · 1mo ago1
CUDA Without NVIDIA: What Actually Works

CUDA Without NVIDIA: What Actually Works

Article · 1mo ago0
Demystifying the NVIDIA DGX Spark for API Developers

Demystifying the NVIDIA DGX Spark for API Developers

Article · 1mo ago0
Popping the CPU-GPU Latency Bubble in Inference

Popping the CPU-GPU Latency Bubble in Inference

Article · 1mo ago2
OpenAI Jalapeno and the Shift to Custom Inference Silicon

OpenAI Jalapeno and the Shift to Custom Inference Silicon

Article · 2mos ago7