Skip to content
Topic

#Gpu

21 articles on Gpu — news, releases, guides and analysis from the SourceFeed engine.

QEMU Finally Gets Real DirectX 11 Acceleration
Article 5d ago 0

QEMU Finally Gets Real DirectX 11 Acceleration

UTM's Triton driver implements the Windows DDI layer itself, unlocking accelerated desktops and games in open-source VMs.

Lenn Voss
Your Agents Are Waiting on the CPU, Not the GPU

Your Agents Are Waiting on the CPU, Not the GPU

Article · 5d ago5
An 8B Fine-Tune Now Fits in 4 GB of VRAM

An 8B Fine-Tune Now Fits in 4 GB of VRAM

Article · 1w ago1
AirLLM's 4GB 70B Trick Is Real, and Beside the Point

AirLLM's 4GB 70B Trick Is Real, and Beside the Point

Article · 1w ago1
KNOD Wants to Turn Your AMD GPU Into a DPU

KNOD Wants to Turn Your AMD GPU Into a DPU

Article · 2w ago0
Autoscale GPU Inference on EKS with Karpenter and Spot Instances

Autoscale GPU Inference on EKS with Karpenter and Spot Instances

Tutorial · 2w ago0
PyTorch Monarch Escapes the CUDA Bubble

PyTorch Monarch Escapes the CUDA Bubble

Article · 2w ago0
The 500-Line Renderer Worth Writing Yourself

The 500-Line Renderer Worth Writing Yourself

Article · 3w ago0
Firefox Bets on Vulkan to Fix Linux Video Decoding

Firefox Bets on Vulkan to Fix Linux Video Decoding

Article · 3w ago1
CUDA Without NVIDIA: What Actually Works

CUDA Without NVIDIA: What Actually Works

Article · 4w ago0
Demystifying the NVIDIA DGX Spark for API Developers

Demystifying the NVIDIA DGX Spark for API Developers

Article · 1mo ago0
Popping the CPU-GPU Latency Bubble in Inference

Popping the CPU-GPU Latency Bubble in Inference

Article · 1mo ago2
OpenAI Jalapeno and the Shift to Custom Inference Silicon

OpenAI Jalapeno and the Shift to Custom Inference Silicon

Article · 1mo ago7
Serve an Open-Source LLM at Scale with vLLM on a Rented GPU Instance

Serve an Open-Source LLM at Scale with vLLM on a Rented GPU Instance

Tutorial · 1mo ago0
The Architecture of Monopoly: Inside NVIDIA's Supercomputing Hegemony

The Architecture of Monopoly: Inside NVIDIA's Supercomputing Hegemony

Article · 1mo ago0
Running 70B Models on 4GB VRAM: The AirLLM Layer-Swap Hack

Running 70B Models on 4GB VRAM: The AirLLM Layer-Swap Hack

Article · 1mo ago1