Install
Tutorials
Tutorials in GPUs: the news, the names and what changed.
- 15 Tracked terms
- Last 30 days Feed window
What this topic collects on
An article joins this feed when it matches these terms. Each one is also a search of its own.
Related topics
Latest in Tutorials
Write the reference before debugging the shader
2+ hour, 40+ min ago (262+ words) The Flash Attention 2 path was easier to finish after the algorithm existed in plain JavaScript first. The forward pass used one workgroup per batch, head, and query tile. It walked keys and values in blocks while keeping the online-softmax state…...
Ollama GPU requirements: VRAM, RAM, and supported GPUs
14+ hour, 40+ min ago (1814+ words) Ollama GPU requirements range from about 3‑4 GB of VRAM for a small 3B model to around 50 GB for a 70B model, at Q4_K_M quantization (a widely used compressed version of a model) with an 8K context. The main factors that determine where you land…...
Pulse: A New VHDL Simulator
20+ hour, 34+ min ago (35+ words) With VHDL being arguably more deterministic and bullet-proof than Verilog, it's good to see another open source VHDL simulator joining the fray that is not a variation of ghdl. Written by [Óscar Grim......
What Is FSR? A Practical Look at FSR Video Upscaling for Browser-Based Tools
1+ day, 16+ hour ago (369+ words) FSR stands for FidelityFX Super Resolution, an upscaling technology built by AMD. It started as a gaming feature: render a game at a lower resolution (say 1080p), then let FSR reconstruct the missing detail to display it at a higher resolution…...
Building a Stable AI Agent Workflow in NVIDIA's NeMo Agent Toolkit: A Practical Journey
1+ day, 23+ hour ago (80+ words) Frank Morales Aguilera, BEng, MEng, SMIEEE Founder & CEO, SOMALA | Former Boeing Associate Technical Fellow | …...
Semi-Tensor Decomposition Slashes Neural Network Training Time and Memory
2+ day, 8+ hour ago (55+ words) Training large neural networks has long been a battle against two unforgiving constraints: time and memory. Every epoch of learning demands billions of floating-point operations, and every hidden layer multiplies the storage burden on hardware that is already stretched to…...
ROCm vs Vulkan for AMD Local LLM Hosting: 2026 Guide
2+ day, 19+ hour ago (1662+ words) ROCm and Vulkan both accelerate AMD GPUs for local LLM hosting, but they are not interchangeable. The right choice depends on the engine, GPU, and workload. In local LLM hosting the two backends sit at different layers. ROCm is AMD's…...
Command A vs Nex-N2.5-Pro - AI Model Comparison
5+ day, 18+ hour ago (81+ words) OpenRouter Command A vs Nex-N2.5-Pro: side-by-side summary Command A and Nex-N2.5-Pro are available through the OpenRouter API, so switching between them takes a model slug change rather than a new integration. Command A, from Cohere, has a 256,000-token…...
Everthine uses IFA 2026 to define its standalone AI companion platform
6+ day, 15+ hour ago (23+ words) At IFA 2026, Everthine used its European debut to define itself as a standalone AI companion platform, separating from its parent brand Lepro....
How To Get The Best Gaming Performance With Nvidia Control Panel
1+ week, 12+ hour ago (493+ words) bgr.com How To Get The Best Gaming Performance With Nvidia Control Panel PC hardware upgrades are, in many cases, prohibitively expensive these days. Some crafty enthusiasts have even gone so far as to make their own RAM in response…...