Bookmarks
2026-08-31
2026-04-10
2026-02-06
2026-02-05
2025-12-25
2025-12-19
2025-12-09
2025-11-17
2025-10-062
2025-09-01
2025-08-19
2025-08-18
2025-07-2410
- Zig in Depth: Vectors and SIMD
- 03 CUDA Fundamental Optimization Part 1
- ARM Assembly: Lesson 1 (MOV, Exit Syscall)
- Refterm Lecture Part 5 - Parsing with SIMD
- CppCon 2016: Timur Doumler “Want fast C++? Know your hardware!"
- C++ cache locality and branch predictability
- Zen 5 And AI Doom w/ Casey Muratori
- The Tech Poutine #23: AMD's Moving to 2nm
- Why GPU Programming Is Chaotic
- Custom Floating-Point Formats
2025-07-2235
- HOW TRANSISTORS REMEMBER DATA
- [PLDI24] Descend: A Safe GPU Systems Programming Language
- Dylan Patel - Inference Math, Simulation, and AI Megaclusters - Stanford CS 229S - Autumn 2024
- Mini Project: How to program a GPU? | CUDA C/C++
- Jim Keller: Moore’s Law is Not Dead
- How CPU Memory & Caches Work - Computerphile
- The ARC Prize 2024 Winning Algorithm
- Endgame: Big Tech Bytes the Dust - Jim Keller, Tenstorrent, Tesla, Apple, AMD, Intel #261
- Designing in 2023: 10 Problems to Solve w/ Jim Keller
- What Jim Keller Sees that Others Miss – DemystifySci #326
- One System, Eight Tenstorrent Wormholes
- RISC-V Day '23 Summer The Future of RISC-V and RISC-V AI (Jim Keller | CEO, Tenstorrent)
- Tenstorrent: Relegating the Important Stuff to the Compiler
- Live at NVIDIA GTC with Acquired
- Past, Present & Future of AI Compute (Panel) | Beyond CUDA Summit 2025
- NVIDIA Doesn't Care About GPUs
- The Exact Moment AMD Beat Intel
- Getting Started with Multi-GPU Scaling: Distributed Libraries | NVIDIA GTC 2025
- TPU V4 and Trends in Accelerator Hardware - Mike Hutton
- Fujitsu’s New ARM Chip: Focused, Fast, and Unlike Anything Else
- The State of Silicon and the GPU Poors - with Dylan Patel of SemiAnalysis
- Mojo meets AMD MI300X: Modular GPU Kernel Hackathon Highlights 🔥
- Must Know Technique in GPU Computing | Episode 4: Tiled Matrix Multiplication in CUDA C
- Dylan Patel (SemiAnalysis) on Multi-Datacenter Training @ Decentralized AI Day 2025
- Inside a Real High-Frequency Trading System | HFT Architecture
- Scaling Computing Performance Beyond the End of Moore’s Law: Song Han
- Dylan Patel: GPT4.5's Flop, Grok 4, Meta's Poaching Spree, Apple's Failure, and Super Intelligence
- FPGA in HFT Systems Explained | Why Reconfigurable Hardware Beats CPUs
- #22 Dylan Patel: China’s Robotics Dominance; AI Infrastructure Breakdown
- Hardware and software [2024 edition]
- Memristors for Analog AI Chips
- Jeff Dean (Google): Exciting Trends in Machine Learning
- 1.2 - Racing Down the Slopes of Moore’s Law (Bram Nauta)
- How does Groq LPU work? (w/ Head of Silicon Igor Arsovski!)
- Intel's Crazy Plan for AI Chips IS WORKING! (Supercut)
2025-07-092
2025-06-27
2025-06-26
2025-05-29
2025-05-16
2025-04-22
2025-04-193
2025-04-10
2025-04-05
2025-03-293
2025-03-28
2025-03-27
2025-03-174
2025-02-26
2025-02-25
2025-02-15
2024-11-29
2024-11-19
2024-11-17
2024-09-23
2024-07-30
2024-07-294
2024-07-233
2024-07-21
2024-07-16
2024-07-11
2024-07-10
2024-07-092
2024-07-08
2024-07-01
2024-06-11
2024-05-28
2024-05-27
2024-03-06
Subcategories
- ai_accelerators (29)
- arm (4)
- floating_point (5)
- gpus (31)
- memory_models (15)
- optimization (17)
- vectorization (8)