Podcast Episodes
Back to SearchRISC-V Mentorship Taught Me the RISC-V ISA Is Far More Than a Reference Manual
## Episode Summary In this episode, we cover: - **RISC-V Mentorship Taught Me the RISC-V ISA Is Far More Than a Reference Manual** (riscv_news) - **R…
3 months, 1 week ago
3DLS: A 3D Logic-Stacked Architecture for Disaggregated LLM Serving
## Episode Summary In this episode, we cover: - **3DLS: A 3D Logic-Stacked Architecture for Disaggregated LLM Serving** (arXiv) - **KernelSight-LM: A…
3 months, 1 week ago
High-Performance NTT Accelerators for PQC leveraging Unified Redundant Arithmetic and Fine-Tuned Microarchitecture
## Episode Summary In this episode, we cover: - **High-Performance NTT Accelerators for PQC leveraging Unified Redundant Arithmetic and Fine-Tuned Mi…
3 months, 1 week ago
TraceLab: Characterizing Coding Agent Workloads for LLM Serving
## Episode Summary In this episode, we cover: - **TraceLab: Characterizing Coding Agent Workloads for LLM Serving** (arXiv) - **Croc: Training the Ne…
3 months, 1 week ago
HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerators
## Episode Summary In this episode, we cover: - **HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerato…
3 months, 1 week ago
CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregation
## Episode Summary In this episode, we cover: - **CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregati…
3 months, 1 week ago
SOLAR: AI-Powered Speed-of-Light Performance Analysis
## Episode Summary In this episode, we cover: - **SOLAR: AI-Powered Speed-of-Light Performance Analysis** (arXiv) - **Accelerating Disaggregated RL f…
3 months, 1 week ago
The Serialized Bridge: Understanding and Recovering LLM Serving Performance under Blackwell GPU Confidential Computing
## Episode Summary In this episode, we cover: - **The Serialized Bridge: Understanding and Recovering LLM Serving Performance under Blackwell GPU Con…
3 months, 2 weeks ago
CVA6-RT: an Open-Source Time-Predictable RV64 Processor for Mixed-Criticality Systems
## Episode Summary In this episode, we cover: - **CVA6-RT: an Open-Source Time-Predictable RV64 Processor for Mixed-Criticality Systems** (arXiv) - *…
3 months, 2 weeks ago
Cache-Resident LLM Inference in GB-Scale Last-Level Caches
## Episode Summary In this episode, we cover: - **Cache-Resident LLM Inference in GB-Scale Last-Level Caches** (arXiv) - **Energy-Efficient CNN Accel…
3 months, 2 weeks ago