Podcast Episodes

Back to Search
RISC-V Mentorship Taught Me the RISC-V ISA Is Far More Than a Reference Manual

## Episode Summary In this episode, we cover: - **RISC-V Mentorship Taught Me the RISC-V ISA Is Far More Than a Reference Manual** (riscv_news) - **R…

3 months, 1 week ago

Short Long
View Episode
3DLS: A 3D Logic-Stacked Architecture for Disaggregated LLM Serving

## Episode Summary In this episode, we cover: - **3DLS: A 3D Logic-Stacked Architecture for Disaggregated LLM Serving** (arXiv) - **KernelSight-LM: A…

3 months, 1 week ago

Short Long
View Episode
High-Performance NTT Accelerators for PQC leveraging Unified Redundant Arithmetic and Fine-Tuned Microarchitecture

## Episode Summary In this episode, we cover: - **High-Performance NTT Accelerators for PQC leveraging Unified Redundant Arithmetic and Fine-Tuned Mi…

3 months, 1 week ago

Short Long
View Episode
TraceLab: Characterizing Coding Agent Workloads for LLM Serving

## Episode Summary In this episode, we cover: - **TraceLab: Characterizing Coding Agent Workloads for LLM Serving** (arXiv) - **Croc: Training the Ne…

3 months, 1 week ago

Short Long
View Episode
HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerators

## Episode Summary In this episode, we cover: - **HBM Is Not All You Need: Efficient Disaggregated LLM Serving across Memory-heterogeneous Accelerato…

3 months, 1 week ago

Short Long
View Episode
CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregation

## Episode Summary In this episode, we cover: - **CrossPool: Efficient Multi-LLM Serving for Cold MoE Models through KV-Cache and Weight Disaggregati…

3 months, 1 week ago

Short Long
View Episode
SOLAR: AI-Powered Speed-of-Light Performance Analysis

## Episode Summary In this episode, we cover: - **SOLAR: AI-Powered Speed-of-Light Performance Analysis** (arXiv) - **Accelerating Disaggregated RL f…

3 months, 1 week ago

Short Long
View Episode
The Serialized Bridge: Understanding and Recovering LLM Serving Performance under Blackwell GPU Confidential Computing

## Episode Summary In this episode, we cover: - **The Serialized Bridge: Understanding and Recovering LLM Serving Performance under Blackwell GPU Con…

3 months, 2 weeks ago

Short Long
View Episode
CVA6-RT: an Open-Source Time-Predictable RV64 Processor for Mixed-Criticality Systems

## Episode Summary In this episode, we cover: - **CVA6-RT: an Open-Source Time-Predictable RV64 Processor for Mixed-Criticality Systems** (arXiv) - *…

3 months, 2 weeks ago

Short Long
View Episode
Cache-Resident LLM Inference in GB-Scale Last-Level Caches

## Episode Summary In this episode, we cover: - **Cache-Resident LLM Inference in GB-Scale Last-Level Caches** (arXiv) - **Energy-Efficient CNN Accel…

3 months, 2 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us