Podcast Episodes
Back to SearchSLAC: Access-Driven CPU-to-GPU Side-channel Attacks via System-Level Cache on Apple Silicon
## Episode Summary In this episode, we cover: - **SLAC: Access-Driven CPU-to-GPU Side-channel Attacks via System-Level Cache on Apple Silicon** (arXi…
1 month, 3 weeks ago
Why Do Prefetchers Fail? Let Agents Answer
## Episode Summary In this episode, we cover: - **Why Do Prefetchers Fail? Let Agents Answer** (arXiv) - **Machine Shape and Hierarchical Blocking: A…
1 month, 3 weeks ago
APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference
## Episode Summary In this episode, we cover: - **APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference** (arXiv) - **Aneto: Pred…
1 month, 3 weeks ago
ArchAgent v2: A Case Study with the Data Prefetching Championship
## Episode Summary In this episode, we cover: - **ArchAgent v2: A Case Study with the Data Prefetching Championship** (arXiv) - **The Fallacy of Inde…
1 month, 4 weeks ago
A Centralized Performance Monitoring Architecture for Heterogeneous Multicore SoCs
## Episode Summary In this episode, we cover: - **A Centralized Performance Monitoring Architecture for Heterogeneous Multicore SoCs** (arXiv) - **MC…
2 months ago
Heterogeneous LLM Serving with General-Purpose Processing-Near-Memory for Retrieval-Based Sparse Attention
## Episode Summary In this episode, we cover: - **Heterogeneous LLM Serving with General-Purpose Processing-Near-Memory for Retrieval-Based Sparse At…
2 months ago
PLoRA: An NDP-Enhanced Pooled-Memory System for Cost-Efficient Multi-LoRA Serving
## Episode Summary In this episode, we cover: - **PLoRA: An NDP-Enhanced Pooled-Memory System for Cost-Efficient Multi-LoRA Serving** (arXiv) - **Bre…
2 months ago
Architectural Implications of Agentic AI Workflows
## Episode Summary In this episode, we cover: - **Architectural Implications of Agentic AI Workflows** (arXiv) - **On the Limits of Machine-Learned R…
2 months ago
LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension
## Episode Summary In this episode, we cover: - **LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension** (ar…
2 months ago
Themis: Software-Defined Hardware Prefetching
## Episode Summary In this episode, we cover: - **Themis: Software-Defined Hardware Prefetching** (arXiv) - **Beyond Static Policies: Dynamic Selecti…
2 months ago