Podcast Episodes

Back to Search
SLAC: Access-Driven CPU-to-GPU Side-channel Attacks via System-Level Cache on Apple Silicon

## Episode Summary In this episode, we cover: - **SLAC: Access-Driven CPU-to-GPU Side-channel Attacks via System-Level Cache on Apple Silicon** (arXi…

1 month, 3 weeks ago

Short Long
View Episode
Why Do Prefetchers Fail? Let Agents Answer

## Episode Summary In this episode, we cover: - **Why Do Prefetchers Fail? Let Agents Answer** (arXiv) - **Machine Shape and Hierarchical Blocking: A…

1 month, 3 weeks ago

Short Long
View Episode
APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference

## Episode Summary In this episode, we cover: - **APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference** (arXiv) - **Aneto: Pred…

1 month, 3 weeks ago

Short Long
View Episode
ArchAgent v2: A Case Study with the Data Prefetching Championship

## Episode Summary In this episode, we cover: - **ArchAgent v2: A Case Study with the Data Prefetching Championship** (arXiv) - **The Fallacy of Inde…

1 month, 4 weeks ago

Short Long
View Episode
A Centralized Performance Monitoring Architecture for Heterogeneous Multicore SoCs

## Episode Summary In this episode, we cover: - **A Centralized Performance Monitoring Architecture for Heterogeneous Multicore SoCs** (arXiv) - **MC…

2 months ago

Short Long
View Episode
Heterogeneous LLM Serving with General-Purpose Processing-Near-Memory for Retrieval-Based Sparse Attention

## Episode Summary In this episode, we cover: - **Heterogeneous LLM Serving with General-Purpose Processing-Near-Memory for Retrieval-Based Sparse At…

2 months ago

Short Long
View Episode
PLoRA: An NDP-Enhanced Pooled-Memory System for Cost-Efficient Multi-LoRA Serving

## Episode Summary In this episode, we cover: - **PLoRA: An NDP-Enhanced Pooled-Memory System for Cost-Efficient Multi-LoRA Serving** (arXiv) - **Bre…

2 months ago

Short Long
View Episode
Architectural Implications of Agentic AI Workflows

## Episode Summary In this episode, we cover: - **Architectural Implications of Agentic AI Workflows** (arXiv) - **On the Limits of Machine-Learned R…

2 months ago

Short Long
View Episode
LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension

## Episode Summary In this episode, we cover: - **LACE: Large Language Model Aided Multi-Agent Framework for Agile RISC-V Instruction Extension** (ar…

2 months ago

Short Long
View Episode
Themis: Software-Defined Hardware Prefetching

## Episode Summary In this episode, we cover: - **Themis: Software-Defined Hardware Prefetching** (arXiv) - **Beyond Static Policies: Dynamic Selecti…

2 months ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us