Episode Details

Back to Episodes

APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference

Published 1 month, 3 weeks ago
Description
## Episode Summary In this episode, we cover: - **APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference** (arXiv) - **Aneto: Predicting System Performance by Exploiting Cross-Workload Regularity** (arXiv) - **RISC-V: The Open-Source Revolution in CPU Architecture - design-reuse.com** (google_arch) - **Conflict to Compliance: RISC-V Extension Migration Across Spec, HW, and SW** (riscv_news) - **China's LineShine tops supercomputer ranking with all-CPU architecture - China Daily** (google_arch) --- *Sponsored by Ada, Ago Consulting, and Zen Semiconductor*
Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us