Episode Details
Back to EpisodesAPEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference
Published 1 month, 3 weeks ago
Description
## Episode Summary
In this episode, we cover:
- **APEX: Adaptive Expert Prefetching for Memory-Efficient Edge MoE Inference** (arXiv)
- **Aneto: Predicting System Performance by Exploiting Cross-Workload Regularity** (arXiv)
- **RISC-V: The Open-Source Revolution in CPU Architecture - design-reuse.com** (google_arch)
- **Conflict to Compliance: RISC-V Extension Migration Across Spec, HW, and SW** (riscv_news)
- **China's LineShine tops supercomputer ranking with all-CPU architecture - China Daily** (google_arch)
---
*Sponsored by Ada, Ago Consulting, and Zen Semiconductor*