Podcast Episodes
Back to SearchAMD reveals CPU architecture roadmap through 2028, following Zen 6 'Venice' launch — Zen 7 'Florence' to de... - Yahoo Tech
## Episode Summary In this episode, we cover: - **AMD reveals CPU architecture roadmap through 2028, following Zen 6 'Venice' launch — Zen 7 'Florenc…
2 months, 2 weeks ago
NextSilicon Productizes Arbel RISC-V Core Into 64-Core Enterprise Processor - Quantum Zeitgeist
## Episode Summary In this episode, we cover: - **NextSilicon Productizes Arbel RISC-V Core Into 64-Core Enterprise Processor - Quantum Zeitgeist** (…
2 months, 2 weeks ago
SEAM-V: A Hybrid-Decoupled RISC-V Vector Processor with Backend-Visible EP Context for Sustained Vector Throughput
## Episode Summary In this episode, we cover: - **SEAM-V: A Hybrid-Decoupled RISC-V Vector Processor with Backend-Visible EP Context for Sustained Ve…
2 months, 2 weeks ago
DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction
## Episode Summary In this episode, we cover: - **DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction** (arXiv) …
2 months, 2 weeks ago
A Quantized Native Runtime for On-Device Semantic Audio Generation
## Episode Summary In this episode, we cover: - **A Quantized Native Runtime for On-Device Semantic Audio Generation** (arXiv) - **Pattern-Guided Des…
2 months, 3 weeks ago
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel
## Episode Summary In this episode, we cover: - **Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verifie…
2 months, 3 weeks ago
CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation
## Episode Summary In this episode, we cover: - **CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator …
2 months, 3 weeks ago
Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit
## Episode Summary In this episode, we cover: - **Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Lim…
2 months, 3 weeks ago
CLIP-3D: Closed-Loop Evaluation of Performance and Physical Constraints for 3D ICs
## Episode Summary In this episode, we cover: - **CLIP-3D: Closed-Loop Evaluation of Performance and Physical Constraints for 3D ICs** (arXiv) - **At…
2 months, 3 weeks ago
FlashAccel: Leveraging High-Bandwidth Flash for High-Throughput LLM Inference
## Episode Summary In this episode, we cover: - **FlashAccel: Leveraging High-Bandwidth Flash for High-Throughput LLM Inference** (arXiv) - **IRONSmi…
2 months, 3 weeks ago