Podcast Episodes

Back to Search
AMD reveals CPU architecture roadmap through 2028, following Zen 6 'Venice' launch — Zen 7 'Florence' to de... - Yahoo Tech

## Episode Summary In this episode, we cover: - **AMD reveals CPU architecture roadmap through 2028, following Zen 6 'Venice' launch — Zen 7 'Florenc…

2 months, 2 weeks ago

Short Long
View Episode
NextSilicon Productizes Arbel RISC-V Core Into 64-Core Enterprise Processor - Quantum Zeitgeist

## Episode Summary In this episode, we cover: - **NextSilicon Productizes Arbel RISC-V Core Into 64-Core Enterprise Processor - Quantum Zeitgeist** (…

2 months, 2 weeks ago

Short Long
View Episode
SEAM-V: A Hybrid-Decoupled RISC-V Vector Processor with Backend-Visible EP Context for Sustained Vector Throughput

## Episode Summary In this episode, we cover: - **SEAM-V: A Hybrid-Decoupled RISC-V Vector Processor with Backend-Visible EP Context for Sustained Ve…

2 months, 2 weeks ago

Short Long
View Episode
DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction

## Episode Summary In this episode, we cover: - **DSTAR: Accelerating Diffusion Transformers via Spatial and Temporal Redundancy Reduction** (arXiv) …

2 months, 2 weeks ago

Short Long
View Episode
A Quantized Native Runtime for On-Device Semantic Audio Generation

## Episode Summary In this episode, we cover: - **A Quantized Native Runtime for On-Device Semantic Audio Generation** (arXiv) - **Pattern-Guided Des…

2 months, 3 weeks ago

Short Long
View Episode
Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verified-Knowledge Flywheel

## Episode Summary In this episode, we cover: - **Smarter and Cheaper at Once: Byte-Exact KV-Cache Grafting Turns a Frozen Small Model into a Verifie…

2 months, 3 weeks ago

Short Long
View Episode
CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator Disaggregation

## Episode Summary In this episode, we cover: - **CODA: Algorithm-Hardware Co-design for Edge Video Diffusion via NMP-Enabled Compute-Cache Operator …

2 months, 3 weeks ago

Short Long
View Episode
Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Limit

## Episode Summary In this episode, we cover: - **Full-Pipeline Inference Optimization for MiMo-V2.5 Series: Pushing Hybrid SWA Efficiency to the Lim…

2 months, 3 weeks ago

Short Long
View Episode
CLIP-3D: Closed-Loop Evaluation of Performance and Physical Constraints for 3D ICs

## Episode Summary In this episode, we cover: - **CLIP-3D: Closed-Loop Evaluation of Performance and Physical Constraints for 3D ICs** (arXiv) - **At…

2 months, 3 weeks ago

Short Long
View Episode
FlashAccel: Leveraging High-Bandwidth Flash for High-Throughput LLM Inference

## Episode Summary In this episode, we cover: - **FlashAccel: Leveraging High-Bandwidth Flash for High-Throughput LLM Inference** (arXiv) - **IRONSmi…

2 months, 3 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us