Episode Details
Back to Episodes
OpenAI’s custom inference chip & MoE fine-tuning gets faster - AI News (Jun 26, 2026)
Published 1 month, 2 weeks ago
Description
Please support this podcast by checking out our sponsors:
- KrispCall: Agentic Cloud Telephony - https://try.krispcall.com/tad
- Invest Like the Pros with StockMVP - https://www.stock-mvp.com/?via=ron
- Discover the Future of AI Audio with ElevenLabs - https://try.elevenlabs.io/tad
Support The Automated Daily directly:
Buy me a coffee: https://buymeacoffee.com/theautomateddaily
-
- KrispCall: Agentic Cloud Telephony - https://try.krispcall.com/tad
- Invest Like the Pros with StockMVP - https://www.stock-mvp.com/?via=ron
- Discover the Future of AI Audio with ElevenLabs - https://try.elevenlabs.io/tad
Support The Automated Daily directly:
Buy me a coffee: https://buymeacoffee.com/theautomateddaily
Today's topics:
OpenAI’s custom inference chip - OpenAI and Broadcom revealed “Jalapeño,” a purpose-built inference accelerator aimed at lowering LLM serving cost and boosting performance-per-watt at data-center scale.
MoE fine-tuning gets faster - NVIDIA and Hugging Face pushed MoE training forward with NeMo AutoModel, highlighting throughput and memory gains that could reduce the GPU barrier for fine-tuning large MoE LLMs.
Apple’s AI-first Mac roadmap - A report says Apple may skip higher-end M6 variants and jump pro Macs to AI-heavier M7 Pro/Max/Ultra chips, signaling on-device AI as a core silicon priority.
Gemini adds built-in computer use - Google folded “computer use” into Gemini 3.5 Flash, making it easier to build agents that can see interfaces and take actions while adding safeguards against prompt injection.
Amazon versus Perplexity’s agent browser - Amazon sued Perplexity over its Comet agentic browser, raising big questions about bot disclosure, user-agent spoofing, and who’s accountable when AI acts inside logged-in sessions.
Anthropic alleges mass distillation attack - Anthropic told the U.S. Senate it believes Alibaba-linked operators ran a large-scale model distillation campaign using fraudulent accounts, escalating the policy fight over AI capability “theft.”
Diffusion to engine-ready 3D geometry - Google Research introduced FLAT, a way to decode diffusion video latents directly into explicit triangle geometry, potentially shortening the path from generative models to engine-friendly 3D assets.
Qwen’s AgentWorld simulation models - Qwen’s AgentWorld models aim to predict environment changes from actions, positioning “world models” as a foundation for stronger planning, simulation training, and agent evaluation benchmarks.
Humans step back into factories - Ford is rehiring veteran engineers after AI quality tools fell short, underscoring that automation still struggles with messy real-world manufacturing and diagnosis.
Prompt-injection stress test by email - The hackmyclaw.com experiment drew thousands of prompt-injection attempts by email, and while secrets weren’t exfiltrated, it revealed real operational risks like account suspensions and context contamination.
AI kids’ books slip marketplace checks - A writer found disturbing image failures in an AI-made children’s bestseller on Amazon, spotlighting weak safety and editorial filters in high-volume AI content marketplaces.
Gaming rumor: Fable sighting in Bedrock - A viral post claims “Fable 5” resurfaced inside Amazon Bedrock Chat, an unverified platform ‘sighting’ that can fuel speculation about backend listings and upcoming game announcements.
-
Listen Now
Love PodBriefly?
If you like Podbriefly.com, please consider donating to support the ongoing development.
Support Us