Podcast Episodes

Back to Search
EP021: How Small Teams Use AI Gateways to Ship Faster

Five patterns small teams (2–10 engineers) use to ship AI products faster with an API gateway: unified key management, instant model switching withou…

5 months, 2 weeks ago

Short Long
View Episode
EP020: The Future of AI APIs — Predictions for 2026-2027

Five bold predictions for where AI APIs are heading: model prices drop 80-90%, context windows hit 1M tokens standard, multimodal becomes default, MC…

5 months, 2 weeks ago

Short Long
View Episode
EP019: Function Calling Across Different AI Models

Every major AI model supports function calling now — but the implementations vary wildly. We compare GPT-4o, GPT-4.1, Claude Opus 4, Gemini 2.5 Pro, …

5 months, 3 weeks ago

Short Long
View Episode
EP018: AI Model Benchmarks That Actually Matter

Most AI benchmarks — MMLU, HumanEval, GSM8K — don't predict real-world performance. We break down why benchmark scores mislead developers, and reveal…

5 months, 3 weeks ago

Short Long
View Episode
EP017: LangChain + Crazyrouter — Build Agents with 627 Models

How to combine LangChain with Crazyrouter for multi-model AI agents. We cover zero-migration setup with ChatOpenAI, cost optimization through per-ste…

5 months, 3 weeks ago

Short Long
View Episode
EP016: Streaming vs Non-Streaming API Calls — Performance Deep Dive

Should you stream your LLM API responses or wait for the full result? We break down Time to First Token, perceived latency, the real cost implication…

5 months, 3 weeks ago

Short Long
View Episode
EP015: AI API Security Best Practices for Production Apps

Five essential security practices every developer needs when shipping AI-powered apps to production. We cover API key management, spending caps, prom…

5 months, 3 weeks ago

Short Long
View Episode
EP014: Top 5 AI Coding Assistants in 2026

We rank the top five AI coding assistants in 2026 — from Amazon Q Developer and Windsurf at the bottom, through GitHub Copilot, all the way up to Cur…

5 months, 3 weeks ago

Short Long
View Episode
EP013: Self-Hosting vs API Gateway — The Real Cost of Running Your Own AI Models

Should you self-host AI models or use an API gateway? We break down the true costs of running Llama 4 and DeepSeek on your own GPUs versus paying per…

5 months, 3 weeks ago

Short Long
View Episode
EP012: Building AI Agents with Tool Calling — A Practical Guide

Tool calling is what turns a chatbot into an agent. We break down how GPT-5, Claude, and Gemini handle function calling, walk through building your f…

5 months, 3 weeks ago

Short Long
View Episode

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us