Episode Details

Back to Episodes

AI Textbooks, Enterprise Agents, and Code Security Benchmarks | UpNext AI – August 13, 2026

Episode 75 Published 1 week, 6 days ago
Description

Today on UpNext AI: a practitioner’s view of where AI writing still falls short, a demanding benchmark for enterprise agents, and new evidence on the limits of automated code-vulnerability detection.

Covered stories:
- Nathan Lambert on using AI to write a technical textbook—and why long-form technical writing remains difficult for current models.
- VAKRA, a benchmark testing agents that must work across APIs, documents, and tool-use policies.
- VICBench, a new multi-language benchmark for tracing code vulnerabilities to the commits that introduced them.
- OpenAI-backed Thrive Holdings raises $2 billion for enterprise AI.
- Google launches the Pixel 11 lineup with new AI features and the Tensor G6 chip.
- Anthropic hires legal-tech founder Robert Mahari to lead Claude’s work with law practices.

Sources:
- https://www.interconnects.ai/p/i-wrote-an-ai-textbook-how-long-until
- https://arxiv.org/abs/2608.12282v1
- https://arxiv.org/abs/2608.12246v1
- https://techcrunch.com/2026/08/12/openai-backed-thrive-holdings-raises-2b-to-bring-ai-to-the-enterprise/
- https://www.theverge.com/gadgets/975237/google-pixel-11-pro-comparison-specs-price-features
- https://the-decoder.com/legal-startup-founder-robert-mahari-joins-anthropic-to-lead-claudes-push-into-law-practices/

Listen Now

Love PodBriefly?

If you like Podbriefly.com, please consider donating to support the ongoing development.

Support Us