Episode Details
Back to EpisodesTest-Time Compute Scaling for Robot Policies (DIRECT)
Published 3 months, 3 weeks ago
Description
Larger models + more thinking + more context improve performance on some prompts but not others; a learned router enables better performance/latency trade-offs.