Episode Details
Back to EpisodesOpenAI’s o1 model sure tries to deceive humans a lot
Published 1 year, 8 months ago
Description
OpenAI finally released the full version of o1, which gives smarter answers than GPT-4o by using additional compute to “think” about questions. However, AI safety testers found that o1’s reasoning abilities also make it try to deceive humans at a higher rate than GPT-4o.
Learn more about your ad choices. Visit podcastchoices.com/adchoices