Skip to content
Launchpad Library logo

model behavior

2 free resources on this topic. Everything here is free and hand-picked. You can also search within this topic.

FreeArticle
Technology & Ethics

Sycophancy in GPT-4o: What Happened and What We're Doing About It

OpenAI's April 29, 2025 explanation of why it rolled back a ChatGPT update that made the model overly flattering and agreeable. It says the update leaned too heavily on short-term thumbs-up feedback, and lists the fixes it planned.

Why I recommend it: A company explaining its own mistake, so read it as OpenAI's account, not an independent review. Useful for seeing why a chatbot that always agrees with you isn't a reliable advisor.

#ai ethics#ai safety#chatgpt#model behavior#openai#sycophancy
openai.comAdded Sep 23, 20260 opens
FreeArticle
Technology & Ethics

An Alien Mind

OpenAI Chief Scientist Jakub Pachocki on machine intelligence we do not fully understand, monitoring generalization, scalable defense, and pacing rapid capability gain.

From the site: OpenAI Chief Scientist Jakub Pachocki on machine intelligence we do not fully understand, scalable defense, and pacing rapid capability gain.

Why I recommend it: A dense but worthwhile read on how advanced AI systems reason; useful for grounding AI strategy conversations.

#ai#ai ethics#ai-ethics#ai-safety#article#cognition#free#model-behavior#openai#reading#research#tech-ethics#technology
OpenAIAdded Sep 7, 20260 opens