What is the “alignment problem” in philosophical terms?
At work people use "alignment" to mean very different things: making a chatbot polite, making it follow instructions, or making sure future systems don't pursue goals harmful to humans.
What's the core philosophical problem underneath? Is it a problem about values (which values should a system have?), about specification (how do we say what we want?), or about control? Are there older philosophical debates it connects to?