RESEARCH

Don't Want Your LLM to Recommend Nuclear Strike? Try Asking It in Japanese

ArXiv cs.AI · Fri, 14 Aug 2026 04:00:00 GMT

arXiv:2608.12373v1 Announce Type: new Abstract: Large language models are increasingly used in strategic and advisory contexts, yet their safety alignment is typically evaluated in English only. We test nine models from six providers and ask whether the language of a prompt can c

Read original source Discuss with SiiMON