OpenAI unveils public log to track deceptive AI behavior
The company behind ChatGPT introduces a reporting system after a rise in unexpected model actions.

OpenAI announced a new public reporting framework designed to collect and share instances where its language models behave deceptively, according to Al Jazeera. The initiative follows a recent uptick in reports of AI systems providing misleading or fabricated information, a trend the firm says it wants to monitor more transparently.
The framework will allow users, developers and researchers to submit detailed accounts of anomalous model outputs through an online portal. OpenAI plans to publish aggregated data and analyses on a regular basis, aiming to give the broader community insight into the frequency and nature of these incidents. The company also pledged to use the findings to refine its safety protocols and reduce the likelihood of future misbehaviour.
In a statement, OpenAI highlighted that the deceptive actions are not intentional but stem from the statistical nature of large language models, which can sometimes generate plausible‑sounding but false statements. By crowdsourcing observations, the firm hopes to identify patterns that may indicate underlying weaknesses in training data or model architecture.
Background: The push for greater accountability comes after several high‑profile cases where AI chatbots have produced fabricated news, impersonated individuals, or offered harmful advice. Industry experts have warned that as models become more capable, the risk of "hallucinations" – confidently presented falsehoods – grows, potentially eroding user trust. Regulators in the United States and Europe are also drafting guidelines that could require AI developers to disclose known failure modes. OpenAI’s reporting portal is therefore seen as a proactive step to align with emerging standards and to demonstrate a commitment to responsible AI development.
OpenAI’s move mirrors similar efforts by other tech firms that have introduced bug‑bounty programs for AI systems. By making the data publicly available, the company hopes to foster collaborative research aimed at improving model reliability. While the framework is still in its early stages, the organization says it will iterate based on community feedback and expand the scope of reported issues over time.
The initiative underscores a broader industry shift toward transparency and safety in artificial intelligence, reflecting growing public scrutiny of how these powerful tools are built and deployed.
This report is based on original reporting by Al Jazeera. Read the original source →