OpenAI recently decided not to release a new artificial intelligence model, citing significant safety concerns and its inability to follow complex instructions. This move, coupled with Anthropic's explicit warnings about AI's existential risks, underscores the growing urgency for a more cautious and responsible approach to developing algorithmic technologies.
What happened
OpenAI's decision to halt an advanced model before its launch sent ripples through the industry. A top executive at the company told the Wall Street Journal that the model in question displayed a "poor aptitude for following orders," raising significant doubts about its reliability and safety in real-world scenarios. This move, also reported by TechCrunch AI, highlights the inherent challenges in developing increasingly autonomous and complex artificial intelligence systems. A model's ability to adhere to instructions is fundamental for its safe deployment and to prevent unexpected or harmful behaviors.
Concurrently, Anthropic, another leading AI firm with a strong focus on safety, released a prospectus to investors. While detailing exponential growth and significant losses in the order of tens of billions of dollars annually, the prospectus also included an explicit warning that its own AI could pose an existential risk to humanity TechCrunch AI. This document, intended for a financial audience, not only reveals the rapid expansion of the sector but also presents an unprecedented ethical and financial dilemma, forcing investors to confront long-term scenarios that extend beyond mere economic returns.
Why it matters
These events are not isolated; they serve as a wake-up call for the entire AI industry and for society. OpenAI's decision to withdraw a model before release is a strong signal: the race for innovation must be balanced with unprecedented rigor in terms of safety and control. It's no longer just about optimizing performance, but about ensuring that systems do not generate unpredictable or harmful outcomes—a problem that goes beyond simple bug fixes. A model's inability to follow instructions can lead to serious consequences, from the spread of misinformation to the manipulation of critical processes.
Anthropic's statements elevate the debate on AI risks to an even more explicit level, embedding it within an economic and corporate responsibility context. The question of "who should be held accountable when an AI Agent (accidentally) acts maliciously?" becomes central, as discussed in recent articles Hacker News AI filtered. The impact on labor and society is clear: unpredictable systems can cause economic, social, and even physical harm, eroding public trust and jeopardizing the responsible adoption of AI. These episodes underscore the fragility of public trust and the need for robust accountability mechanisms to prevent widespread misuse or incidents.
The HDAI perspective
These developments reinforce Human Driven AI's conviction that technological development must be intrinsically linked to principles of ethical AI and robust governance. It's not merely about preventing technical errors, but about establishing a clear and transparent framework of responsibility that places humans at the center. Transparency and accountability are fundamental pillars for building a future where AI serves humanity, improving lives rather than creating new threats. Companies' capacity for self-regulation is being tested, and their willingness to prioritize safety over release speed is a crucial indicator of their ethical commitment.
Events involving OpenAI and Anthropic underscore the urgency of an open and constructive dialogue among developers, legislators, academics, and civil society. It is essential that discussions about the risks and benefits of AI are not confined to research labs or corporate boardrooms but become part of an informed public debate. This is a central theme that will be explored at the HDAI Summit 2026 in Pompeii, where Italian and international experts will discuss strategies for AI governance that are effective and oriented towards collective well-being. It is imperative that companies do not merely acknowledge risks but implement proactive assessment and mitigation mechanisms, placing human impact at the core of every development decision and ensuring that innovation is sustainable and safe.
What to watch
Attention will increasingly shift towards companies' ability to self-regulate and towards legislative pressure, such as the EU AI Act, to impose safety and transparency standards. It will be crucial to observe how AI giants balance the drive for innovation with the need for effective governance and greater accountability. International collaboration and the exchange of best practices will be essential to address challenges that transcend national borders. The future of AI will depend on our collective ability to guide this technology with wisdom and responsibility.

