US Department of Commerce to Evaluate AI Models from Tech Giants in Bid for Enhanced Safety

Ryan Patel, Tech Industry Reporter
5 Min Read
⏱️ 4 min read

**

In a significant move aimed at bolstering public safety and accountability, the US Department of Commerce has announced a new programme to assess artificial intelligence (AI) models developed by major tech firms including Google, Microsoft, and xAI. This initiative, spearheaded by the Centre for AI Standards and Innovation (CAISI), marks a notable shift in regulatory strategy, expanding upon previous agreements made during the Biden Administration with other AI companies like OpenAI and Anthropic.

Expanded Testing Framework for AI Tools

Under this new framework, the aforementioned tech companies have agreed to voluntarily submit their AI models for rigorous testing and evaluation before public release. CAISI’s director, Chris Fall, emphasised the importance of these collaborations, stating, “These expanded industry collaborations help us scale our work in the public interest at a critical moment.” The evaluations will encompass areas such as testing, collaborative research, and the development of best practices related to commercial AI systems.

Among the tools being assessed, Google’s Gemini—a chatbot that has gained traction across various Google products—is noteworthy for its utilisation in US defence and military applications. Similarly, Microsoft’s CoPilot and xAI’s Grok, which has garnered attention for controversial content moderation issues, will also face scrutiny.

A Shift from Previous Administration Policies

This proactive approach represents a departure from the policies of the previous Trump administration, which adopted a largely laissez-faire stance towards AI regulation. Under Trump, executive orders were implemented to minimise regulatory barriers to AI development, promoting the idea that the US could dominate the field through reduced oversight. However, as the military’s reliance on AI continues to grow and concerns about the safety of advanced models come to the forefront, the current administration appears to be reassessing this hands-off approach.

In a recent corporate blog post, Microsoft highlighted its ongoing commitment to testing AI models, asserting that “testing for national security and large-scale public safety risks necessarily must be a collaborative endeavour with governments.” This sentiment underscores a growing recognition among tech giants of the need for a cooperative framework with regulatory bodies in navigating the complexities of AI safety.

Previous Evaluations and Unreleased Models

CAISI has previously conducted evaluations on 40 different AI tools, some of which remain unreleased due to safety concerns. While the centre did not disclose specifics on which models were held back, the mere existence of such evaluations speaks volumes about the heightened scrutiny facing AI advancements. Last month, senior officials from the Trump administration met with Anthropic CEO Dario Amodei, even as the company faces a lawsuit from the US Department of Defense over its reluctance to remove safety barriers for government use of its models.

This evolving landscape suggests that the administration is now more inclined to impose checks and balances on AI technology, particularly as new powerful models emerge that raise ethical and safety questions.

The Role of the Private Sector in AI Regulation

The decision to involve private companies in the safety testing of AI tools highlights the critical role that the tech sector must play in ensuring responsible development and deployment of AI technologies. As the capabilities of these models become increasingly sophisticated, the potential risks associated with their misuse or unintended consequences also escalate.

Notably, the shift in policy reflects a growing consensus that collaboration between the government and tech firms is essential for mitigating risks while fostering innovation. This partnership could ultimately lead to the establishment of robust safety standards that not only protect the public but also enable the responsible advancement of AI technologies.

Why it Matters

The new testing programme initiated by the US Department of Commerce signifies a pivotal moment in the governance of artificial intelligence, signalling a move towards more rigorous oversight in an industry that has historically operated with minimal regulation. As AI becomes woven into the fabric of various sectors—including defence, healthcare, and finance—ensuring the safety and ethical implications of its deployment is paramount. This initiative not only aims to safeguard public interests but also sets a precedent for global AI governance, positioning the US at the forefront of a critical dialogue on the responsible use of emerging technologies.

Share This Article
Ryan Patel reports on the technology industry with a focus on startups, venture capital, and tech business models. A former tech entrepreneur himself, he brings unique insights into the challenges facing digital companies. His coverage of tech layoffs, company culture, and industry trends has made him a trusted voice in the UK tech community.
Leave a Comment

Leave a Reply

Your email address will not be published. Required fields are marked *

© 2026 The Update Desk. All rights reserved.
Terms of Service Privacy Policy