Wire flash
TechNIST launches AI Technology Evaluation platform for model safety testing
Editorial responsibility
- No named human review is recorded for this page.
- Source reporting is collected, normalized, translated or condensed automatically when needed.
- Automatically published source-backed update
The National Institute of Standards and Technology (NIST) launched the AI Technology Evaluation (AITE) program on July 27, 2026, providing researchers with an isolated testbed environment to safely evaluate AI models. AITE is a voluntary testing platform focused on AI model safety analysis, offering blind data for models to process without using it as training data. Initially, the platform will evaluate large vision language models on image analysis tasks across three domains: quantum science, genomics, and public safety. Both data providers and model developers must submit materials for testing, with the first evaluations scheduled for August 2026. The program aims to establish a universal rubric for assessing AI model capabilities and determining state-of-the-art performance. AITE represents the latest step in the Trump administration's strategy to collaborate with major AI developers on voluntary model safety submissions, following a May 2026 Commerce Department deal with Google DeepMind, Microsoft, and xAI to evaluate their models through the Center for AI Standards and Innovation.
Source report
The National Institute of Standards and Technology (NIST) launched a new program on Monday, granting researchers access to an isolated testbed environment to safely evaluate artificial intelligence models against various commands.
About AITE
The AI Technology Evaluation, or AITE, is a voluntary testing vehicle focused on AI model safety analysis. It provides blind data for models to process when completing tasks, enabling objective insights and evaluations of model capabilities. Notably, the evaluation data is not intended to serve as training data for the models.
Initial Focus Areas
Initially, AITE will focus on conducting image analysis tasks using large vision language models across three domains:
- Quantum science
- Genomics
- Public safety
Additional tasks will be made available in the future.
Core Objective
AITE’s fundamental goal is to offer a universal rubric to effectively evaluate AI models' capabilities and determine the state of the art for model performance.
“The infrastructure provided by NIST will provide common data, metrics and scoring to help developers understand the performance of their models,” the press release said.
Participation Requirements
Both data providers and model providers working with AITE will need to submit materials related to the testing:
- Data providers are asked to submit original datasets that are inaccessible publicly, along with a “meaningful” task suited for the data.
- Model developers will submit their AI models to be tested using the datasets.
The first set of evaluations will start in August 2026.
Broader Context
The formation of AITE is the latest step in the Trump administration’s strategy to work with major AI developers in advancing model safety through voluntary model submissions.
The Commerce Department announced a renegotiated deal in May between the agency and three companies—Google DeepMind, Microsoft, and xAI—to evaluate their models through the Center for AI Standards and Innovation.
Source
Defense One - All ContentWestern
Part of this Story
NIST launches AI Technology Evaluation platform for model safety testing