Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

NIST unveils new AI evaluation platform

Дата публикации: 27-07-2026 20:56:00

The AI Technology Evaluation will provide exclusive data to grade models’ performance in select areas.

Основное содержимое страницы с новостью.

860x394.jpg?1785230250

R. Wilson / NIST

The National Institute of Standards and Technology launched a new program on Monday granting researchers access to an isolated testbed environment to safely evaluate artificial intelligence models against various commands. 

The AI Technology Evaluation, or AITE, is a voluntary testing vehicle focused on AI model safety analysis. It provides blind data for models to process when completing tasks to gain objective insights and conduct evaluations of model capabilities. Notably, the evaluation data is not intended to serve as training data for the models.

Initially, AITE will focus on conducting image analysis tasks using large vision language models across three domains: quantum science, genomics and public safety. More tasks will be available in the future.

AITE’s fundamental goal is to offer a universal rubric to effectively evaluate AI models' capabilities and determine the state of the art for model performance. 

“The infrastructure provided by NIST will provide common data, metrics and scoring to help developers understand the performance of their models,” the press release said.

Both data providers and model providers working with AITE will need to submit materials related to the testing. Data providers are asked to submit original datasets that are inaccessible publicly, along with a “meaningful” task suited for the data. 

Model developers, likewise, will submit their AI models to be tested using the datasets. The first set of evaluations will start in August 2026. 

The formation of AITE is the latest step in the Trump administration’s strategy to work with major AI developers in advancing model safety through voluntary model submissions.

The Commerce Department announced a renegotiated deal in May between the agency and three companies—Google Deepmind, Microsoft and xAI—to evaluate their models through the Center for AI Standards and Innovation.

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1ГП тестирует искусственный интеллект для оценки коррупционных факторов в нормативных актах0016-11-2021
2New tool detects hidden bias in medical AI07.6622-07-2026
3Bill Proposes Federal Agency to Focus on AI0726-06-2026
4Nvidia unveils new AI model and expands Japan’s physical AI ecosystem0516-07-2026
5ВТБ разработает «этического цензора» для контроля ИИ-помощников0506-10-2025
6NtechLab: искусственный интеллект будет следить за безопасностью во дворах регионов РФ0020-03-2025
7"СберМобайл" представил обновленную платформу искусственного интеллекта вещей с AI-агентом0004-09-2025
8New AI center works to bridge government and AI industry06.0530-07-2026
9ЕС представил новую киберстратегию0016-12-2020
10Stable-GFlowNet uncovers more hidden weaknesses in generative AI models07.4929-07-2026

Классификация: Пресс-релизы. Схожих патентов: 0. Схожих новостей: 10. Тональность: 0. Информативность: 8.19. Источник: www.defenseone.com.