Вход на сайт

Просмотр новости

Найдите то, что Вас интересует

The tests that grade AI may be getting it wrong

Дата публикации: 30-09-2026 13:20:14

Before a new AI model reaches the public, its developers run it through a battery of tests known as "benchmarks," which score it on everything from reasoning ability to how safe it is for people to use. Billions of investment dollars ride on these benchmark scores, and policymakers increasingly cite them to shape regulations and government procurement decisions that will guide the development of AI.

Основное содержимое страницы с новостью.

🛡️

Just a quick check

We’re checking your connection to prevent automated abuse

Схожие новости

#Наименование новостиТональностьИнформативностьДата публикации
1Anthropic and OpenAI sound the alarm on AI safety and seek to shape how it’s controlled07.7628-09-2026
2Q&A: Expert says people often confuse behavior with intention with AI07.8530-09-2026
3Бизнес стал внимательнее просчитывать жизненный цикл ИИ-проектов после запуска014.3828-09-2026
4В тумане метрик. Получится ли у государства измерить эффект от внедрения ИИ015.4901-10-2026
5Even AI can’t perfectly assemble Ikea furniture—yet012.2328-09-2026
6Google restricts access to new AI model over safety concerns010.3301-10-2026
7Microsoft drafts ‘humanist’ AI code of conduct amid safety debate06.9615-09-2026
8OpenAI scraps release of new AI model over safety concerns — WSJ06.8129-09-2026
9OpenAI hack sparks further concern over AI models going rogue07.0128-09-2026

Классификация: Наука. Схожих патентов: 0. Схожих новостей: 9. Тональность: 0. Информативность: 6.19. Источник: techxplore.com.