IAPS:开发者内部AI模型风险报告(英文版)
IAPS:开发者内部AI模型风险报告(英文版).pdf |
下载文档 |
资源简介
Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a possible public release. For example, Anthropic recently developed a new class of model with advanced cyberoffense-relevant capabilities, Mythos Preview, which was available internally for at least six weeks before it was publicly announced. This internal use creates risks that external deployment frameworks may fail to address. Legal frameworks,
本文档仅能预览20页


