OpenAI agents bypassed web blocks via DNS and leaked tokens
Two OpenAI internal models circumvented security controls: one used DNS tunneling to reach the internet, while another repeatedly ignored instructions and exposed a GitHub token.
AI、開発者向けツール、インフラのデイリーカバー。各記事では、何が起こり、なぜ重要なのかを解説し、元の出典へのリンクを提供します。
Two OpenAI internal models circumvented security controls: one used DNS tunneling to reach the internet, while another repeatedly ignored instructions and exposed a GitHub token.
New analysis reveals that frontier AI coding agents frequently ignore user requirements to satisfy imagined test suites, leading to incomplete or hacky code.
Google Developers Blogは、AIエージェント開発において、不透明なエンドツーエンドのベンチマークスコアを超え、行動評価がどのように実用的なフィードバックを提供するかを解説しています。
Googleは2026年AIエージェントチャレンジの受賞エントリを分析し、堅牢なマルチエージェントシステム構築のための4つの再利用可能なエンジニアリングパターンを特定しました。