検証: Piのadversarial-qaスキルを検証した

00:01:24 ・ source: https://zenn.dev/mskbhd/articles/lab-142-adversarial-qa

Transcript

hostenToday we explore Pi's adversarial-qa skill, which finds bugs by trying to break code with executable counterexamples.
guestjaPiというプラットフォームの adversarial-qa スキルについて見ていきます。これは実行可能な反例を使ってコードのバグを探す手法です。
hostenSurprisingly, this skill isn't code—it's a SKILL.md workflow document telling agents how to attack.
guestja驚くことに、これはコードライブラリではなく、AIエージェントに『どう攻めるか』の手順を指示するドキュメント形式なんです。
hostenThe author tested seven techniques like property checking, fuzzing, and mutation testing on buggy sort functions.
guestja著者は7つの手法(プロパティチェック、ファジング、ミューテーションテスト等)を実装してバグを検出しました。
hostenEach technique caught different bugs—combining them gave complementary coverage.
guestja各手法がそれぞれ異なるバグを検出し、組み合わせることで相補的なカバレッジが得られました。
hostenThe philosophy is 'falsify, not reassure'—actively hunt bugs instead of politely confirming code works.
guestja基本思想は『バグの不在を確認するのではなく、バグを積極的に探す』という反証主義です。
hostenThis approach challenges traditional code review AIs by proving behavior through concrete counterexamples, not just analysis.
guestja従来のコードレビューAIの読み取り中心の手法に対し、実行可能な反例で振る舞いを反証するというまったく新しいアプローチです。