検証: 個人AIエージェント基盤『三層構造』を11アプリ・27スキルで作って動かした
00:01:12 ・ source: https://zenn.dev/mskbhd/articles/lab-221-exp-note
Transcript
hostenToday we examine how one engineer built an AI agent platform with 11 apps and 27 skills.
guestja1人のエンジニアが、11個のアプリと27個のスキルを持つAIエージェント基盤を実装した検証記事ですね。
hostenHe discovered that scaling from 5 to 11 apps revealed a critical design flaw with source reference types.
guestja5個から11個にアプリを増やしたときに初めて、出典IDの型が統一されていないという設計の弱さが表面化しました。
hostenHe also found that only 2 of 10 automations were truly idempotent when tested multiple times.
guestja10個の自動化プロセスのうち、実は2個しべき冪等性がなく、3個は重複レコードを作り続けることが判明しました。
hostenDespite 23 passing tests and strong quality checks, external review found 3 critical blind spots he missed.
guestja23個のテストがすべてパスしていても、外部レビューで参照の整合性検証が7件全て未保証など、3つの重大な問題が見つかりました。
hostenHis key insight: self-written tests only catch what you thought to test; external review is essential.
guestja最大の学びは、自分で書いたテストは自分が想定した範囲の欠陥しか見つけられず、別の視点からの検査が同じくらい重要だということです。