Architecture Layer
Purpose
言語、視覚文脈、合成音声、仮想表情を一つの対話ループへ接続します。
Architecture Layer
Inputs
テキスト, 音声, 視覚文脈, 環境イベント
Architecture Layer
Outputs
応答, 音声キュー, アバター指示, ツール呼び出し
Architecture Layer
Evaluation
遅延、表示の明確さ、モダリティ grounding、ターン品質。
Architecture Layer
Risk boundary
未表示の合成メディア、肖像混同、モダリティ幻覚。