ジャーナル
ジャーナル
エージェントと構造化データに関するフィールドノート——モデルがSQLを生成するときに何が壊れるか、もっともらしい数値がなぜ回答ではないか、そして根拠を示せる数値を出力するために何が必要か。
IIエントリー一覧
The 15%-off dashboard: wrong joins corrupt analytics silently
A generated join that fans out rows doesn't crash. It inflates a metric by 15% and lets everyone keep steering by it. Silent wrongness is the real cost of text-to-SQL.
Replay What Your Agent Answered: AI Audit Trails That Hold
Only 17% of organizations can reconstruct what their agents did. From 2 August 2026 the EU AI Act expects the record to exist. Replay is how it holds.
The confidently wrong revenue number: why LLMs can't add
A model hands you a revenue figure with perfect posture and no arithmetic behind it. The research says the failure is structural — and so is the fix.
Prompt injection is the new SQL injection — the fix isn't SQL
P2SQL, CVE-2024-5565, EchoLeak, the Supabase leak: hostile SQL now arrives in the model's output. The fix is an agent with no syntax to inject into.
Read-only that wasn't: the week AI agents deleted production
An agent deleted a production database. Another destroyed user files. A third shipped a wipe order to a million installs. One July week — one shared write path.
検証されることを
前提に書く。
このジャーナルのすべての数値はケイパビリティ契約に遡ることができます
IVテーマ一覧