ボトルネックはモデルではありません。 ハーネスです.
ハーネスエンジニアリング:エージェントが信頼性とガバナンスを持って本番環境に到達できるよう、モデルを取り囲むものを構築する新興の規律。
- AI責任者 · Google Cloud LATAM
- 元NTT DATA
- エージェント型AI
- 対話型AI
- ロボティクス
- プロダクト
- エンタープライズ・トランスフォーメーション
EUREKA、隔週で受信トレイにお届けします
AI、エージェント、ロボティクス — デモから現実世界でのインパクトへと移行するためのアイデア、講演、分析。英語とスペイン語で提供。
Substackで確認を行います。スパムはありません。いつでも購読解除できます。
私が確信している5つのテーゼ
AIで起きていることをどう読み解くか、および真の機会がどこにあるかを形作る信念。
- 01
ボトルネックはモデルではなく、ハーネスにある。
AIシステムの真の性能は、モデルそのものよりも、ツール、コンテキスト、検証、メモリ、制限など、モデルを取り巻く環境に依存します。
- 02
コパイロットからデジタル従業員へ。
自律型エージェントはもはや支援にとどまりません。プロセス全体をエンドツーエンドで実行します。これにより、価値、リスク、そしてROIの測定方法が変わります。
- 03
ロボティクスはついにその頭脳を手に入れた。
基盤モデルこそ、ロボティクスが40年間待ち望んでいたものです。身体性AI(Embodied AI)は、産業、科学、社会における次のフロンティアです。
- 04
成果あたりのコストが、どのエージェントが生き残るかを決定する。
トークン、リトライ、人の目による監視。すべてのエージェントにP&L(損益)が存在します。優位性とは、単にエージェントを保有していることではなく、どのエージェントが自活できているかを把握し、それを証明できることにあります。
- 05
人々に届かなければ、それはまだデモに過ぎない。
価値ある応用AI = ビジネス + 社会 + 人々。これら3つすべてが揃わなければ、真のインパクトは生まれません。
ここから始める
エージェント型AI、基盤モデル、そしてそれらをリリースする現実など、現在の考えを捉えた厳選された記事。
The agentic enterprise needs an immune system
How to design AI agents that can act without concentrating trust: identity, least privilege, containment, oversight, and recovery.
The Stopwatch and the Exam
Static AI benchmarks did not die; they stopped being enough. A grounded look at the move from capability to agency, the harness problem, and where to put attention now. Backed by a public catalog of 69 agentic benchmarks.
Cognitive Architecture and Emergent Phenomena in Advanced AI
A synthesis of current mechanistic and behavioral research on how cognitive structures and emergent behavior arise in modern AI systems.
The next generation of AI: Self-Improvement and Autonomous Learning
Self-improvement and autonomous learning in AI — and the road to an intelligence explosion.
執筆テーマ
ライブラリ全体で繰り返し登場するテーマ。任意のトピックでライブラリをフィルタリングできます。
プロトタイプ、ベンチマーク、ツール
AI、エージェント、複雑なシステムを探索するためのインタラクティブな実験。まずはオープンなAIエージェント・ベンチマーク・トラッカーから。
インタラクティブなサンドボックス:6つの軸でエージェントを構成し、OWASP LLM Top 10の脅威を選択し、エージェンティック・コントロール・マトリクスの18のコントロールのオン/オフを切り替えて、カバレッジ(埋められた象限、空のままの象限、対応するフレームワーク)を確認できます。残存リスクスコアは意図的に排除されています。
OpenAI、Anthropic、Google、Microsoft、Amazon、SpaceXAIからオープンソースの最前線に至る24のAIエージェントを、6つの直交する軸とアクションサーフェスにわたるベクトルとしてエンコードするインタラクティブなタキソノミー。フィルター可能なテーブル、A×T×Iガバナンス・リスク・マトリクス、機械読み取り可能なJSON APIを備えています。
レコメンデーション
私が立ち上げたチーム、共に働いたエグゼクティブ、そして一緒にプロダクトをリリースしたエンジニアからの声。
“During the time we worked together he demonstrated a remarkable vision to apply technology in order to solve real daily problems. His passion and proactivity makes him achieve what sometimes seems impossible. Very focused on customer needs, always a pleasure to work with Santiago!”
Agustin Aznar GabásHead of Retail (Spain, Portugal & Italy) · Executive Committee“Santi worked with me for 10 years. His ability to set up teams and his foresight in detecting new tech trends, that will disrupt the market is outstanding. He ran the product management of a set of payment products, which are successfully used today by Spanish Merchants, Banks, and Users.”
José Manuel ReyesPayment Processing Specialist“Santi is constantly surprising with innovative and disruptive products with huge potential to change the market. He is always ahead of the pack in terms of technology and how it can disrupt businesses.”
Angel Luis de la MoyaProduct Manager · Getnet Europe (a Santander Company)“I worked on Santi's team for 4 years. He has a surprising drive, he is able to break down all the barriers to launch a product. Due to his software engineering background he understands the technical challenges, which makes communication between clients and technical staff easy.”
Héctor Abraham Morillo PrietoSoftware Engineer · Automattic
発信チャネル
同じアイデアを異なるフォーマットで。読む、見る、聴くなど、ご自身のスタイルに合ったチャネルをお選びください。
運営者について
物理学者としてのバックグラウンドを持ち、現在はGoogle CloudのLATAMにおけるAI Forward Deployed Engineers(FDE)部門責任者。AIの変革を理解し、実際にプロダクトを届ける際に何が有効かを共有するために執筆しています。
