Eris - #31
Open
grandchildrice wants to merge 77 commits into
Open
Conversation
- ASCON: Agentic Simulation for eCONomics competition - Slides: title, executive summary, background, competition design, team structure, sponsor benefits x2, packages, budget, roadmap, references, closing - Fix SL04 build error (removed invalid closing div tags) - Add slides/ directory with SL01-SL12.md - Update slides.md title and reduce to 12 slides only
- Action Title: convert all topic titles to assertive sentence titles - Skim Test: titles now tell a complete story when read in sequence - SCR structure: SL03 follows Situation/Complication/Question/Answer format - 3-Second Test: each slide conveys one clear idea - Bullet points minimized: tables used throughout for clarity - Speaker notes added/updated for all slides - SL09: added note that EF/Dentsu Research are not confirmed
- タイトル・コンセプトを「AIエージェントがテストしてプロトコル改善やセキュリティイシューを発見するテスト用L2 Rollup」に変更 - SL01: カバースライドをV2コンセプト(24/7 Stress Test, OpStack L2)に刷新 - SL02: エグゼクティブサマリーをV2の価値提供(DeFiテスト基盤)に更新 - SL03: 新規追加 — 静的テストから動的・自律的テストへのパラダイムシフトを説明 - SL04: アーキテクチャ図をDeFiプロバイダー→L2→AIエージェントのフローに変更 - SL05: 旧スライド削除(V1コンセプトの重複スライド) - SL10: スポンサーメリットをパラメータ最適化・セキュリティ発見に更新 - SL11: EIP規格テスト・分散型AIホスティング促進の価値を追加 - SL12: スポンサーティアにコントラクトデプロイ優先権を追加 - SL14: Phase詳細にL2基盤構築・EIPフィードバックを追記 - SL15: 参考文献にDeFiセキュリティ事例(Uniswap Bug Bounty等)を追加 - SL16: クロージングをDeFiプロトコルの未来を創る訴求に変更
- SL01: カバーを「AIエージェントが利益を競うDeFiテスト用公共財 L2 Rollup」に変更 - SL02: エグゼクティブサマリーのHowに「コントラクトをデプロイするだけで参加可能」を明記 - SL03: Answerを「コンペを通じた動的テスト環境(公共財)」として再定義 - SL04: アーキテクチャ図のL2を「テスト用公共財」と明記、03を「研究のための公共財」に変更 - SL10: スポンサーメリットの説明に「利益を競うAI」という文脈を追加 - SL11: メリット③を「DeFi研究のための公共財」に変更、④をEIP規格テストに整理 - SL12: タイトルを「DeFiの未来を創る公共財の実現者」に変更
- slides.md: タイトルを「AIエージェントが利益を競うDeFiテスト用コンペ基盤」に変更 - SL01: カバーを「一度構築すれば永続的に拡張できる」コンセプトに刷新、∞ Expandableを追加 - SL02: Howに「コントラクトをデプロイするだけ」、Visionに「永続的に拡張できるテスト基盤」を追記 - SL03: ASCON V2 → ASCONに統一 - SL04: アーキテクチャ図を「Step1: 基盤構築→Step2: 第1回ASCON→Step3+: コントラクト追加」の3ステップ構造に刷新 - SL10〜SL12: ASCON V2 → ASCONに統一 - SL15〜SL16: ASCON V2 → ASCONに統一
- SL04: scale-200を削除、mt/mbをコンパクトに、カードのpaddingとフォントサイズを縮小 - SL15: 1カラム→2カラムレイアウトに変更し全体のはみ出しを解消
Restructured the deck around a clear, fact-first arc: an accessible problem → the solution → the competition → fact-based sponsor value. Problem (two fact-based slides): - SL02b: crypto theft hit $3.4B in 2025, yet the worst attack classes (MEV, oracle manipulation, economic attacks) are explicitly out of scope for bug bounties — and the human audit model is collapsing (curl ended its bounty after AI-slop spam, HackerOne +210% AI reports, Code4rena announced wind-down). - SL02c (new): AI agents are already moving real money (Google AP2, x402; Gartner: $15T B2B by 2028) but there is no safe public arena to test agent-vs-agent behaviour before it touches the real economy. Solution: SL03 is now a full-bleed immersive scene — ASCON, the world's first blockchain for LLMs — using the company-deck city engine ported in as public/city, ethereum variant, with AI-agent inhabitants added: cyan Trader agents walking the avenues, red hooded Hacker agents that lunge at them and rob them. Sponsor value (SL10) rewritten in the same fact style — each pillar leads with a figure (唯一 / 100+ / 1,430 / $3.4B). Cover, executive summary and the sources slide updated to match. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Three corrections from review: - Naming: the blockchain L2 is "Eris"; "ASCON" is the recurring competition that feeds it. Renamed every L2/platform reference to Eris across the deck; kept ASCON only where it means the competition. - Agents: Trader, Hacker and Verifier all share one goal — maximising their own profit. Verifier is no longer framed as altruistic "defending"; it earns bounties by finding and proving vulnerabilities. It is the relentless competition between profit-seeking agents that hardens the city. - Sponsor-facing language (SL12-14): dropped internal fundraising framing. SL13 no longer states that platinum money is "all allocated to the prize" — the funding-source mapping and buffer line are removed; it is now a plain, sponsor-facing budget-transparency slide. SL12 drops the "prize allocation" perk; SL14 roadmap reworded (募集→確定) and the stray RWA mention removed. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
SL04 (the ASCON competition) was a wall of prose. Rebuilt it around a flywheel diagram — five numbered stages on a dashed ring (hold a competition → implementations gather → Eris hardens → discovery quality rises → industry attention → next round draws more) with the first- competition facts as compact figures beside it. SL10's title no longer talks about "showing value with facts" — it just states it: the four assets a sponsor walks away with. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Brings scripts/refine.sh, refine-step.mjs and the recursive/visual self-improvement prompts over from company-deck so the ascon-proposal deck has the same refine loop available. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Direct polish pass (the autonomous refine loop's slidev-export renderer hangs in this sandbox, so applied by hand): - SL05: crisper title — the EIP testbed framed as a concrete "only place in the world", not "副次機能:…としても機能します". - SL16: closing reframed from "DeFiプロトコルの未来" to the deck's actual thesis — a world where AI agents run the economy safely (Eris + ASCON). Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
p2-4 — cut the text, make the eye-flow obvious: - SL02 executive summary: the 7-row table is now a left-to-right flow of three cards — 課題 → Eris(解決) → ASCON(エンジン) — with Eris as the emphasised focal point and a thin Who/Vision/Bonus footer. - SL02b / SL02c: each opens on one huge number ($3.4B / $15T) as the entry point, then a single supporting band, then compact fact cards, then a bold conclusion bar — prose cut to labels. p9-12 — the four separate achievement slides are merged into one (SL07): a 2×2 grid that maps each result onto what it proves for Eris — Hacker AI, Verifier AI, Trader/insight, and the proof-engine core — each with its peer-reviewed venue. SL08/SL09/SL09b removed. SL13 budget: dropped the "we disclose the use of funds" pitch — it is just a plain breakdown now. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
- SL02c: rebuilt — the box-salad is gone. A clean left/right contrast
split by a single hairline: left = the reality (the $15T economy is
already here, AP2/x402), right = the gap (no safe place to test it),
carried by type hierarchy and whitespace, not borders.
- SL05: the EIP testbed is now a 4-step flow diagram (new EIP → deploy
on Eris → 100+ agents stress it → instant data), not prose.
- SL07: collapsed to three cards — Trader / Hacker / Verifier — with
theorem-proving folded into Verifier. Each card leads with the role
in plain words and explains the result for non-experts ("DEX" → "an
online exchange anyone can use"). Inline sources added.
- Sources: the standalone sources slide is moved to the very end as an
Appendix (after the closing); every content slide already cites
inline via SourceCite.
- New SL14b: "what the competition produces" — four ongoing research
themes (DeFi design proposals, new security baselines, zero-day
discovery, AI-agent-economy analysis).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Every content slide already cites its sources inline via SourceCite, so the standalone appendix list was redundant. Removed it; the deck now ends on the closing slide. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
- SL01: dropped the background photo and dark overlay; the cover is now a clean white slide with black serif type, matching the deck. - SL16: added a line break before the contact email so it sits clear of the organisation name. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
SL14b reorganised so the research output is honest about likelihood: - Certain — just holding ASCON yields the security-baseline paper and benchmark from the Hacker×Verifier match data. - Conditional — DeFi design proposals and zero-day write-ups happen if the imbalances / vulnerabilities are actually found. - Future — the AI-agent-economy analysis comes once data accumulates. Three columns, certainty decreasing left to right, the committed deliverable emphasised. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
Under the conditional ("発見があり次第") tier of SL14b, added a third
output: an active-cyber-defense dataset. Dynamic, attack-related data is
scarce in the field — Eris, where Hacker AIs attack continuously for
real stakes, can harvest real attack data and publish it as a research
dataset.
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
The "certain" tier of SL14b gains a second guaranteed output: with many AI implementations and their match results in hand, the competition itself yields research on winning-strategy analysis and on combining strong agents into higher-performance ones — no discovery required. Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
# Conflicts: # bun.lock # package.json # prompts/10_Recursive_Self_Improvement.md # prompts/11_Visual_Self_Improvement.md # scripts/refine-step.mjs # scripts/refine.sh
…n-agent into ascon-proposal
- p13 クロージングを簡潔化(テストベッド表現・体言止め・「一般社団法人」削除) - p14(参考)を削除し SL01-13 に - 表紙(p1)とクロージング(p13)の下部中央に Eris ロゴ(文字付きSVG)+Nyx ロゴを 横並び表示(区切り線なし)。この2枚は右下フッターを非表示に - global-bottom: フッター判定をテンプレートへ移し、中間スライドの Nyx ロゴ+ページ番号を復活(script setup 内 $nav 参照の誤作動を修正) - public/logos/eris_logo.svg 追加 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 起承転結ラベル・メタ語り(本題に入る前に/中心命題/前ページで等)を削除 - 各ノートを話す順の箇条書きに。事実・数字・固有名詞は保持 - 投影スライドには出ない(presenter view 専用) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- AIエージェントを node とし、戦略ごとに 1 node + service への labeled edge - 実サービス 5 種(Uniswap V3 / Balancer v2 / Curve / Aave v3 / GMX v2)+詐欺コントラクトを node 化 - 戦略は examples/agents 準拠で分割(即時流動性供給 / デルタ中立LP / マーケットメイク / 裁定 / レバレッジ運用 / 清算 / パーペチュアル / トレンド / 平均回帰 / 現物ヘッジ) - ハッキング node(Curve を攻撃)と詐欺・欺瞞(deploy → 別 AI が誤って利用)を severe で表現 - node サイズ統一・矢印は枠手前で停止・戦略名は和文で平易化・sources/薄文字は削除 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- p10 の後ろに実験設定ページを挿入(自作シミュレータの自己改善ループ) - SL09 の loop 図をそのまま流用し、検証=onchain state、意図=LLM が戦略別にパラメータ調整 tx を提出、に位置づけ調整 - 左リスト=戦略の調整(version 更新 / threshold↑ / size・slippage 締め / rollback) - 右リスト=onchain state フィードバック(約定・revert / PnL self↔frozen / ポジション / fair gap) - 上の青弧=検証→意図のフィードバック Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- タイトルを「自作シミュレータ実験: 設定」に変更 - 意図側リストを調整パラメータの網羅に変更(entry 閾値 / size・leverage / slippage・fee 上限 / LP レンジ・リスク guard) - PnL の (self↔frozen) 注記を削除 - SourceCite を削除 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 自作シミュ実験を「主張のある研究知見」に格上げする related-work 位置づけ - 命題(falsifiable)を bg-2 thesis として冒頭提示 - 4 系統の先行研究 × [示したこと / Eris が再現 / Eris の進化] の比較表 - El Farol / Minority Game(1994–97) - LLM Cannot Self-Correct(ICLR'24, Huang+) - Generation–Verification Gap(ICLR'25, 2412.02674) - LLM トレーディング agent(TradingAgents, 2412.20138) - 進化列を accent で payoff として強調 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 主役を「競争するAIは儲け方より損の減らし方を学ぶ」という一文の insight に - 個体レベル(自制=安心 / accent)→ 系全体(流動性枯渇=未解決リスク / severe)の micro→macro 対比を図に - 先行研究は「理論が予測 → Eris が実証」の1行 credibility footer に圧縮 - 専門語を平易化(GV-gap 等は footer のみ) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 実験結果の図を挿入(setup と insight の間): - 結果01: 図1(Dynamic vs Fixed) と 図2(変更回数 vs ΔPnL) を横並び/「優劣を決めるのは自己改善の回数ではない」 - 結果02: 図3(revert 率)/「勝った戦略の共通パターン: revert回数の減少」 - 結果03: 図4(パラメータ推移)/「勝った戦略の共通パターン: リスク削減・選別」 - 画像は public/images/eris/ に配置。img の max-width:none で原寸スケール可能に - p10: 各戦略名の下に code identifier を併記(crossvenue·cvbal / jitlp / dnlp / fairmm / aaveloop / liquidator / gmxperp / gmxtrend / gmxrev / spothedge)→ 結果図の凡例と接続 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- タイトルを1行に短縮(AIの優劣を決めるのは、自己改善の回数ではない。) - フック strip を追加:「いちばん多く自己改善した戦略が、いちばん負けた」(gmxtrend 43回・最下位) → 効くのは回数でなく〈変更の向き〉。結果02/03(リスク削減・選別)への橋渡し - 図は max-height 320 に調整しフックの余白を確保 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 「いちばん多く自己改善した戦略が…」フックを削除 - 図1/図2 を max-height 366・gap 縮小で拡大(横並びの幅制約内で最大化) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 主張を「自己改善は 生成ではなく 選別(リスク削減)に向かう」に - 3つの先行研究が別々の理由で同一方向を指す収束図: - GV-gap(ICLR'25)→検証で勝つ / Minority Game・El Farol(1994-97)→自制が勝つ / LLM自己修正(ICLR'24)→外部信号で効く - 中央 hub=選別・リスク削減(accent)。右へ exit=開いた問い「全員が同時に自制したら?→Eris が次に測る」(MGが残す問い・severe破線) - footer=我々の実証(revert↓ / 量でなく向き / Self–Frozen 反実仮想統制) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- ICLR'24/'25(LLM一般論)が古い/自明との指摘を受け、金融寄りの現行研究に差し替え - 3潮流: トレーディングエージェント(TradingAgents/AlphaAgent, alpha decay) / LLM市場シミュ(TwinMarket/ABM-LLM, 創発バブル) / 経済シミュ古典(Minority Game/El Farol, crowding) - 収束結論: 賢い自己改善=選別・リスク削減「構造が知能の向きを決める」 - 見出しを「賢い自己改善も、市場は選別へ引き戻す。」に - 開いた問いを市場シミュ研究の脆弱性ミクロ起源に接続。footer=我々の差分(敵対競争+本物の検証+Self-Frozen統制で向きを初測定) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- カードを「どんな研究か→何がわかったか」の平易な2段に - AIに売買させる研究→儲け口は競争ですぐ枯れる - 市場をAIで丸ごと再現→バブルや暴落がひとりでに起きる - 経済モデルの古典→儲けの奪い合いは早い者勝ち - 専門語(alpha decay/crowding/創発/選別)を日常語に翻訳。論文名は小タグへ格下げ - 見出し「賢いAIほど、新しい儲けより〈守り〉を選ぶ。」/ hub「〈守り〉に入る」 - 正確な研究内容はスピーカーノートへ(口頭補足想定) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 「3研究が収束」を解消し、左→右の流れに再構成 - ① 先行研究(理論・部分実証):平易な3findings+論文名タグ - ② Eris の貢献(accent強調):①を本物のDeFi市場で「再現」かつ自己改善の〈向き〉を「実測(強化)」→賢いAIも守りに入る - ③ 前向きな射程(accent破線):理論止まりも Eris なら測れる(severe不安でなくcapabilityとして) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 主張を「作り物でない市場で、理論を再現した」に - 左=先行2系統(bot研究/sim研究)、各gapをsevereで明示 - アーム=Erisが足すもの(+マクロ / +本物の機構・転移) - 右=Erisの達成(accent): 本物のDeFiで敵対競争を再現→理論を再現 - footer=DeFiだからできた(認識論キッカー): 作り物のマクロは仮定が混じる/本物の機構から創発したマクロは現実に転移 - 考察・仮説はp16へ送る方針 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 横=理論↔実装、縦=モデル↔現実 の4象限マップに - 理論×モデル: 経済モデルの古典 / 実装×モデル: LLM経済シミュ / 実装×現実: AI trading bot - 理論×現実(空白): Eris をハイライト(accent-soft zone+ringマーカー) - 「マクロも測れる」等の明言を削除(視聴者に空白象限から気付かせる) - 前向きの接続を「現実の検証環境で意図と検証をどこまで洗練できるか」に変更(意図/検証ループへ戻す) - SourceCite を追加(TradingAgents/AlphaAgent/TwinMarket/ABM-LLM/Minority Game・El Farol) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 理論↔実装だと Eris が中間で曖昧 → 横軸を「協調・単体↔敵対競争」に - 協調×モデル: LLM経済シミュ / 敵対×モデル: 経済モデルの古典 - 協調×現実: AI trading bot / 敵対×現実: Eris(空き角を独占) - 各点に「研究名+示したこと」を併記、象限を拡大 - 「ここから問う」callout削除。見出し「本物の敵対競争だから、理論が現実に現れた。」 - Eris=Minority Gameを本物にした、の対角ストーリーが立つ Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- Eris の丸ポチを Eris ロゴ(accent-soft カード)に差し替え(他点は dot のまま) - 再現できた理由を平易に追記:「勝つAIほど、危ない取引を減らしたから」 - タイトルを「賢いAIでも、経済の法則からは逃げられない。」にNHE強化 (構造=経済法則を敵に置く/provocativeだが誠実/暴走等の overclaim は回避) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 論点①: 速さでなく適応(vs bot)/論点②: 多様性とコストの経済学(実開発と地続き) - 仮説として明示し、Erisで確かめにいける、で締め Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- Screen Recording(.mov 62MB/3152×1768) を mp4 に変換(ブラウザchrome をcrop・1600幅・無音・faststart・2.5MB) - poster フレームを抽出 - SL10b1: 実験設定の直後にデモ(Self/Frozen 並走・エージェントメッシュ・tx フィード)。autoplay/loop/muted - 「そして、実際に走らせた。」で setup→demo→結果 へ繋ぐ Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 2論点の考察を、Next Step の4論点に集約(まとめ) - 01 botと何が違うか / 02 戦略の多様化はどう起きるか - 03 守りのAIに渡す認証・権限は(KYA/委任=意図側設計) - 04 役割の多様化は何をもたらすか(ハッカーAI・別手段で稼ぐAI・α一発狙い) - 2×2 numbered grid。Eris セクションの forward 締めに Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 各 ns-e を「問いの説明」→「本当にそうなる?と思わせる具体的仮説」に書き換え - 01 平常→急変の対比でAIの適応力 / 02 勝ち筋=情報効率の意外性 - 03 自制するなら実データで権限を決められる期待 / 04 1体の逸脱で自制は崩れるかの緊張 - 論点タイトルは維持 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 各カードを「前提→展開→知りたいこと」の段階的説明に拡張(枠を下に拡大) - 詩的な em-dash 演出を排し、具体(計算コスト/KYA/契約の穴 等)で接地 - ns-e 12.5px・カード縦伸長で4〜5行を収容 Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 「設計が要る」止まりから具体へ: - 静的な許可リストでなく振る舞い連動の動的認可 - 自制している間だけ権限拡大、revert増で自動で絞り剥奪 - 前提=守っている事実を検証できること(KYA+自制の実績を信頼の根拠に) Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
- 「多くを任せられる」から方向転換:リスク低減が最適解なら、権限管理はセキュリティに近い問題に - 核心=AIの逸脱をアルファ(機会)かリスク(脅威)か見分ける「検証能力」。それが権限の上限を決める Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
No description provided.