vibes-plug 2.5.0 → 2.11.0

This diff represents the content of publicly available package versions that have been released to one of the supported registries. The information contained in this diff is provided for informational purposes only and reflects changes between package versions as they appear in their respective public registries.
Files changed (177) hide show
  1. package/.claude/rules/vibes-plug-core.md +32 -0
  2. package/.cursor/rules/vibes-plug-core.mdc +51 -0
  3. package/.cursorrules +42 -0
  4. package/AGENTS.md +37 -7
  5. package/BLUEPRINT.md +309 -217
  6. package/CHANGELOG.md +133 -1
  7. package/CLAUDE.md +70 -0
  8. package/LICENSE +1 -1
  9. package/README.md +641 -415
  10. package/index.js +19 -0
  11. package/package.json +44 -8
  12. package/plugin.json +24 -7
  13. package/scripts/generate_swarm_gif.py +295 -0
  14. package/scripts/install.js +201 -0
  15. package/skills/accessibility-testing-expert/SKILL.md +116 -0
  16. package/skills/ai-cost-token-optimizer/SKILL.md +82 -52
  17. package/skills/ai-evals-benchmark-expert/SKILL.md +188 -0
  18. package/skills/ai-llm-integration-expert/SKILL.md +185 -178
  19. package/skills/ai-media-generation-expert/SKILL.md +172 -0
  20. package/skills/ai-prompt-engineering-expert/SKILL.md +84 -0
  21. package/skills/angular-expert/SKILL.md +148 -0
  22. package/skills/api-design-expert/SKILL.md +6 -3
  23. package/skills/api-gateway-proxy-expert/SKILL.md +81 -0
  24. package/skills/app-analyzer-optimizer/SKILL.md +6 -3
  25. package/skills/apple-ecosystem-expert/SKILL.md +144 -141
  26. package/skills/{asisten_ramah → asisten-ramah}/SKILL.md +7 -1
  27. package/skills/astro-framework-expert/SKILL.md +200 -0
  28. package/skills/async-queue-temporal-expert/SKILL.md +210 -24
  29. package/skills/authentication-identity-expert/SKILL.md +278 -275
  30. package/skills/auto-doc-updater/SKILL.md +7 -1
  31. package/skills/autonomous-chaos-monkey/SKILL.md +63 -63
  32. package/skills/autonomous-red-teamer/SKILL.md +172 -28
  33. package/skills/autonomous-tdd-debugger/SKILL.md +70 -64
  34. package/skills/background-jobs-queue-expert/SKILL.md +235 -0
  35. package/skills/biome-linter-formatter-expert/SKILL.md +89 -0
  36. package/skills/blockchain-web3-expert/SKILL.md +115 -0
  37. package/skills/bootstrap-to-modern/SKILL.md +9 -6
  38. package/skills/brainstorming/SKILL.md +58 -50
  39. package/skills/browser-automation-expert/SKILL.md +197 -21
  40. package/skills/bun-runtime-expert/SKILL.md +7 -1
  41. package/skills/chatbot-messaging-expert/SKILL.md +114 -0
  42. package/skills/ci-cd-devops-architect/SKILL.md +45 -36
  43. package/skills/cloud-hosting-expert/SKILL.md +7 -1
  44. package/skills/coderabbit/SKILL.md +7 -1
  45. package/skills/compliance-gdpr-privacy-expert/SKILL.md +85 -0
  46. package/skills/cron-scheduler-expert/SKILL.md +303 -297
  47. package/skills/data-pipeline-etl-expert/SKILL.md +84 -0
  48. package/skills/data-telemetry-expert/SKILL.md +7 -1
  49. package/skills/data-visualization-expert/SKILL.md +154 -0
  50. package/skills/database-migration-versioning-expert/SKILL.md +90 -0
  51. package/skills/database-orm-expert/SKILL.md +13 -3
  52. package/skills/dependency-upgrade-migrator/SKILL.md +300 -294
  53. package/skills/design-system-architect/SKILL.md +278 -259
  54. package/skills/desktop-electron-expert/SKILL.md +128 -0
  55. package/skills/documentation-site-expert/SKILL.md +59 -0
  56. package/skills/doku-mcp-server/SKILL.md +7 -1
  57. package/skills/doku-payment-gateway/SKILL.md +7 -1
  58. package/skills/domain-driven-design-expert/SKILL.md +82 -0
  59. package/skills/e2e-testing-expert/SKILL.md +7 -1
  60. package/skills/ecommerce-expert/SKILL.md +87 -0
  61. package/skills/edge-serverless-db-expert/SKILL.md +98 -42
  62. package/skills/email-notification-expert/SKILL.md +367 -361
  63. package/skills/error-resilience-expert/SKILL.md +485 -479
  64. package/skills/event-driven-architect/SKILL.md +7 -1
  65. package/skills/feature-flag-analytics-expert/SKILL.md +65 -45
  66. package/skills/file-upload-media-expert/SKILL.md +436 -430
  67. package/skills/firebase-security-expert/SKILL.md +7 -1
  68. package/skills/form-validation-expert/SKILL.md +406 -400
  69. package/skills/fullstack-expert/SKILL.md +60 -1
  70. package/skills/gemini-agent-booster/SKILL.md +173 -135
  71. package/skills/geospatial-maps-expert/SKILL.md +80 -0
  72. package/skills/global-a11y-i18n-expert/SKILL.md +7 -1
  73. package/skills/glsl-shader-expert/SKILL.md +106 -100
  74. package/skills/go-programming-expert/SKILL.md +21 -15
  75. package/skills/graph-rag-knowledge-expert/SKILL.md +159 -0
  76. package/skills/graphql-apollo-expert/SKILL.md +113 -107
  77. package/skills/headless-cms-expert/SKILL.md +181 -0
  78. package/skills/hig/SKILL.md +7 -1
  79. package/skills/js-backend-expert/SKILL.md +218 -216
  80. package/skills/legacy-code-translator/SKILL.md +70 -64
  81. package/skills/local-slm-edge-ai-expert/SKILL.md +167 -0
  82. package/skills/logging-error-tracking-expert/SKILL.md +343 -337
  83. package/skills/mcp-client-orchestrator/SKILL.md +75 -69
  84. package/skills/mcp-server-architect/SKILL.md +294 -194
  85. package/skills/micro-frontend-architect/SKILL.md +111 -105
  86. package/skills/mobile-expo-expert/SKILL.md +8 -2
  87. package/skills/mobile-push-notification-expert/SKILL.md +70 -50
  88. package/skills/modern-css-native-expert/SKILL.md +189 -0
  89. package/skills/monday-design-aesthetic/SKILL.md +7 -1
  90. package/skills/monorepo-architect/SKILL.md +7 -1
  91. package/skills/mpa-orchestrator/SKILL.md +20 -1
  92. package/skills/multi-agent-orchestration/SKILL.md +254 -234
  93. package/skills/multiple-entry-points/SKILL.md +37 -1
  94. package/skills/mvc-expert/SKILL.md +7 -1
  95. package/skills/n8n-automation-expert/SKILL.md +89 -0
  96. package/skills/nextjs-app-router-expert/SKILL.md +148 -0
  97. package/skills/openapi-swagger-codegen-expert/SKILL.md +67 -0
  98. package/skills/payment-gateway-expert/SKILL.md +85 -1
  99. package/skills/pdf-document-generation-expert/SKILL.md +91 -0
  100. package/skills/performance-web-vitals/SKILL.md +7 -1
  101. package/skills/post-quantum-crypto-migrator/SKILL.md +57 -57
  102. package/skills/prd-architect/SKILL.md +7 -1
  103. package/skills/proactive-background-watcher/SKILL.md +67 -61
  104. package/skills/production-ready-hardener/SKILL.md +461 -455
  105. package/skills/project-context-mapper/SKILL.md +84 -78
  106. package/skills/pwa-offline-first-expert/SKILL.md +185 -0
  107. package/skills/python-programming-expert/SKILL.md +407 -401
  108. package/skills/rate-limit-abuse-prevention/SKILL.md +376 -370
  109. package/skills/realtime-collaboration-expert/SKILL.md +55 -1
  110. package/skills/rich-text-editor-expert/SKILL.md +177 -0
  111. package/skills/rust-programming-expert/SKILL.md +7 -1
  112. package/skills/saas-billing/SKILL.md +7 -1
  113. package/skills/saas-multi-tenant/SKILL.md +7 -1
  114. package/skills/saas-mvp-launcher/SKILL.md +20 -1
  115. package/skills/saas-transformer/SKILL.md +499 -488
  116. package/skills/scalability-clean-code/SKILL.md +7 -1
  117. package/skills/search-engine-expert/SKILL.md +89 -0
  118. package/skills/secure-fuzz-testing/SKILL.md +7 -1
  119. package/skills/self-evolving-memory-graph/SKILL.md +90 -74
  120. package/skills/self-healing-cloud-orchestrator/SKILL.md +57 -57
  121. package/skills/senior-frontend/SKILL.md +141 -161
  122. package/skills/seo/SKILL.md +41 -17
  123. package/skills/session-context-loader/SKILL.md +82 -76
  124. package/skills/session-handoff-resume/SKILL.md +7 -1
  125. package/skills/{skill_baru → skill-baru}/SKILL.md +8 -2
  126. package/skills/solidjs-expert/SKILL.md +80 -0
  127. package/skills/spa-orchestrator/SKILL.md +20 -1
  128. package/skills/sse-websocket-streaming-expert/SKILL.md +93 -0
  129. package/skills/state-management-expert/SKILL.md +7 -1
  130. package/skills/supabase-migration/SKILL.md +47 -1
  131. package/skills/supabase-security-expert/SKILL.md +7 -1
  132. package/skills/svelte-sveltekit-expert/SKILL.md +91 -0
  133. package/skills/svg-animation-motion-expert/SKILL.md +115 -0
  134. package/skills/tailwind-expert/SKILL.md +88 -136
  135. package/skills/tanstack-query-expert/SKILL.md +7 -1
  136. package/skills/tauri-expert/SKILL.md +7 -1
  137. package/skills/token-saver/SKILL.md +1 -1
  138. package/skills/typescript-expert/SKILL.md +12 -6
  139. package/skills/ui-components-expert/SKILL.md +165 -279
  140. package/skills/ui-ux-pro-max/SKILL.md +23 -3
  141. package/skills/vector-db-rag-expert/SKILL.md +175 -19
  142. package/skills/vibe-code-gardener/SKILL.md +1 -1
  143. package/skills/visual-qa-vision-agent/SKILL.md +70 -64
  144. package/skills/voice-ai-realtime-agent/SKILL.md +202 -0
  145. package/skills/vue-frontend-expert/SKILL.md +131 -125
  146. package/skills/wasm-edge-computing-expert/SKILL.md +97 -0
  147. package/skills/web-3d-graphics-expert/SKILL.md +136 -130
  148. package/skills/web-game-engine-expert/SKILL.md +101 -95
  149. package/skills/web-scraper/SKILL.md +157 -207
  150. package/skills/website-design-cloner/SKILL.md +179 -173
  151. package/skills/webxr-ar-vr-expert/SKILL.md +122 -116
  152. package/skills/wordpress-headless-expert/SKILL.md +144 -0
  153. package/skills/zero-to-prod-orchestrator/SKILL.md +52 -27
  154. package/skills/zero-trust-secret-vault/SKILL.md +87 -39
  155. package/.github/ISSUE_TEMPLATE/feature_request.md +0 -20
  156. package/.github/workflows/publish.yml +0 -20
  157. package/CONTRIBUTING.md +0 -199
  158. package/SECURITY.md +0 -21
  159. package/banner.png +0 -0
  160. package/skills/autonomous-swarm-director/SKILL.md +0 -69
  161. package/skills/hyper-context-synthesizer/SKILL.md +0 -55
  162. package/skills/llm-cost-arbitrage-router/SKILL.md +0 -59
  163. package/skills/senior-fullstack/SKILL.md +0 -167
  164. package/skills/senior-fullstack/references/architecture_patterns.md +0 -160
  165. package/skills/senior-fullstack/references/development_workflows.md +0 -222
  166. package/skills/senior-fullstack/references/tech_stack_guide.md +0 -190
  167. package/skills/senior-fullstack/scripts/code_quality_analyzer.py +0 -114
  168. package/skills/senior-fullstack/scripts/fullstack_scaffolder.py +0 -114
  169. package/skills/senior-fullstack/scripts/project_scaffolder.py +0 -114
  170. package/skills/seo-aeo-landing-page-writer/SKILL.md +0 -97
  171. package/skills/seo-geo/SKILL.md +0 -188
  172. package/skills/ui-ux-pro-max/scripts/__pycache__/core.cpython-310.pyc +0 -0
  173. package/skills/ui-ux-pro-max/scripts/__pycache__/core.cpython-312.pyc +0 -0
  174. package/skills/ui-ux-pro-max/scripts/__pycache__/design_system.cpython-310.pyc +0 -0
  175. package/skills/ui-ux-pro-max/scripts/__pycache__/design_system.cpython-312.pyc +0 -0
  176. package/skills/ui_ux_expert/SKILL.md +0 -125
  177. package/vibes-swarm-demo.gif +0 -0
@@ -1,7 +1,7 @@
1
1
  ---
2
2
  name: mpa-orchestrator
3
3
  description: "Orchestrates Multi-Page Application (MPA) architecture within a single repository, integrating with relevant skills / Mengorkestrasi arsitektur Multi-Page Application (MPA) dalam satu repositori, terintegrasi dengan skill relevan lainnya."
4
- author: "Antigravity"
4
+ author: "Roedy Rustam"
5
5
  ---
6
6
 
7
7
  # Multi-Page Application (MPA) Orchestrator (2026 Edition)
@@ -13,6 +13,9 @@ author: "Antigravity"
13
13
  <a name="english"></a>
14
14
  ## English
15
15
 
16
+ ### Orchestration & Integration
17
+ Connects and orchestrates with relevant domain skills like `modern-web-guidance`, `brainstorming`, `zero-to-prod-orchestrator`, and `project-context-mapper` to ensure cohesive execution.
18
+
16
19
  ### Description
17
20
  A structured approach for building and orchestrating Multi-Page Application (MPA) architectures within a single repository. Acts as an orchestrator connecting MPA principles with specialized skills (like `mvc-expert`, `saas-multi-tenant`, `senior-frontend`) to deliver cohesive, server-rendered applications. In 2026, MPAs are increasingly built with **Astro 5** for content-heavy sites or **traditional server frameworks** (Laravel, Django, Go) augmented with **HTMX 2** or **Alpine.js** for interactivity.
18
21
 
@@ -47,6 +50,7 @@ A structured approach for building and orchestrating Multi-Page Application (MPA
47
50
  | React MPA with SSR | **Next.js 15** Pages Router or App Router |
48
51
 
49
52
  ### Orchestration Guidelines
53
+ - **With `modern-web-guidance`**: MUST execute FIRST before any HTML/CSS task to ensure modern frontend standards.
50
54
  - **With `mvc-expert`**: Enforce MVC pattern — Controllers handle logic, Views handle rendering.
51
55
  - **With `saas-multi-tenant`**: Integrate tenant identification in core routing middleware — every page load initializes `tenant_id` context securely.
52
56
  - **With `senior-frontend` / `ui-ux-pro-max`**: Enhance with Alpine.js for reactive UI or HTMX 2 for HTML-driven partial updates — no heavy client-side bundles.
@@ -65,6 +69,9 @@ A structured approach for building and orchestrating Multi-Page Application (MPA
65
69
  <a name="bahasa-indonesia"></a>
66
70
  ## Bahasa Indonesia
67
71
 
72
+ ### Integrasi Orkestrasi
73
+ Terhubung dan mengorkestrasi skill domain yang relevan seperti `modern-web-guidance`, `brainstorming`, `zero-to-prod-orchestrator`, dan `project-context-mapper` untuk memastikan eksekusi yang kohesif.
74
+
68
75
  ### Deskripsi
69
76
  Pendekatan terstruktur untuk membangun dan mengorkestrasi arsitektur Multi-Page Application (MPA) di dalam satu repositori. Bertindak sebagai orkestrator yang menghubungkan prinsip MPA dengan skill spesialis lain untuk menghasilkan aplikasi server-rendered yang kohesif dan modern. Di 2026, MPA semakin banyak dibangun dengan **Astro 5** untuk situs konten-berat atau framework server tradisional yang diperkuat dengan **HTMX 2** atau **Alpine.js**.
70
77
 
@@ -87,6 +94,7 @@ Pendekatan terstruktur untuk membangun dan mengorkestrasi arsitektur Multi-Page
87
94
  | React MPA dengan SSR | **Next.js 15** Pages Router atau App Router |
88
95
 
89
96
  ### Panduan Orkestrasi
97
+ - **Dengan `modern-web-guidance`**: WAJIB dieksekusi PERTAMA KALI sebelum tugas HTML/CSS untuk memastikan standar frontend modern.
90
98
  - **Dengan `mvc-expert`**: Pastikan MPA mengikuti pola MVC secara ketat.
91
99
  - **Dengan `saas-multi-tenant`**: Integrasikan identifikasi tenant di middleware routing inti.
92
100
  - **Dengan `senior-frontend` / `ui-ux-pro-max`**: Tingkatkan dengan Alpine.js atau HTMX 2 untuk interaksi tanpa bundle berat.
@@ -99,3 +107,14 @@ Pendekatan terstruktur untuk membangun dan mengorkestrasi arsitektur Multi-Page
99
107
  - Membangun situs konten, halaman marketing, atau aplikasi kritis SEO.
100
108
  - Menggunakan Astro 5 untuk situs statis atau konten-berat dengan island interaktif opsional.
101
109
  - Menambahkan interaktivitas ke halaman server-rendered dengan HTMX 2 tanpa menulis ulang ke SPA.
110
+
111
+ ---
112
+ ### 📄 Standard Pages Mandate (CRITICAL)
113
+ **MANDATORY**: Whenever you are building a new application, landing page, or website, you MUST automatically create the following standard pages:
114
+ 1. **About Page** (`/about`)
115
+ 2. **Profile Page** (`/profile`)
116
+ 3. **Contact Page** (`/contact`)
117
+ 4. **Terms of Reference / Terms of Service** (`/terms`)
118
+ 5. **Privacy Policy** (`/privacy-policy`)
119
+
120
+ These pages must be generated with standard boilerplate content that can later be customized to fit the specific application. Do not wait for the user to ask for them; they are a strict requirement for all web projects. / **WAJIB**: Otomatis buatkan halaman standar (About, Profile, Contact, Terms, Privacy Policy) pada setiap pembuatan aplikasi/website baru dengan konten boilerplate yang bisa disesuaikan nanti.
@@ -1,234 +1,254 @@
1
- ---
2
- name: multi-agent-orchestration
3
- description: "Expert guide for designing and orchestrating multi-agent systems, agent swarms, graph-based workflows (LangGraph, CrewAI, AutoGen), shared state memory, and human-in-the-loop guardrails in English and Indonesian."
4
- author: "Roedy Rustam"
5
- ---
6
-
7
- # Multi-Agent Orchestration Expert (2026 Edition)
8
-
9
- [English](#english) | [Bahasa Indonesia](#bahasa-indonesia)
10
-
11
- ---
12
-
13
- <a name="english"></a>
14
- ## English
15
-
16
- ### Description
17
- Expert guide for designing, building, and deploying production-grade multi-agent AI systems. Covers agent orchestration frameworks (LangGraph, OpenAI Agents SDK, Google ADK, Mastra.ai, CrewAI, AutoGen), shared state and memory management, tool execution, human-in-the-loop (HITL) guardrails, and observability for agentic workflows.
18
-
19
- **Swarm Synergy:** This skill acts as a powerful orchestrator when combined with `mcp-server-architect` (for external tool integration) and `ai-llm-integration-expert` (for foundation model setup). Together, they form a complete, end-to-end **AI Engineering Swarm**.
20
-
21
- ### Trigger Conditions
22
- - Building autonomous AI agents that execute multi-step tasks.
23
- - Designing systems where multiple specialized AI agents collaborate.
24
- - Implementing graph-based agent workflows with LangGraph or similar frameworks.
25
- - Integrating human-in-the-loop checkpoints for high-stakes decisions.
26
- - Building AI pipelines with tool-calling, RAG retrieval, code execution, or browser control.
27
- - Evaluating and selecting agent frameworks (LangGraph vs OpenAI Agents SDK vs Google ADK).
28
-
29
- ### Agent Framework Comparison (2026)
30
-
31
- | Framework | Language | Best For | Key Differentiator |
32
- |---|---|---|---|
33
- | **LangGraph** | Python / TypeScript | Complex stateful workflows | Graph-based, any LLM, full control |
34
- | **OpenAI Agents SDK** | Python | GPT-5 native agents | Built-in handoffs, tracing, guardrails |
35
- | **Google ADK** | Python | Gemini-powered agents | Multi-agent, Vertex AI, streaming |
36
- | **Mastra.ai** | TypeScript | TS-first agent apps | Built-in memory, evals, RAG, MCP |
37
- | **CrewAI** | Python | Team-of-agents tasks | Role-based agents, easy to start |
38
- | **AutoGen** | Python | Research & LLM evaluation | Conversation-driven agents |
39
-
40
- ### Core Architecture Principles
41
-
42
- #### 1. Agent Roles & Specialization
43
- Design agents with single responsibilities — avoid "do-everything" agents:
44
- - **Orchestrator Agent**: Routes tasks, decomposes goals, delegates to specialists.
45
- - **Specialist Agents**: Domain-specific (research agent, code agent, data analyst, writer).
46
- - **Tool Agents**: Wrap external capabilities (browser agent, SQL agent, file agent).
47
- - **Critic/Validator Agent**: Reviews output of other agents before finalizing.
48
-
49
- #### 2. LangGraph Stateful Graph Workflows
50
- LangGraph models agent workflows as directed graphs with persistent state — ideal for complex, multi-step tasks with branching logic and HITL:
51
- ```python
52
- from langgraph.graph import StateGraph, END
53
- from langgraph.checkpoint.memory import MemorySaver
54
- from typing import TypedDict, Annotated
55
- import operator
56
-
57
- class AgentState(TypedDict):
58
- messages: Annotated[list, operator.add]
59
- task: str
60
- result: str
61
-
62
- def research_node(state: AgentState):
63
- # Call research agent
64
- return {"messages": [research_agent.invoke(state["task"])]}
65
-
66
- def write_node(state: AgentState):
67
- # Call writing agent with research result
68
- return {"result": writing_agent.invoke(state["messages"])}
69
-
70
- def should_revise(state: AgentState) -> str:
71
- # Conditional routing
72
- return "revise" if needs_revision(state["result"]) else "end"
73
-
74
- builder = StateGraph(AgentState)
75
- builder.add_node("research", research_node)
76
- builder.add_node("write", write_node)
77
- builder.add_conditional_edges("write", should_revise, {"revise": "research", "end": END})
78
-
79
- # Persist state for HITL
80
- memory = MemorySaver()
81
- graph = builder.compile(checkpointer=memory, interrupt_before=["write"])
82
- ```
83
-
84
- #### 3. OpenAI Agents SDK — Handoffs & Guardrails
85
- Use the OpenAI Agents SDK for native GPT-5 agent workflows with built-in tracing:
86
- ```python
87
- from agents import Agent, Runner, handoff, input_guardrail, GuardrailFunctionOutput
88
-
89
- # Define specialist agents
90
- researcher = Agent(
91
- name="Researcher",
92
- instructions="Search and retrieve relevant information.",
93
- tools=[web_search, document_retrieval],
94
- )
95
-
96
- writer = Agent(
97
- name="Writer",
98
- instructions="Write high-quality content based on research.",
99
- handoffs=[handoff(researcher, tool_name_override="get_research")],
100
- )
101
-
102
- # Input guardrail to prevent harmful requests
103
- @input_guardrail
104
- async def content_filter(ctx, agent, input) -> GuardrailFunctionOutput:
105
- if contains_harmful_content(input):
106
- return GuardrailFunctionOutput(output_info="Blocked", tripwire_triggered=True)
107
- return GuardrailFunctionOutput(output_info="OK", tripwire_triggered=False)
108
-
109
- # Run with tracing
110
- result = await Runner.run(writer, "Write an article about...", guardrails=[content_filter])
111
- ```
112
-
113
- #### 4. Google ADK — Gemini Multi-Agent
114
- Google Agent Development Kit (ADK) for building Gemini-powered agents with Vertex AI integration:
115
- ```python
116
- from google.adk.agents import Agent
117
- from google.adk.tools import google_search, code_execution
118
-
119
- root_agent = Agent(
120
- model="gemini-2.5-pro",
121
- name="orchestrator",
122
- instruction="Coordinate research and analysis tasks.",
123
- sub_agents=[research_agent, analysis_agent],
124
- tools=[google_search, code_execution],
125
- )
126
- ```
127
-
128
- #### 5. Mastra.ai TypeScript-First Agents
129
- For TypeScript teams, Mastra provides the most complete agentic framework:
130
- ```typescript
131
- import { Agent, MastraMemory } from '@mastra/core';
132
- import { createTool } from '@mastra/core/tools';
133
-
134
- const webSearchTool = createTool({
135
- id: 'web-search',
136
- description: 'Search the web for current information',
137
- inputSchema: z.object({ query: z.string() }),
138
- execute: async ({ context: { query } }) => searchWeb(query),
139
- });
140
-
141
- const researchAgent = new Agent({
142
- name: 'researcher',
143
- instructions: 'Find and summarize information accurately.',
144
- model: { provider: 'ANTHROPIC', name: 'claude-sonnet-4-5' },
145
- tools: { webSearch: webSearchTool },
146
- memory: new MastraMemory({ storage: supabaseStorage }),
147
- });
148
- ```
149
-
150
- #### 6. Human-in-the-Loop (HITL) Guardrails
151
- Mandatory for high-stakes agent actions (financial transactions, email sending, code deployment):
152
- - **Interrupt Checkpoints**: Pause graph execution before irreversible actions.
153
- - **Approval Flows**: Send pending action to a UI for human review before continuing.
154
- - **Confidence Thresholds**: Auto-approve if confidence > 90%, escalate if < 70%.
155
-
156
- #### 7. Agent Memory Architecture
157
- - **Working Memory (In-context)**: Recent messages and task state in the prompt window.
158
- - **Episodic Memory**: Summarized past sessions stored as embeddings (Mem0, MemGPT).
159
- - **Semantic Memory**: Domain knowledge in a vector store (pgvector, Qdrant).
160
- - **Procedural Memory**: Learned tool-use patterns stored as structured data.
161
-
162
- #### 8. Observability & Evaluation
163
- - **LangSmith**: Native tracing for LangGraph, LangChain agents.
164
- - **OpenAI Tracing**: Built-in in OpenAI Agents SDK — view agent runs, handoffs, tool calls.
165
- - **Mastra Evals**: Built-in evaluation framework for Mastra agents.
166
- - **Custom Metrics**: Track task completion rate, tool call accuracy, latency, and cost per run.
167
-
168
- ---
169
-
170
- <a name="bahasa-indonesia"></a>
171
- ## Bahasa Indonesia
172
-
173
- ### Deskripsi
174
- Panduan ahli untuk merancang, membangun, dan men-deploy sistem multi-agen AI tingkat produksi. Mencakup framework orkestrasi agen (LangGraph, OpenAI Agents SDK, Google ADK, Mastra.ai), manajemen state dan memori bersama, eksekusi tool, guardrail human-in-the-loop (HITL), dan observabilitas untuk alur kerja agentik.
175
-
176
- **Sinergi Swarm:** Skill ini bertindak sebagai orkestrator yang sangat *powerful* jika dikombinasikan dengan `mcp-server-architect` (untuk integrasi eksternal tool) dan `ai-llm-integration-expert` (untuk penyiapan foundation model). Bersama-sama, ketiganya membentuk **AI Engineering Swarm** yang komprehensif dari ujung ke ujung.
177
-
178
- ### Kondisi Pemicu
179
- - Membangun agen AI otonom yang mengeksekusi tugas multi-langkah.
180
- - Merancang sistem di mana beberapa agen AI khusus berkolaborasi.
181
- - Mengimplementasikan alur kerja agen berbasis graph dengan LangGraph atau framework serupa.
182
- - Mengintegrasikan checkpoint human-in-the-loop untuk keputusan berisiko tinggi.
183
- - Membangun pipeline AI dengan tool-calling, RAG, eksekusi kode, atau kontrol browser.
184
- - Mengevaluasi dan memilih framework agen yang tepat.
185
-
186
- ### Perbandingan Framework Agen (2026)
187
-
188
- | Framework | Bahasa | Terbaik Untuk | Diferensiasi Kunci |
189
- |---|---|---|---|
190
- | **LangGraph** | Python / TS | Alur kerja stateful kompleks | Berbasis graph, LLM apa saja, kontrol penuh |
191
- | **OpenAI Agents SDK** | Python | Agen GPT-5 native | Handoffs, tracing, guardrails bawaan |
192
- | **Google ADK** | Python | Agen berbasis Gemini | Multi-agen, Vertex AI, streaming |
193
- | **Mastra.ai** | TypeScript | Aplikasi agen TS-first | Memori, evaluasi, RAG, MCP bawaan |
194
- | **CrewAI** | Python | Tugas tim-agen | Agen berbasis peran, mudah dimulai |
195
- | **AutoGen** | Python | Riset & evaluasi LLM | Agen berbasis percakapan |
196
-
197
- ### Prinsip Arsitektur Inti
198
-
199
- #### 1. Peran & Spesialisasi Agen
200
- Rancang agen dengan tanggung jawab tunggal:
201
- - **Orchestrator Agent**: Mendelegasikan tugas ke agen spesialis.
202
- - **Specialist Agents**: Domain-spesifik (agen riset, kode, analis data, penulis).
203
- - **Tool Agents**: Membungkus kemampuan eksternal (browser, SQL, file).
204
- - **Critic/Validator Agent**: Meninjau output agen lain sebelum difinalisasi.
205
-
206
- #### 2. LangGraph Alur Kerja Graf Stateful
207
- LangGraph memodelkan alur kerja agen sebagai graf terarah dengan state persisten — ideal untuk tugas kompleks dengan logika percabangan dan HITL. State disimpan di checkpointer (MemorySaver atau PostgreSQL) untuk resume antar sesi.
208
-
209
- #### 3. OpenAI Agents SDK Handoffs & Guardrails
210
- SDK native untuk agen GPT-5 dengan handoffs agen-ke-agen, tracing bawaan, dan guardrails untuk mencegah output berbahaya.
211
-
212
- #### 4. Google ADK — Agen Gemini Multi-Agent
213
- ADK untuk membangun agen Gemini dengan integrasi Vertex AI, sub-agents, dan tool seperti Google Search dan eksekusi kode.
214
-
215
- #### 5. Mastra.ai Agen TypeScript-First
216
- Framework paling lengkap untuk tim TypeScript: memori bawaan, evaluasi, RAG, dan dukungan MCP native.
217
-
218
- #### 6. Human-in-the-Loop (HITL) Guardrails
219
- Wajib untuk aksi agen berisiko tinggi (transaksi keuangan, pengiriman email, deployment kode):
220
- - **Interrupt Checkpoints**: Jeda eksekusi graf sebelum aksi tidak dapat dibalik.
221
- - **Approval Flows**: Kirim aksi yang menunggu ke UI untuk ditinjau manusia.
222
- - **Confidence Thresholds**: Auto-approve jika keyakinan > 90%, eskalasi jika < 70%.
223
-
224
- #### 7. Arsitektur Memori Agen
225
- - **Working Memory**: Riwayat percakapan recent dalam context window.
226
- - **Episodic Memory**: Sesi masa lalu yang diringkas sebagai embedding (Mem0).
227
- - **Semantic Memory**: Pengetahuan domain dalam vector store (pgvector, Qdrant).
228
- - **Procedural Memory**: Pola penggunaan tool yang dipelajari sebagai data terstruktur.
229
-
230
- #### 8. Observabilitas & Evaluasi
231
- - **LangSmith**: Tracing native untuk LangGraph.
232
- - **OpenAI Tracing**: Bawaan di OpenAI Agents SDK lihat run, handoff, tool call.
233
- - **Mastra Evals**: Framework evaluasi bawaan untuk agen Mastra.
234
- - **Metrik Kustom**: Lacak tingkat penyelesaian tugas, akurasi tool call, latensi, dan biaya per run.
1
+ ---
2
+ name: multi-agent-orchestration
3
+ version: "2.8.0"
4
+ description: "Expert guide for designing and orchestrating multi-agent systems, agent swarms, 2026 Anthropic agentic design patterns, graph-based workflows (LangGraph, OpenAI Agents SDK, Google ADK, Mastra.ai), shared state memory, and human-in-the-loop guardrails in English and Indonesian."
5
+ author: "Roedy Rustam"
6
+ ---
7
+
8
+ # Multi-Agent Orchestration Expert (2026 Edition)
9
+
10
+ [English](#english) | [Bahasa Indonesia](#bahasa-indonesia)
11
+
12
+ ---
13
+
14
+ <a name="english"></a>
15
+ ## English
16
+
17
+ ### Orchestration & Integration
18
+ Connects and orchestrates with relevant domain skills like `brainstorming`, `zero-to-prod-orchestrator`, `ai-llm-integration-expert`, `mcp-server-architect`, and `project-context-mapper` to ensure cohesive execution.
19
+
20
+ ### Description
21
+ Expert guide for designing, building, and deploying production-grade multi-agent AI systems. Covers core agentic design patterns (Prompt Chaining, Routing, Parallelization, Orchestrator-Workers, Evaluator-Optimizer), stateful graph engines (LangGraph, OpenAI Agents SDK, Google ADK, Mastra.ai), shared episodic/semantic memory, tool execution sandboxes, and human-in-the-loop (HITL) guardrails.
22
+
23
+ **Swarm Synergy:** This skill acts as a master orchestrator when combined with `mcp-server-architect` (for external tool integration) and `ai-llm-integration-expert` (for foundation model setup). Together, they form a complete, end-to-end **AI Engineering Swarm**.
24
+
25
+ ### Trigger Conditions
26
+ - Building autonomous AI agents that execute complex, multi-step tasks across several domains.
27
+ - Designing systems where multiple specialized AI agents collaborate, deliberate, and cross-validate.
28
+ - Implementing stateful, graph-based agent workflows with LangGraph, OpenAI Agents SDK, or Google ADK.
29
+ - Implementing Anthropic agentic design patterns: Evaluator-Optimizer loops, Orchestrator-Workers, or Routing.
30
+ - Integrating human-in-the-loop (HITL) pause checkpoints for high-risk actions (code execution, database migrations, financial transactions).
31
+ - Evaluating and selecting agent architectures across Python, TypeScript, and multi-platform swarms.
32
+
33
+ ### Anthropic 2026 Core Agentic Design Patterns
34
+
35
+ Production systems should favor explicit **Workflows** over unbounded autonomous loops where predictability and reliability are required:
36
+
37
+ ```
38
+ 1. PROMPT CHAINING
39
+ [Input] ---> [LLM Step 1] ---> [Gate/Validator] ---> [LLM Step 2] ---> [Output]
40
+
41
+ 2. ROUTING
42
+ [Input] ---> [Classifier/Router] ──┬──> [Specialist Agent A]
43
+ ├──> [Specialist Agent B]
44
+ └──> [Specialist Agent C]
45
+
46
+ 3. PARALLELIZATION (Sectioning & Voting)
47
+ [Input] ──┬──> [Task 1 (Subagent)] ──┐
48
+ ├──> [Task 2 (Subagent)] ──┼──> [Aggregator / Synthesizer]
49
+ └──> [Task 3 (Subagent)] ──┘
50
+
51
+ 4. ORCHESTRATOR-WORKERS (Dynamic Decomposition)
52
+ [Input] ---> [Orchestrator] ──┬──> [Worker 1 (Focused Context)] ──┐
53
+ ├──> [Worker 2 (Focused Context)] ──┼──> [Orchestrator Synthesis]
54
+ └──> [Worker 3 (Focused Context)] ──┘
55
+
56
+ 5. EVALUATOR-OPTIMIZER LOOP (Zero-Tolerance Quality Gate)
57
+ [Input] ---> [Generator Agent] <─────┐ (Feedback Loop)
58
+ │ │
59
+ ▼ │
60
+ [Evaluator / Auditor] ────┘ (Reject / Needs Revision)
61
+
62
+ (Approved)
63
+ [Output]
64
+ ```
65
+
66
+ ### Agent Framework Comparison (2026)
67
+
68
+ | Framework | Language | Best For | Key Differentiator |
69
+ |---|---|---|---|
70
+ | **LangGraph (v0.3+)** | Python / TypeScript | Complex stateful workflows & graphs | Graph-based, persistent checkpointers, time-travel debugging |
71
+ | **OpenAI Agents SDK** | Python | GPT-5 / o-series native agents | Built-in agent handoffs, tracing, and tripwire guardrails |
72
+ | **Google ADK** | Python | Gemini-powered swarms | Native Vertex AI, multi-agent streaming, search grounding |
73
+ | **Mastra.ai** | TypeScript | TS-first web apps & microservices | Built-in memory, evals, RAG, and native MCP support |
74
+ | **CrewAI** | Python | Role-playing business teams | Fast initial prototyping for business analyst teams |
75
+
76
+ ### Core Implementation Guidelines
77
+
78
+ #### 1. LangGraph — Persistent State & HITL Checkpoints
79
+ LangGraph models agent workflows as directed acyclic or cyclic graphs with persistent state:
80
+ ```python
81
+ from langgraph.graph import StateGraph, END
82
+ from langgraph.checkpoint.memory import MemorySaver
83
+ from typing import TypedDict, Annotated
84
+ import operator
85
+
86
+ class AgentState(TypedDict):
87
+ messages: Annotated[list, operator.add]
88
+ task: str
89
+ code_artifact: str
90
+ audit_feedback: str
91
+ approved: bool
92
+
93
+ def generator_node(state: AgentState):
94
+ # Generates or refactors code based on previous feedback
95
+ code = coder_agent.invoke(state["task"], feedback=state.get("audit_feedback"))
96
+ return {"code_artifact": code}
97
+
98
+ def evaluator_node(state: AgentState):
99
+ # Runs automated linter/tests & security review
100
+ audit = auditor_agent.invoke(state["code_artifact"])
101
+ return {
102
+ "audit_feedback": audit.critique,
103
+ "approved": audit.is_passing
104
+ }
105
+
106
+ def route_next(state: AgentState) -> str:
107
+ return END if state["approved"] else "generator"
108
+
109
+ builder = StateGraph(AgentState)
110
+ builder.add_node("generator", generator_node)
111
+ builder.add_node("evaluator", evaluator_node)
112
+ builder.set_entry_point("generator")
113
+ builder.add_edge("generator", "evaluator")
114
+ builder.add_conditional_edges("evaluator", route_next, {"generator": "generator", END: END})
115
+
116
+ # Persist state with checkpointer for HITL interruption before destructive actions
117
+ checkpointer = MemorySaver()
118
+ graph = builder.compile(checkpointer=checkpointer, interrupt_before=["generator"])
119
+ ```
120
+
121
+ #### 2. OpenAI Agents SDK — Agent Handoffs & Guardrails
122
+ Implement native agent handoffs where specialized agents transition control cleanly:
123
+ ```python
124
+ from agents import Agent, Runner, handoff, input_guardrail, GuardrailFunctionOutput
125
+
126
+ researcher = Agent(
127
+ name="Researcher",
128
+ instructions="Research libraries, security advisories, and system specs.",
129
+ tools=[web_search, doc_retrieval],
130
+ )
131
+
132
+ architect = Agent(
133
+ name="Architect",
134
+ instructions="Synthesize technical architecture and delegate research when needed.",
135
+ handoffs=[handoff(researcher, tool_name_override="delegate_research")],
136
+ )
137
+
138
+ @input_guardrail
139
+ async def safety_guardrail(ctx, agent, input_data) -> GuardrailFunctionOutput:
140
+ if contains_destructive_commands(input_data):
141
+ return GuardrailFunctionOutput(output_info="Blocked destructive payload", tripwire_triggered=True)
142
+ return GuardrailFunctionOutput(output_info="Safe", tripwire_triggered=False)
143
+
144
+ result = await Runner.run(architect, "Design high-throughput ingestion pipeline", guardrails=[safety_guardrail])
145
+ ```
146
+
147
+ #### 3. Google ADK — Gemini Multi-Agent Systems
148
+ Orchestrate Gemini 3.x agents with streaming subagent calls and Vertex AI tooling:
149
+ ```python
150
+ from google.adk.agents import Agent
151
+ from google.adk.tools import google_search, code_execution
152
+
153
+ director = Agent(
154
+ model="gemini-3.1-pro",
155
+ name="director",
156
+ instruction="Coordinate domain specialists and synthesize final deliverables.",
157
+ sub_agents=[frontend_agent, backend_agent, security_agent],
158
+ tools=[google_search, code_execution],
159
+ )
160
+ ```
161
+
162
+ #### 4. Mastra.ai TypeScript-Native Agents
163
+ For modern Next.js / Node.js / Bun environments:
164
+ ```typescript
165
+ import { Agent, MastraMemory } from '@mastra/core';
166
+ import { createTool } from '@mastra/core/tools';
167
+ import { z } from 'zod';
168
+
169
+ const researcher = new Agent({
170
+ name: 'researcher',
171
+ instructions: 'Find and summarize accurate technical documentation.',
172
+ model: { provider: 'ANTHROPIC', name: 'claude-3-7-sonnet-20250219' },
173
+ memory: new MastraMemory({ storage: supabaseStorage }),
174
+ });
175
+ ```
176
+
177
+ #### 5. Human-in-the-Loop (HITL) Guardrails
178
+ Mandatory safeguards before executing irreversible operations:
179
+ - **Interrupt Checkpoints**: Halt workflow execution before executing code, migrating databases, or modifying production records.
180
+ - **Approval Dashboards**: Surface diff previews and proposed shell commands to the user or admin before proceeding.
181
+ - **Confidence Gates**: Auto-proceed only when model confidence score is >= 0.90; trigger human escalation otherwise.
182
+
183
+ #### 6. Swarm Circuit Breakers & Fallback Protocols
184
+ - **Retry Caps**: Maximum 2 automated retries per subagent.
185
+ - **Fallback Escalation**: If a specialist agent stalls or loops, the Swarm Director gracefully fallbacks to `fullstack-expert` or requests human guidance.
186
+ - **Checkpoint Persistence**: Always persist intermediate progress to `PROGRESS.md` or `BLUEPRINT.md` so sessions can resume without losing context.
187
+
188
+ ---
189
+
190
+ <a name="bahasa-indonesia"></a>
191
+ ## Bahasa Indonesia
192
+
193
+ ### Integrasi Orkestrasi
194
+ Terhubung dan mengorkestrasi skill domain yang relevan seperti `brainstorming`, `zero-to-prod-orchestrator`, `ai-llm-integration-expert`, `mcp-server-architect`, dan `project-context-mapper` untuk memastikan eksekusi yang kohesif.
195
+
196
+ ### Deskripsi
197
+ Panduan ahli untuk merancang, membangun, dan men-deploy sistem multi-agen AI tingkat produksi. Mencakup pola desain agentik inti (Prompt Chaining, Routing, Parallelization, Orchestrator-Workers, Evaluator-Optimizer), engine graph stateful (LangGraph, OpenAI Agents SDK, Google ADK, Mastra.ai), memori bersama episodik/semantik, sandbox eksekusi tool, dan guardrail human-in-the-loop (HITL).
198
+
199
+ **Sinergi Swarm:** Skill ini bertindak sebagai orkestrator utama jika dipadukan dengan `mcp-server-architect` (untuk integrasi tool eksternal) dan `ai-llm-integration-expert` (untuk konfigurasi foundation model). Bersama-sama, ketiganya membentuk **AI Engineering Swarm** yang tangguh dari awal hingga rilis produksi.
200
+
201
+ ### Kondisi Pemicu
202
+ - Membangun agen AI otonom yang mengeksekusi tugas kompleks multi-langkah lintas domain.
203
+ - Merancang sistem kolaborasi, deliberasi, dan validasi silang antar beberapa agen AI spesialis.
204
+ - Mengimplementasikan alur kerja graph stateful dengan LangGraph, OpenAI Agents SDK, atau Google ADK.
205
+ - Menerapkan 5 pola desain agentik standar: Prompt Chaining, Routing, Parallelization, Orchestrator-Workers, atau Evaluator-Optimizer.
206
+ - Mengintegrasikan pos henti human-in-the-loop (HITL) untuk tindakan berisiko tinggi (eksekusi kode, migrasi database, transaksi keuangan).
207
+ - Memilih dan mengevaluasi arsitektur agen di ekosistem Python, TypeScript, atau multi-platform.
208
+
209
+ ### 5 Pola Desain Agentik Inti (Standar Anthropic 2026)
210
+
211
+ Untuk sistem produksi yang handal, utamakan arsitektur **Workflows** terstruktur daripada loop otonom tanpa batas:
212
+
213
+ 1. **Prompt Chaining**: Memecah tugas menjadi langkah-langkah sekuensial dengan validasi output di setiap transisi.
214
+ 2. **Routing**: Mengklasifikasikan input pengguna dan mengarahkannya ke model atau sub-agen yang memiliki spesialisasi yang tepat.
215
+ 3. **Parallelization (Sectioning & Voting)**: Menjalankan beberapa sub-agen secara simultan untuk tugas independen atau menjalankan ensemble untuk konsensus voting.
216
+ 4. **Orchestrator-Workers**: Agen orkestrator pusat memecah masalah dinamis, mendelegasikannya ke pekerja dengan konteks terfokus, lalu merangkum hasil akhirnya.
217
+ 5. **Evaluator-Optimizer Loop**: Agen pembuat (*generator*) menghasilkan solusi sementara agen penilai (*evaluator*) memberikan audit dan umpan balik hingga standar kualitas terpenuhi.
218
+
219
+ ### Perbandingan Framework Agen (2026)
220
+
221
+ | Framework | Bahasa | Terbaik Untuk | Keunggulan Utama |
222
+ |---|---|---|---|
223
+ | **LangGraph (v0.3+)** | Python / TypeScript | Alur kerja graf stateful kompleks | Berbasis graf, checkpointer persisten, time-travel debugging |
224
+ | **OpenAI Agents SDK** | Python | Agen native GPT-5 / o-series | Handoff antar agen bawaan, tracing, dan guardrail otomatis |
225
+ | **Google ADK** | Python | Swarm agen bertenaga Gemini | Integrasi Vertex AI native, streaming multi-agen, search grounding |
226
+ | **Mastra.ai** | TypeScript | Web apps & microservice TS-first | Memori bawaan, evaluasi otomatis, RAG, dan dukungan MCP native |
227
+ | **CrewAI** | Python | Tim simulasi peran | Cepat untuk membuat prototipe kolaborasi tim bisnis |
228
+
229
+ ### Panduan Implementasi Inti
230
+
231
+ #### 1. LangGraph State Persisten & Checkpoint HITL
232
+ Memodelkan alur agen sebagai graf terarah dengan state bersama dan penyimpanan checkpoint:
233
+ - Simpan state di database (PostgreSQL / MemorySaver) agar alur kerja dapat dijeda dan dilanjutkan kapan saja.
234
+ - Terapkan `interrupt_before` sebelum node yang menjalankan perintah destruktif untuk meminta persetujuan manusia (*Human-in-the-loop*).
235
+
236
+ #### 2. OpenAI Agents SDK — Handoffs & Guardrails
237
+ Terapkan transisi kendali yang mulus antar agen dengan fungsi `handoff` bawaan serta pasang filter `guardrail` pada input dan output untuk mencegah eksekusi instruksi berbahaya.
238
+
239
+ #### 3. Google ADK — Multi-Agent Gemini
240
+ Bangun hierarki agen dengan model Gemini 3.x, di mana root agent mengoordinasikan sub-agents untuk riset, eksekusi kode, dan pembuatan dokumen.
241
+
242
+ #### 4. Mastra.ai — Solusi TypeScript Penuh
243
+ Gunakan Mastra untuk ekosistem Next.js dan Node.js: sediakan memori persisten ke Supabase/PostgreSQL, integrasikan tool MCP secara langsung, dan manfaatkan framework evaluasi bawaan.
244
+
245
+ #### 5. Guardrails Human-in-the-Loop (HITL)
246
+ Pengamanan wajib sebelum melakukan tindakan yang tidak dapat dibatalkan:
247
+ - **Pos Henti Interupsi**: Hentikan eksekusi sebelum menjalankan skrip shell berbahaya, migrasi skema tabel, atau memodifikasi data produksi.
248
+ - **Tinjauan Pratinjau**: Tampilkan ringkasan perbedaan (*diff*) kepada pengguna sebelum modifikasi dieksekusi.
249
+ - **Ambang Keyakinan**: Otomatis lanjutkan hanya jika skor keyakinan model >= 0.90; eskalasikan ke manusia jika berada di bawah ambang batas.
250
+
251
+ #### 6. Circuit Breakers & Protokol Pemulihan Swarm
252
+ - **Batas Percobaan Ulang**: Maksimal 2 kali perbaikan otomatis per sub-agen.
253
+ - **Eskalasi Fallback**: Jika agen spesialis mengalami kendala konteks atau gagal berulang kali, Swarm Director segera mengalihkan tugas ke `fullstack-expert` atau meminta masukan pengguna.
254
+ - **Persistensi Kemajuan**: Simpan selalu checkpoint di `PROGRESS.md` atau `BLUEPRINT.md` agar alur kerja dapat dilanjutkan secara efisien tanpa token berlebih.