复制安装命令
用 Codex 或 Claude 安装复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它先审查 Skill 页面再帮你安装。
复制前请先查看来源、License 和安全提示。
Essays and writing behind this toolkit live at vexjoy.com.
用 Codex 或 Claude 安装复制这段 Prompt,粘贴到 Codex、Claude 或其他助手里,让它先审查 Skill 页面再帮你安装。
复制前请先查看来源、License 和安全提示。
来源文件:README.md
Essays and writing behind this toolkit live at vexjoy.com.
AI agents skip steps.
"Looks correct" replaces running tests. "Trivial change" replaces verification. The agent confidently ships broken code because nothing structurally prevented it from skipping the work.
Harnesses have a second problem: given only a skill list, they do not route eagerly enough, or correctly enough. Good skills sit unused. So this toolkit connects the skills, agents, and workflows we want directly into the harness, automatically. You don't have to understand what is here. Say what you want in plain English and you get all the value we have put into it: the right specialist with the right methodology, behind gates that demand exit codes, not assertions.
44 domain agents, 122 workflow skills, 78 hooks, 136 scripts. Agents carry knowledge, skills enforce methodology, hooks block incomplete work, scripts handle determinism.
Works across Claude Code (/do), Codex ($do), Factory (/do), Reasonix (/do).
$ claude
> /do debug this Go test
Routing: go-engineer + systematic-debugging
Phase 1/4: Reproduce: running test, capturing failure...
Phase 2/4: Hypothesize: 3 candidates from stack trace...
Phase 3/4: Verify: isolated root cause in connection pool timeout
Phase 4/4: Fix: patch applied, test passing, PR opened
✓ Delivered: PR #847, fix connection pool timeout in health check
The router reads intent, picks a Go agent paired with a debugging skill, and runs the full lifecycle. You typed one sentence. The system did the rest.
ROUTE PLAN EXECUTE VERIFY DELIVER RECORD
┌──────┐ ┌──────┐ ┌──────┐ ┌──────┐ ┌──────┐ ┌──────┐
│ /do │───▶│ Task │───▶│Agent │───▶│Tests │───▶│ PR │───▶│Route │
│Router│ │ Plan │ │+Skill│ │Gates │ │Branch│ │Result│
└──────┘ └──────┘ └──────┘ └──────┘ └──────┘ └──────┘
This is the single thing that separates it from "agent with a system prompt."
| Agent Says | What Happens |
|---|---|
| "Code looks correct, skip tests" | Exit gate requires test output. Blocked. |
| "Trivial change, no verification" | Hook blocks completion without evidence. |
| "Similar to before" | Skill demands case-specific proof. |
| "User is in a hurry" | Protocol overrides time pressure. |
| "I'm confident" | Gate demands exit code, not assertion. |
Hooks fire automatically. Gates block completion. Skills encode counter-arguments at every skip-worthy step. The agent verifies or it doesn't finish.
For what I do, the difference is enormous. If you're doing simple single-file edits, maybe less so.
The same routing serves knowledge work. The content engine researches, drafts in a calibrated voice, validates against 397 AI patterns, and repurposes finished pieces for each platform. /html turns any request into a single self-contained HTML file: report, slide deck, prototype, data viz, diagram. Non-engineers who try the toolkit consistently name the HTML artifacts as the thing they love. No code, no setup beyond the installer.
Changes to the toolkit itself ship with evidence. New skills get blind A/B tests against a no-skill baseline before merge. Routing and writing-standard decisions carry measured verdicts; PHILOSOPHY.md cites the numbers. Experiments that lost go into the negative-results registry, what-didnt-work.md; the registry now covers routing reversals, unvalidated A/B citations, and disabled lint rules alongside the original program refutations.
The automated nightly evolution loop (/evolve, writes to evolution-reports/) ran regularly through mid-May 2026. It is currently dormant; recent evidence has come from manual PRs instead.
git clone https://github.com/notque/vexjoy-agent.git ~/vexjoy-agent
cd ~/vexjoy-agent
./install.sh
Links into ~/.claude/ and mirrors into ~/.codex/, ~/.factory/, ~/.reasonix/ — each mirror only when that runtime is detected (its command on PATH or its home dir already exists). The installer asks symlink (live updates via git pull) or copy (stable snapshot).
Want only part of the toolkit? Run ./install.sh --configure to pick which skills, agents, and hooks install, or copy .local.example/profile.yaml to .local/profile.yaml and edit. No profile file = full install, unchanged behavior. Credit: @thomasvan. Details: .local.example/README.md.
| CLI | Entry Point |
|---|---|
| Claude Code | /do |
| Codex | $do |
| Factory | /do |
| Reasonix | /do |
Full setup: docs/start-here.md
Mirrors agents, skills, and supported hooks into ~/.codex/. The original six-hook allowlist was correct for Codex v0.114, when tool hooks only intercepted Bash. Current support requires Codex v0.144.1+ and classifies the 74 Claude hook registrations as 26 native, 35 adapter-backed, and 13 unsupported (61 supported). These are registration counts, not unique hook files. The installer also preserves explicit per-subagent model routing for GPT-5.6 Sol by setting the MultiAgent V2 compatibility keys documented in openai/codex#31814.
Codex now exposes apply_patch to tool hooks. VexJoy's adapter converts each patch operation into the Write/Edit payload expected by existing guards, but it cannot intercept writes performed through unified_exec, unmatched MCP tools, WebSearch, or other unsupported tool paths. PreCompact and Stop adapters also receive less telemetry than Claude Code: Codex does not provide Claude's conversation_history or session_data. This is expanded compatibility, not full Claude parity.
After install or any hook-definition change, run /hooks in Codex and review the new definitions before trusting them. Codex hash-trusts hook commands and skips changed, unreviewed definitions.
Gemini CLI support removed (deprecated upstream, transitioned to Antigravity CLI); Antigravity support pending CLI maturity. Per Google's transition announcement, Gemini CLI stops serving requests on 2026-06-18 for Google AI Pro / Ultra and free Gemini Code Assist for individuals. Gemini API integrations (image-gen backends, sprite pipeline, GEMINI_API_KEY) are unaffected and stay in the toolkit.
If a prior install mirrored into ~/.gemini/, remove the stale mirrors with:
rm -rf ~/.gemini/skills ~/.gemini/agents ~/.gemini/hooks ~/.gemini/scripts ~/.gemini/antigravity/plugins/vexjoy-agent
Mirrors agents (as "droids"), skills, and all hooks into ~/.factory/. Hook config merges into ~/.factory/settings.json with paths rewritten.
Mirrors skills, scripts, and the allowlisted hooks (scripts/reasonix-hooks-allowlist.txt) into ~/.reasonix/ (no agent or custom-command surface, so neither is installed; the /do router rides in as a skill). Reasonix fires only 4 events (PreToolUse, PostToolUse, UserPromptSubmit, Stop), so only hooks for those events are allowlisted. Hook config is written to the hooks key of ~/.reasonix/settings.json in Reasonix's native flat shape (one entry per hook, match regex over the tool name); the generator builds absolute python3 commands, so no path rewrite is applied. MCP/model/permissions in ~/.reasonix/config.json are user-owned and left untouched.
The toolkit supplies its own routing, domain knowledge, methodology, and enforcement. The default system prompt duplicates most of that.
claude --system-prompt "."
Strips built-in tool-use instructions. The toolkit's agents, skills, hooks, and CLAUDE.md provide equivalent coverage.
| Layer | Count | Does |
|---|---|---|
| Agents | 44 | Domain knowledge: idiom tables, failure mode catalogs, error-to-fix mappings |
| Skills | 122 | Phased methodology with gates. Can't skip steps. Each phase has exit criteria requiring evidence. |
| Hooks | 78 | Fire on lifecycle events. Block incomplete work. Zero LLM cost. |
| Scripts | 136 | Determinism: test runners, linters, validators. No LLM judgment. |
Full skill catalog: docs/skills.md.
┌─────────────────────────────────────────────────┐
│ SKILL.md │
│ ┌─ Frontmatter ─────────────────────────────┐ │
│ │ triggers, pairs_with, success-criteria │ │
│ └────────────────────────────────────────────┘ │
│ Reference Loading Table (conditional imports) │
│ Phased Instructions (numbered, with gates) │
│ Verification (evidence requirements) │
└─────────────────────────────────────────────────┘
A game built entirely by Claude Code using these agents, skills, and pipelines:
I just want to use it Install, learn /do, done.
I do knowledge work Writing, research, data analysis, moderation, HTML artifacts. No code.
I'm a developer Architecture, extension points, adding agents and skills.
I'm an AI power user Routing tables, pipelines, hooks, telemetry DB.
I'm an AI agent Machine-dense inventory. Tables, paths, schemas.
I'm on LinkedIn 🚀 Thought leadership. Agree? 👇
Full design philosophy: PHILOSOPHY.md
One report-only script surfaces upkeep work; it prints a digest and never edits, deletes, or blocks.
python3 scripts/stale-skill-scan.py --top 20 ranks stale skills and agents as pruning candidates. Run it quarterly; see docs/deprecation-template.md.Scheduled work follows the same boundary as everything else: judgment uses agents; repeatable plumbing uses scripts.
| Need | Use |
|---|---|
| Run a deterministic command on a schedule | scripts/agent-scheduler.py with runner: "command" |
| Run an agent judgment on a schedule, webhook, or file change | scripts/agent-scheduler.py with the default runner: "claude" |
| Install or remove a user crontab entry safely | scripts/crontab-manager.py |
| Audit shell cron reliability | cron-automation |
| Keep one interactive objective moving until criteria verify | objective-loop |
See CONTRIBUTING.md.
MIT. See LICENSE.
name: business-ops
description: "Business operations: strategy, technology, growth, competitive intelligence, support, finance, HR, legal, operations, sales, productivity, product management."
user-invocable: false
allowed-tools:
- Read
- Write
- Bash
- Grep
- Glob
- Edit
routing:
triggers:
# Strategy/CEO
- "should we"
- "evaluate opportunity"
- "trade-off"
- "worth it"
- "invest in"
- "strategy"
# Technology/CTO
- "build vs buy"
- "vendor evaluation"
- "adopt"
- "technology choice"
- "tech stack"
# Growth/CMO
- "grow audience"
- "growth"
- "brand"
- "positioning"
- "community building"
# Competitive
- "competitive analysis"
- "market landscape"
- "differentiation"
# Evaluation
- "feasibility"
- "effort estimate"
- "ROI"
- "priority"
- "project evaluation"
- "go no go"
- "viability"
# Customer Support
- "customer support"
- "ticket triage"
- "support response"
- "knowledge base"
- "KB article"
- "escalation"
- "customer research"
# Finance
- "finance"
- "journal entry"
- "reconciliation"
- "variance analysis"
- "financial statements"
- "financial audit"
- "month-end close"
- "SOX"
# HR
- "HR"
- "human resources"
- "recruiting"
- "performance review"
- "compensation"
- "hiring"
- "onboarding"
- "org planning"
# Legal
- "legal"
- "contract review"
- "compliance check"
- "NDA"
- "legal risk"
- "legal brief"
- "vendor check"
- "german compliance"
- "DSGVO"
- "GoBD"
- "TDDDG"
- "AI Act compliance"
- "eIDAS"
# Operations
- "operations"
- "vendor review"
- "runbook"
- "process documentation"
- "risk assessment"
- "capacity plan"
- "change management"
- "compliance tracking"
# Sales
- "sales"
- "call prep"
- "pipeline review"
- "forecast"
- "draft outreach"
- "prospect research"
- "competitive intelligence"
# Productivity
- "productivity"
- "task management"
- "daily plan"
- "weekly review"
- "meeting agenda"
- "focus time"
- "goal setting"
- "status update"
- "time management"
- "prioritize tasks"
- "standup"
- "retrospective"
# Product Management
- "product management"
- "feature spec"
- "PRD"
- "roadmap"
- "stakeholder update"
- "user research"
- "sprint planning"
- "product metrics"
not_for: "micro library choices (use decision-helper), writing content, SEO of specific posts, or tactical marketing competitive analysis (use marketing) — this is executive strategy, not campaign execution. Code security audits, vulnerability scanning, or auth-flow reviews (use security-review) — only financial/accounting audit and SOX compliance. Code performance review (use reviewer-code) — this covers people performance reviews and HR operations. Software task specs, requirements, or plan-lifecycle management (use planning) — this skill prioritizes and tracks work, not specs. UX design methodology, wireframes, or accessibility audits (use design) — this handles product strategy, roadmaps, user research for feature prioritization."
complexity: Medium
category: decision-support
pairs_with:
- marketing
- data-analysisUmbrella skill for all business functions: executive strategy (CEO/CTO/CMO), competitive intelligence, project evaluation, customer support, finance, HR, legal, operations, sales, productivity, and product management. Each domain loads its own reference files on demand — this skill detects the mode, loads the right references, and executes the appropriate framework.
Scope: Business decisions and operational workflows. Use decision-helper for technical architecture micro-choices, domain agents for code, voice-writer for content, and systematic-debugging for debugging.
Classify the user's request into exactly one mode before proceeding. If the request spans multiple modes, choose the primary one and note the secondary.
| Mode | Signal Phrases | Reference |
|---|---|---|
| STRATEGY | Market entry, partnerships, resource allocation, opportunity, "should I/we", strategic pivots, investment | references/csuite.md |
| TECHNOLOGY | Build vs buy, vendor, SaaS, tech stack, architecture, adopt, technology choice | references/csuite.md |
| GROWTH | Content strategy, audience, SEO, marketing, brand, community, positioning, channel | references/csuite.md |
| COMPETITIVE | Competitor, competition, market landscape, differentiation, positioning against, market share | references/csuite.md |
| EVALUATION | Feasibility, effort estimate, ROI, priority, go/no-go, viability, "is it worth it" | references/csuite.md |
| SUPPORT | Customer support, ticket triage, support response, knowledge base, KB article, escalation | references/customer-support.md |
| FINANCE | Finance, journal entry, reconciliation, variance analysis, financial statements, audit, SOX, month-end close | references/finance.md |
| HR | HR, recruiting, performance review, compensation, hiring, onboarding, org planning | references/hr.md |
| LEGAL | Legal, contract review, compliance check, NDA, legal risk, legal brief, vendor check, DSGVO, GoBD | references/legal.md |
| OPERATIONS | Operations, vendor review, runbook, process documentation, risk assessment, capacity plan, change management | references/operations.md |
| SALES | Sales, call prep, pipeline review, forecast, draft outreach, prospect research, competitive intelligence | references/sales.md |
| PRODUCTIVITY | Productivity, task management, daily plan, weekly review, meeting agenda, focus time, goal setting, standup | references/productivity.md |
| PRODUCT | Product management, feature spec, PRD, roadmap, stakeholder update, user research, sprint planning, metrics | references/product-management.md |
Load references based on the detected mode. Load only the references required by the mode.
| Signal | Mode | Reference |
|---|---|---|
| Market entry, partnerships, resource allocation, opportunity | STRATEGY | references/strategic-frameworks.md, references/decision-matrices.md |
| Build vs buy, vendor, SaaS, tech stack, architecture | TECHNOLOGY | references/tco-framework.md, references/vendor-evaluation.md |
| Content, audience, SEO, marketing, brand, community | GROWTH | references/audience-segmentation.md, references/channel-evaluation.md |
| Competitor, market landscape, positioning, differentiation | COMPETITIVE | references/competitive-mapping.md, references/market-positioning.md |
| Feasibility, effort, ROI, priority, go/no-go | EVALUATION | references/feasibility-scoring.md, references/roi-frameworks.md |
| Ticket triage, support response, KB article, escalation, customer research | SUPPORT | references/customer-support.md |
| Journal entry, reconciliation, variance, financial statements, audit, SOX | FINANCE | references/finance.md |
| Recruiting, performance review, compensation, hiring, onboarding, org planning | HR | references/hr.md |
| Contract review, compliance check, NDA, legal risk, legal brief, DSGVO, GoBD | LEGAL | references/legal.md |
| Vendor review, runbook, process documentation, risk assessment, capacity plan, change management | OPERATIONS | references/operations.md |
| Call prep, pipeline review, forecast, draft outreach, prospect research | SALES | references/sales.md |
| Task management, daily plan, weekly review, meeting agenda, goal setting, standup | PRODUCTIVITY | references/productivity.md |
| Feature spec, PRD, roadmap, stakeholder update, user research, sprint planning, metrics | PRODUCT | references/product-management.md |
For each mode, load the corresponding reference file for the full framework and instructions:
references/csuite.mdreferences/customer-support.mdreferences/finance.mdreferences/hr.mdreferences/legal.mdreferences/operations.mdreferences/sales.mdreferences/productivity.mdreferences/product-management.md| Error | Cause | Solution |
|---|---|---|
| Too many options | 5+ options creating paralysis | Eliminate obviously inferior options first. Get to 2-4 before running full framework. |
| Not enough information | User cannot answer framing questions | Identify 2-3 critical unknowns. Recommend time-boxed research sprint before deciding. |
| Analysis paralysis | Keeps adding criteria or second-guessing | Apply reversibility test. If reversible, recommend best current option with checkpoint. |
| Emotional attachment | User has already decided, wants validation | Name the pattern directly. Ask: stress-test the choice, or genuinely evaluate all options? |
| Reference | When to Load | Content |
|---|---|---|
references/csuite.md | Any executive strategy mode | Full C-suite decision support frameworks: STRATEGY, TECHNOLOGY, GROWTH, COMPETITIVE, EVALUATION |
references/strategic-frameworks.md | STRATEGY mode | Porter's Five Forces, SWOT scoring, OKR alignment matrices |
references/decision-matrices.md | STRATEGY mode | Weighted decision matrices, ICE/RICE scoring, pre-mortem templates |
references/tco-framework.md | TECHNOLOGY mode | TCO templates, hidden cost checklists, migration cost models |
references/vendor-evaluation.md | TECHNOLOGY mode | Vendor scorecards, RFP criteria, red flag detection, contract checklist |
references/audience-segmentation.md | GROWTH mode | ICP scoring matrix, persona templates, segmentation frameworks |
references/channel-evaluation.md | GROWTH mode | Channel scoring matrices, CAC/LTV models, funnel stage mapping |
references/competitive-mapping.md | COMPETITIVE mode | Landscape map templates, feature matrices, activity tracker |
references/market-positioning.md | COMPETITIVE mode | Positioning maps, differentiation scoring, win/loss frameworks |
references/feasibility-scoring.md | EVALUATION mode | Three-dimension feasibility model, confidence calibration, decision tree |
references/roi-frameworks.md | EVALUATION mode | T-shirt sizing, three-point estimation, risk-adjusted NPV |
references/customer-support.md | SUPPORT mode | Triage, response drafting, KB articles, escalation, customer research |
references/finance.md | FINANCE mode | Journal entries, reconciliation, variance analysis, financial statements, audit/SOX |
references/hr.md | HR mode | Recruiting, performance management, compensation, org planning, people analytics |
references/legal.md | LEGAL mode | Contract review, compliance, NDA triage, risk assessment, legal writing |
references/operations.md | OPERATIONS mode | Runbooks, risk assessment, vendor management, process docs, change management, compliance |
references/sales.md | SALES mode | Call prep, pipeline analysis, outreach, competitive intelligence, forecasting |
references/productivity.md | PRODUCTIVITY mode | Task management, daily/weekly planning, meeting optimization, status updates, goals |
references/product-management.md | PRODUCT mode | Feature specs, roadmaps, stakeholder updates, research synthesis, metrics, sprint planning |
评论 (0)
暂无评论,成为第一个评论者吧!