mirror of
https://github.com/boshu2/agentops.git
synced 2026-09-14 15:08:13 +08:00
e6a1b0f54e
When a request includes coding and a retrospective, the retrospective can accidentally become a prerequisite for the code judgment it needs to analyze. Clarify the existing Plan, Validate, Postmortem and optional goal guidance so code acceptance, delivery facts and requested analysis have distinct consumers. Overall completion still requires every requested deliverable; explicitly requested interim analysis states its cutoff and pending checks. Tighten the existing coordination guidance: pass focused context, assign integration and review ownership, act on a known blocking CI failure before the entire run ends, and use native waits for unchanged pending state. Skill Eval now explicitly distinguishes following an instruction from proving benefit, and fresh context from small input. Preserve BDD/DDD, required independent judgment, caller authority and protected evidence storage. AGENTS.md stays within its existing 250-line limit; no skills, controllers or gates are added. Validation: exact-head CI passed, all 28 selected gates passed, and all 50 root/product-boundary tests passed. The local aggregate passed with its existing absent legacy-integration skip. Fresh independent review returned PASS over all 37 changed paths with no findings or unchecked candidate criteria, including current generated parity. Earlier trigger, description-budget and terminology failures were repaired without changing tests or limits. Static checks and document review establish contract consistency, not measured model uplift or token savings.
13 KiB
13 KiB
Skill Router
34 live skills. Choose guidance for a concrete task need; no skill is mandatory. A clear task can proceed in the native agent. Read a skill only when its description fits. Names and descriptions below come from each source SKILL.md; explicit invocation remains available.
Intent, implementation and final judgment
| Skill | Use it for |
|---|---|
| implement | Implement accepted behavior, repair defects or execute a selected wave with per-lane evidence. Use when: coding is authorized and ready; return facts, not a binding verdict. |
| plan | Define intended behavior, review write scope and assess reversible decisions. Use when: acceptance or approach is unclear before coding; stop once actionable. |
| validate | Freshly judge a finished change against original acceptance before merge. Use when: independent proof is needed; author tests cannot issue PASS. Triggers: "check this change". |
Engineering specialists
| Skill | Use it for |
|---|---|
| doc | Write grounded docs, READMEs, repo instructions or continuity handoffs. Use when: these documents are requested; no reports as a routine completion ritual. |
| domain | Clarify domain terms, bounded contexts and repository conventions. Use when: naming, rule ownership or Go and other language standards are unclear; avoid a broad survey. |
| refactor | Simplify structure, interfaces or responsibilities while preserving behavior. Use when: a focused refactor is requested; feature changes need their own intent. |
| research | Trace code or test a recurring pattern to answer one cited question. Use when: uncertainty needs evidence. Not for external feature teardowns; use reverse-engineer. |
| reverse-engineer | Tear down an authorized competitor repo, binary or product into a feature inventory and adoption choices. Use when: comparing an external system; local questions go to Research. |
| security | Review code or scan for security vulnerabilities, secrets, dependencies and prompt risks. Use when: concrete exposure needs assessment; never silently change policy. |
| skill-builder | Create, adapt, consolidate or repair skill packages and projections. Use when: authoring guidance, descriptions or structure; Skill Eval measures behavioral benefit. |
| skill-eval | Measure whether a skill helps a named task or needs revision or removal. Use when: a bounded routing or coding evaluation is requested; conformance alone cannot show benefit. |
| test | Write behavioral tests, practice TDD or inspect important coverage gaps. Use when: test design or missing proof needs work; running an existing suite needs no skill. |
Memory on demand
| Skill | Use it for |
|---|---|
| memory | Recall reviewed lessons or deliberately mine and curate experience. Use when: prior evidence can change an action, or learning is requested; no mandatory recall or lesson. |
Deliberate planning and review strategies
| Skill | Use it for |
|---|---|
| council | Compare independent views on a consequential or contested decision. Use when: the caller selects multiple judges; evidence resolves disagreement, not voting. |
| craft-goal | Draft or lint a bounded persistent goal above a bead graph of RPI experiments. Use when: this goal workflow is explicitly selected; shaping a single change belongs to Plan. |
| idea-genie | Generate evidenced options or challenge an idea. Use when: deciding what to build or comparing alternatives; exploration does not authorize implementation. |
| postmortem | Analyze outcomes or an interim cutoff. Use when: a postmortem is explicitly requested; consumes available judgment, never gates code acceptance or requires a lesson. |
| premortem | Challenge a rollout plan with one fresh judge before implementation; identify what could make it fail. Not for finished-code judgment. Triggers: "one judge", "challenge this plan". |
| reality-check | Check whether a claimed shipped feature, repo state or goal status holds up in evidence. Use when: comparing a claim with what exists; a gap report is not a verdict. |
| rpi | Apply the outcome-to-judgment charter. Use when: the caller explicitly selects RPI; ordinary coding, delegation and native goals do not require this workflow. |
Explicit tool and runtime adapters
| Skill | Use it for |
|---|---|
| account-rotation | Switch coding-agent accounts and verify runtime identity. Use when: the caller requests an account change; never rotate automatically to evade a quota. |
| agent-mail | Coordinate selected writers with Agent Mail messages and advisory file reservations. Use when: this adapter is requested; mail does not own tracker status. |
| agent-native | Dispatch independent tasks to parallel workers or selected persistent roles. Use when: delegation is authorized with disjoint scopes; execution does not validate output. |
| agy-native | Run a supplied task in AGY Antigravity and collect its result. Use when: the caller selects AGY; never a fallback for native coding. |
| cass | Search agent session logs and cited episodes with CASS. Use when: past prompts, decisions or failures may answer a question; repeated text is not a proven lesson. |
| cc-hooks | Configure Claude Code hooks and narrow enforcement guards. Use when: the caller requests hook installation, repair or policy changes; a hook is not required to use other skills. |
| codex-exec | Run one prompt through headless Codex and capture its result. Use when: requesting a single noninteractive Codex process. Not for worker batches or retries. |
| dcg | Diagnose a Destructive Command Guard block or configure its rules. Use when: DCG rejected an operation or policy work is requested; never disguise commands to bypass it. |
| ms | Find and load guidance with the meta_skill search engine. Use when: searching a skill corpus; CASS owns past sessions and Skill Builder owns package authoring. |
| ntm | Operate selected NTM agent panes and inspect native state. Use when: persistent tmux roles are requested; pane liveness and prompt delivery are not validation. |
| rch | Offload one build through RCH or diagnose its remote compiler. Use when: remote compilation is selected; report errors without creating a retry controller. |
| sbh | Inspect disk pressure with SBH and perform an authorized recovery action. Use when: storage diagnosis or SBH recovery is requested; inspection does not authorize deletion. |
| using-flywheel | Operate the Agentic Coding Flywheel through its native workflow. Use when: the caller explicitly selects this factory; convergence and closed work do not prove semantic acceptance. |
| using-gc | Operate Gas City through its Mayor, registry packs and native run state. Use when: the caller explicitly selects Gas City; factory completion does not replace independent judgment. |
Complete inventory
| Skill | Tier | Disposition | Hard dependencies | Capabilities | Effects |
|---|---|---|---|---|---|
account-rotation |
execution | keep_optional_adapter |
- | account_rotation |
rotate_agent_account |
agent-mail |
execution | keep_optional_adapter |
- | agent_mail |
write_agent_mail_records, install_precommit_guard, authorized_destructive_reset |
agent-native |
meta | keep_optional_adapter |
- | role_dispatch, observe_workers, handoff, dispatch_once |
manage_runtime_sessions, invoke_selected_executor |
agy-native |
cross-vendor | keep_optional_adapter |
- | dispatch_explicit_packet, provide_fresh_context |
start_agy_session |
cass |
execution | keep_optional_adapter |
- | cass |
rebuild_local_index, sync_remote_sources, download_semantic_model |
cc-hooks |
execution | keep_optional_adapter |
- | cc_hooks |
write_hook_config, append_guardrail_telemetry, write_session_sentinel |
codex-exec |
orchestration | keep_optional_adapter |
- | codex_exec |
run_codex_process, sandbox_tiered_workspace_and_network_effects |
council |
judgment | keep_strategy |
- | collect_independent_judgments, synthesize_disagreement |
write_advisory_council_report |
craft-goal |
judgment | keep_strategy |
- | goal_prompt_design, goal_prompt_lint |
- |
dcg |
execution | keep_optional_adapter |
- | dcg |
write_dcg_config |
doc |
product | keep_specialist |
- | doc, initialize_missing_docs, write_session_handoff |
write_documentation, write_requested_handoff, create_requested_evidence_directory |
domain |
knowledge | keep_specialist |
- | domain, clarify_domain_language, reconcile_domain_names |
update_existing_domain_contracts |
idea-genie |
execution | keep_strategy |
- | generate_evidenced_options, dueling_idea_genies |
write_idea_portfolio |
implement |
execution | keep |
- | execute_one_experiment, collect_factual_evidence |
modify_declared_subject, derive_subject_manifest |
memory |
execution | keep_off_path |
- | recall_applicable_context, mine_supported_observations, curate_topic_pages, toil_mining |
write_protected_drafts, update_authorized_topic_pages, write_requested_toil_report |
ms |
execution | keep_optional_adapter |
- | ms |
spawn_search_server, write_feedback_outcomes, rebuild_search_index |
ntm |
execution | keep_optional_adapter |
- | ntm |
manage_ntm_panes, dispatch_pane_commands |
plan |
execution | keep |
- | shape_intent, define_acceptance, bound_write_scope |
update_intent_source |
postmortem |
judgment | keep_strategy |
- | postmortem |
write_postmortem_report |
premortem |
judgment | keep_strategy |
- | challenge_plan |
write_advisory_plan_review |
rch |
execution | keep_optional_adapter |
- | rch |
remote_compilation_offload, authorized_remote_daemon_worker_mutation |
reality-check |
judgment | keep_strategy |
- | compare_claim_to_evidence, measure_declared_goals, report_native_status |
write_advisory_gap_report, write_goal_snapshot, write_requested_rendered_spec |
refactor |
execution | keep_specialist |
- | refactor |
modify_source_files |
research |
execution | keep_specialist |
- | research, codebase_recon, pattern_mining |
write_research_report, write_recon_pack, write_pattern_evidence |
reverse-engineer |
execution | keep_specialist |
- | reverse_engineer |
clone_upstream_repo, authorized_binary_execution, write_teardown_artifacts |
rpi |
meta | keep_strategy |
plan, implement, validate |
own_authorized_outcome, report |
dispatch_core_phases |
sbh |
execution | keep_optional_adapter |
- | sbh |
delete_reclaimable_files, release_disk_ballast, modify_host_storage_config |
security |
product | keep_specialist |
- | security |
write_scan_artifacts |
skill-builder |
meta | keep_specialist |
- | skill_builder, heal_skill, export_skill, distill_expertise |
write_skill_source, write_build_report, regenerate_skill_projections, repair_skill_projections, write_converted_skill_projection, write_advisory_proposal |
skill-eval |
meta | keep_specialist |
- | author_seeded_probe, run_probe_tier, evaluate_skill_decision |
write_probe_package, dispatch_probe_producer |
test |
execution | keep_specialist |
- | test |
write_test_files, write_test_evidence, modify_source_files |
using-flywheel |
execution | keep_optional_adapter |
- | route_to_native_flywheel_workflow, expose_agentops_skills, observe_flywheel_runtime |
- |
using-gc |
execution | keep_optional_adapter |
- | dispatch_explicit_packet, observe_gc_runtime, inspect_pack_registries, drive_mayor_door |
operate_gas_city, configure_codex_trust |
validate |
judgment | keep |
- | compute_subject_identity, judge_acceptance, return_validation_result, persist_verdict |
write_verdict_artifact |