Skills Agentes

Convex Self Heal

Convierte un error de producción en un PR de fix triado, con causa raíz identificada, reparado y certificado (tsc + rehearsal + reproduce-then-gone) para que un humano lo mergee — nunca hace merge automático.

Oficial

Reemplaza a: Flujos de error→PR de Sentry/Datadog/Vercel que tratan el backend como opaco y no certifican el diff antes del review humano

Estrellas
59

en todo el repo

Actividad
55

0–100, la ruta de este skill

Actualizado
el mes pasado

último commit aquí

Commits
1

últimos 90 días

Contexto
1.4k tok

49 tok en reposo

Paquete
1 archivo

5 KB

Instalar

Funciona con cualquier agente que lea SKILL.md

npx -y skills add get-convex/agent-skills --skill convex-self-heal --agent claude-code

Se instala solo en este repositorio.

Qué hace

  • Toma un error de producción y lo triaja, root-causea, repara en una rama y certifica el fix (tsc + rehearsal + reproduce-then-gone) antes de abrir un PR
  • Orquesta sentinel (captura) → findings bus (diagnóstico) → fixers (repara) → migrate-rehearse/tsc/probe (certifica) → PR humano (decide) → deploy-guard (promueve)
  • Clasifica errores como transitorios, de config o defectos de código/schema, y solo repara los últimos
  • Tras el merge y deploy, revisa la tabla sentinel + logs para confirmar que el error deja de repetirse, y reabre si persiste
  • Solo auto-prepara clases de fix pre-aprobadas (validator, índice, ownership-check, backfill no destructivo); deriva lo destructivo o ambiguo a un humano

Úsalo cuando

  • Aparece un error nuevo o recurrente en producción capturado por sentinel y hace falta triaje, root-cause y fix certificado
  • Se necesita un PR de reparación con evidencia de certificación (tsc, rehearsal, reproduce-then-gone) para que un humano lo revise y mergee
  • Después de un merge, hay que confirmar que la firma de error dejó de aparecer en prod

No lo uses cuando

  • No hay sentinel instalado (el skill ofrece instalarlo y se detiene, no hay nada que curar sin captura)
  • El error es un blip transitorio de red (se reintenta/ignora, no se abre PR)
  • La causa raíz es incierta (se detiene y reporta en vez de proponer un fix no certificado)

Qué lo activa

Di cualquiera de estas frases y el agente debería cargar este skill.

  • “Tenemos un error recurrente en sentinel, investígalo y prepara un fix certificado”
  • “Revisa este error de producción y abre un PR con la reparación si se puede certificar”
  • “Confirma si el error de la última semana dejó de aparecer tras el deploy”

SKILL.md

En inglés

Gated production self-healing loop

Sentry/Datadog/Vercel can go error→investigate→draft-PR, but they treat the backend as opaque and stop at the human merge gate with an unverified diff. Convex can do the step they can't: because the error rows live in the user's own deployment and the fix can be rehearsed on a preview of that deployment, the platform certifies the fix against real invariants before anyone reviews it. This capability is the composition capstone — it wires sentinel (capture) → the findings bus (diagnose) → the fixers (repair) → migrate-rehearse/tsc/probe (certify) → a human PR (decide) → deploy-guard (promote). The human keeps the merge button; the machine does everything up to and including proving the fix works.

Workflow

  1. GUARD: deploy-guard — this loop reads prod and PROPOSES prod changes; classify + announce the deployment and get the standing consent for the loop's scope up front (what classes of fix it may auto-prepare vs must always defer). Never auto-merge; the human merge is the fixed boundary.
  2. CAPTURE: require sentinel (prod errors in the user's own deployment, redacted at write time). If absent, offer to install it and stop — there is nothing to heal without capture.
  3. TRIAGE a new/ recurring error: pull it via the official MCP (data/run-once-query over the sentinel table, or the monitor's prod_error event). Classify: transient (retry/ignore — do NOT open a PR for a one-off network blip), config (env/secret — hand to env, never guess a secret), or a code/schema defect (proceed).
  4. ROOT-CAUSE on the findings bus: run the relevant audit pass on the implicated function — convex-insights (the failing requests + stacks), convex-advisor (if it's a read-limit/OCC cause), convex-reviewer/convex-authz (if it's a logic/authz defect). Produce a bus finding with evidence (the stack + the reproducing input) and a fixCapability. If root cause is unclear, STOP and report — a wrong fix is worse than an open error.
  5. REPAIR via the finding's fixCapability (convex-authz, reviewer fixers, convex-expert for perf) on a branch — never on prod directly.
  6. CERTIFY against the backend's own invariants BEFORE proposing (this is the differentiator — do not skip any that apply): (a) tsc --noEmit clean; (b) if the fix touches schema/data, run it through migrate-rehearse on a preview seeded with a prod snapshot — the schema-conformance gate must pass on real-shaped data; (c) reproduce-then-confirm-gone: replay the error's triggering input against the fixed code (a convex-test case or an MCP run on the preview) and assert the failure no longer occurs; (d) no-regression: the finding must be gone AND no new bus finding introduced on the touched function. A fix that fails any applicable certification is NOT proposed — it's reported as 'attempted, could not certify' with what failed.
  7. PROPOSE, never merge: open a PR (or a diff for review) containing the fix, the certification evidence (tsc result, rehearsal outcome, the reproduced-then-gone assertion), the original error + finding, and the reversibility note. Label the change class. The human reviews and merges.
  8. PROMOTE on merge via deploy-guard's prod consent; after deploy, re-check the sentinel table + logs (failures) to confirm that error signature stops recurring (do NOT use insights for this — it tracks only OCC/read-limit perf events, not arbitrary error signatures) — the loop is only closed when the error stops recurring in prod. If it recurs, reopen with the new evidence.
  9. BOUND it: only classes the user pre-approved in step 1 are auto-prepared (default-safe set: validator fixes, missing-index adds, ownership-check adds, non-destructive backfills); anything destructive, security-sensitive beyond an added check, or ambiguous is always deferred to explicit human direction. Log every action to an append-only record so the loop is auditable.

Rules

  • The human keeps the merge button — this loop prepares and certifies fixes, it NEVER auto-merges or auto-deploys to prod (matches the industry boundary: no credible system ships unattended prod auto-merge).
  • Certify before proposing: tsc + (schema→migrate-rehearse on a prod-snapshot preview) + reproduce-then-confirm-the-failure-is-gone + no new bus finding. An uncertified fix is reported as 'could not certify', never proposed as done.
  • Triage first: transient blips get retried/ignored, config errors go to env (never guess a secret), only real code/schema defects enter the repair loop.
  • Repair on a branch/preview, never on prod directly; promote only through deploy-guard's fresh prod consent.
  • Only pre-approved fix classes are auto-prepared (default-safe: validator/index/ownership/non-destructive backfill); destructive or ambiguous changes are always deferred to the human.
  • Close the loop for real: after merge+deploy, confirm the error signature stops recurring via the sentinel table + logs (not insights, which only sees perf events); reopen if it persists.
  • Every action is logged to an append-only, auditable record; data residency stays in the user's own deployment (sentinel discipline).
  • If root cause is unclear, STOP and report — an uncertain fix is worse than an open, visible error.

Reproducido de get-convex/agent-skills bajo licencia Apache-2.0. Leer esta página en markdown.

Archivos

1 archivo en el paquete. Solo se lee SKILL.md al activarse — las referencias se cargan si el skill decide que las necesita.

Antes de instalar

Requiere sentinel instalado para capturar errores de producción en el propio deployment del usuario.

Detalles

Creador
get-convex
Licencia
Apache-2.0
Recursos incluidos
Solo SKILL.md
Código fuente
Ver SKILL.md

Más de get-convex/agent-skills

Este repo incluye 33 skills. Si instalas uno, normalmente ya tienes los demás. Ver el pack agent-skills entero y su comando de instalación

Audita y refuerza la autorización de una app Convex: impersonación por identity-from-arg, checks de ownership por documento faltantes, queries públicas que filtran datos por un id del cliente y escrituras en un contenedor ajeno.

Costo de contexto al activarse
2.3k tok
Tamaño del paquete
1 archivo
Última actualización
hace 22 días
Oficialseguridad

Diseña y construye backends reactivos y type-safe de nivel producción en Convex: schema, queries/mutations/actions, índices, auth, storage, scheduling, multiplayer en tiempo real y workflows LLM/agentes.

Costo de contexto al activarse
972 tok
Tamaño del paquete
1 archivo
Última actualización
hace 22 días
Oficialdesarrollo apis

Levanta un template Next.js + Convex barebones a partir de una idea en una frase.

Costo de contexto al activarse
1.1k tok
Tamaño del paquete
3 archivos
Última actualización
hace 22 días
Oficialdesarrollo apis

Especialista en el backend de Convex: código dentro de convex/ (funciones, schemas, índices, queries, mutations, actions, endpoints HTTP, cron jobs, storage, auth y componentes).

Costo de contexto al activarse
1.1k tok
Tamaño del paquete
1 archivo
Última actualización
hace 22 días
Oficialbases de datos

Añade un backend de agente de IA / RAG (@convex-dev/agent) a la app Convex.

Costo de contexto al activarse
430 tok
Tamaño del paquete
1 archivo
Última actualización
hace 29 días
Oficialdesarrollo apis

Construye componentes Convex reutilizables con tablas aisladas y APIs de cara a la app. Útil para nuevos componentes, módulos de backend reutilizables, integraciones o límites de componentes.

Costo de contexto al activarse
2.6k tok
Tamaño del paquete
7 archivos
Última actualización
el mes pasado
Oficialherramientas desarrollo

Skills relacionados

Wizard

267k

Genera un wizard bash interactivo que guía a un humano por los pasos que solo él puede dar. Úsalo para aprovisionar infraestructura, credenciales o secretos de CI, o una migración puntual.

Costo de contexto al activarse
1k tok
Tamaño del paquete
3 archivos
Última actualización
el mes pasado
devops infraestructura

Valida el paquete de preview de Next.js específico de un commit y dispara manualmente toda la suite de tests de despliegue vía el workflow test_e2e_deploy_release.yml de GitHub Actions.

Costo de contexto al activarse
1.2k tok
Tamaño del paquete
2 archivos
Última actualización
hace 29 días
Oficialdevops infraestructura

Canary

134k

Monitoreo canary post-deploy: vigila la app en producción tras el despliegue. (gstack)

Costo de contexto al activarse
14.8k tok
Tamaño del paquete
2 archivos
Última actualización
el mes pasado
Permisos
devops infraestructura