Skills Agentes

Citation Fixer

Audita y corrige el formato de citas en las páginas del brain, asegurando que cada hecho tenga [Source: ...]; resuelve referencias a tweets sin URL vía la API de X.”

Solicitasearchget_pageput_pagelist_pages
Estrellas
28.9k

en todo el repo

Actividad
42

0–100, la ruta de este skill

Actualizado
hace 3 meses

último commit aquí

Commits
0

últimos 90 días

Contexto
1.5k tok

71 tok en reposo

Paquete
2 archivos

6 KB

Instalar

Funciona con cualquier agente que lea SKILL.md

npx -y skills add garrytan/gbrain --skill citation-fixer --agent claude-code

Se instala solo en este repositorio.

Este skill makes network requests.

Qué hace

  • Escanea páginas del brain buscando citas [Source: ...] faltantes o mal formateadas
  • Reescribe citas malformadas para que coincidan con conventions/quality.md
  • Detecta referencias a tweets/posts sin URL y las resuelve vía la API de X
  • Construye enlaces deterministas https://x.com/<handle>/status/<id> verificados
  • Reporta conteo de páginas escaneadas, citas encontradas, corregidas y huecos restantes

Úsalo cuando

  • Se pide auditar o corregir citas en páginas del brain
  • Hay referencias a tweets/posts sin URL que necesitan resolverse
  • Se ejecuta un barrido semanal por cron de páginas con referencias rotas
  • Otras skills como enrich o media-ingest necesitan validar citas antes de hacer commit

No lo uses cuando

    Qué lo activa

    Di cualquiera de estas frases y el agente debería cargar este skill.

    • Corrige las citas de esta página
    • Haz una auditoría de citas en el brain
    • Revisa las referencias a tweets rotas y resuélvelas
    • Ejecuta un barrido de citation-fixer en las páginas recientes

    SKILL.md

    En inglés

    Citation Fixer Skill

    Convention: see conventions/quality.md for the canonical citation format every fix should match.

    Output rule: all links MUST be deterministic (built from API data, not composed by LLM). See _output-rules.md.

    Contract

    This skill guarantees:

    • Every brain page is scanned for citation compliance.
    • Missing citations are flagged with specific location.
    • Malformed citations are fixed to match the standard format.
    • (v0.25.1) Tweet / post references without URLs are resolved via X API and patched with deterministic https://x.com/<handle>/status/<id> links.
    • Results reported with counts (scanned, fixed, remaining).

    Phases

    1. Scan pages. List pages and read each one, checking for inline [Source: ...] citations.
    2. Identify issues:
      • Facts without any citation
      • Citations missing date
      • Citations missing source type
      • Citations with wrong format
      • (v0.25.1) Tweet references without x.com URLs
    3. Fix format issues. Rewrite malformed citations to match conventions/quality.md.
    4. (v0.25.1) Resolve tweet references via the X API integration.
    5. Report results. Count: pages scanned, citations found, issues fixed, tweets resolved, remaining gaps.

    Tweet resolution pipeline (v0.25.1 extension)

    For each broken tweet reference, follow this chain. The actual API call goes through whatever X integration the host has configured (typical shape: a recipe under recipes/x-api/ with handle / search-all endpoints).

    Step 1: Identify broken references

    Scan the page for patterns that indicate tweet references without URLs:

    • Contains words like tweeted, posted, said on X, RT, retweet, X post
    • Contains quoted text that looks like a tweet (short, punchy, often starts with a quote)
    • Has [Source: ... X/Twitter ...] without an x.com URL
    • References engagement metrics (likes, impressions) without a link

    Step 2: Extract searchable content

    From each broken reference, extract:

    • The handle (if mentioned: @<username>)
    • The quoted text (if available)
    • The approximate date (often present in surrounding timeline entries)

    Step 3: Search for the actual tweet

    Use the host's X API integration. Query patterns:

    # Handle + quoted text:
    from:<handle> "<exact quote fragment>"
    
    # Quoted text only:
    "<exact quote fragment>"
    
    # Original of a retweet:
    "<exact quote>" -is:retweet
    

    Step 4: Verify and extract metadata

    Once a candidate is found:

    • Confirm the text matches the quoted fragment.
    • Pull the tweet id, author handle, engagement metrics (likes / RTs / impressions).
    • Construct the URL: https://x.com/<handle>/status/<tweet_id>.

    Step 5: Patch the brain page

    Replace the broken citation with a proper one:

    Before:

    "<quote fragment>" [Source: <some hand-wavy attribution>]
    

    After:

    "<full verified quote>" — <N> likes, <N> RTs, <N> impressions
    [Source: [X/<handle>, YYYY-MM-DD](https://x.com/<handle>/status/<tweet_id>)]
    

    Batch mode

    When sweeping many pages:

    Find candidate pages

    # Pages mentioning tweets but with no x.com links
    for f in $(find . -name "*.md" -not -path "./node_modules/*"); do
      refs=$(grep -ci "tweet\|posted\|x post\|RT\|retweet\|said on X" "$f")
      links=$(grep -c "x.com/.*/status/" "$f")
      if [ "$refs" -gt 2 ] && [ "$links" -eq 0 ]; then
        echo "$f"
      fi
    done
    

    Priority order

    1. Recently created / updated pages — fresh broken refs are easiest to resolve while context is fresh.
    2. High-traffic pages (frequent reads / writes from other skills).
    3. Everything else — bulk cleanup over time.

    Rate limiting

    • X API: respect the host's tier limits; don't hammer.
    • Target ~50 pages per batch run.
    • 1-3 API calls per page (search + verify).
    • Batch-commit every 10-20 pages so a partial failure doesn't lose progress.

    Output format

    Citation Audit Report
    =====================
    Pages scanned:        N
    Citations found:      N
    Issues fixed:         N
    Tweet links resolved: N
    Remaining gaps:       N (pages with uncitable facts)
    

    Anti-Patterns

    • ❌ Inventing citations for facts that have no source. Flag them.
    • ❌ Removing facts that lack citations (flag them; don't delete).
    • ❌ Fixing citations without reading the full page context.
    • ❌ Batch-fixing without checking quality on a sample first (see conventions/test-before-bulk.md).
    • ❌ Composing tweet URLs by guessing the tweet id. Always go through the X API; deterministic links only.

    Integration

    This skill can be called:

    • Manually — "fix citations on this page"
    • As a batch cron — weekly sweep of pages with broken refs
    • By other skillsenrich or media-ingest can call citation-fixer before commit to validate output

    Metrics

    If running as a recurring batch, track state in a small JSON file under ~/.gbrain/citation-fixer-state.json:

    {
      "last_run": "2026-04-15T...",
      "pages_scanned": 0,
      "citations_fixed": 0,
      "tweet_links_resolved": 0,
      "citations_unresolvable": 0,
      "pages_remaining": 1424
    }
    

    Output Format

    The skill's output shape is documented inline in the body sections above (see "Output", "Brain page format", or equivalent). The literal section header here exists for the conformance test (test/skills-conformance.test.ts).

    Reproducido de garrytan/gbrain bajo licencia MIT. Leer esta página en markdown.

    Archivos

    2 archivos en el paquete. Solo se lee SKILL.md al activarse — las referencias se cargan si el skill decide que las necesita.

    Antes de instalar

    Requiere la integración con la API de X/Twitter del host para resolver referencias a tweets.

    Detalles

    Creador
    garrytan
    Categoría
    Documentos
    Licencia
    MIT
    Recursos incluidos
    Incluye scripts o referencias
    Repositorio
    garrytan/gbrain
    Código fuente
    Ver SKILL.md

    Etiquetas

    Más de garrytan/gbrain

    Este repo incluye 75 skills. Si instalas uno, normalmente ya tienes los demás.

    Setup

    28.9k

    Configura GBrain con auto-aprovisionamiento de Supabase o PGLite, inyección en AGENTS.md y primera importación.

    Costo de contexto al activarse
    7.4k tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 4 días
    bases de datos

    Chequeos de salud del brain: aplicación de back-links, auditoría de citas, validación de filing, detección de info obsoleta, páginas huérfanas y benchmarks.

    Costo de contexto al activarse
    5k tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 4 días
    productividad

    Migra un brain de gbrain-base a la taxonomía de 14 tipos canónicos de gbrain-base-v2 usando gbrain onboard --check y el handler Minion unify-types.

    Costo de contexto al activarse
    3.2k tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 5 días
    bases de datos

    Cuándo y qué recuperar: abre la página del brain de una entidad relevante antes de responder desde memoria.

    Costo de contexto al activarse
    740 tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 1 hora
    productividad

    Operaciones del brain: búsqueda primero, ciclo leer-enriquecer-escribir, atribución de fuentes, enriquecimiento ambiental y back-linking. Leer antes de cualquier interacción con el brain.

    Costo de contexto al activarse
    2.6k tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 3 días
    productividad

    Importa exports de ChatGPT, Claude y Perplexity y transcripciones de sesiones como páginas fechadas en conversations/, valida y extrae hechos, y mantiene el archivo sin huecos con detección y backfill.

    Costo de contexto al activarse
    5k tok
    Tamaño del paquete
    2 archivos
    Última actualización
    hace 4 días
    productividad

    Skills relacionados

    Toma un libro (EPUB/PDF) y genera un análisis personalizado capítulo a capítulo: cada capítulo se conserva en detalle y se refleja en la vida real del lector usando el contexto de su brain.

    Costo de contexto al activarse
    6.5k tok
    Tamaño del paquete
    2 archivos
    Última actualización
    hace 9 días
    documentos

    Genera un PDF de calidad de publicación desde cualquier brain page usando el binario make-pdf de gstack; la brain page siempre es la fuente de verdad, el PDF es solo una renderización.

    Costo de contexto al activarse
    1.7k tok
    Tamaño del paquete
    2 archivos
    Última actualización
    hace 3 meses
    documentos

    Reglas de decisión para archivar páginas nuevas del brain según el tema principal, no el formato ni la fuente. Referencia para todas las skills de escritura.

    Costo de contexto al activarse
    407 tok
    Tamaño del paquete
    1 archivo
    Última actualización
    hace 4 meses
    documentos