Skills Agentes

Voice Note Ingest

Ingiere una nota de voz preservando el fraseo exacto (nunca parafraseado) y la enruta a originals/, concepts/, people/, companies/, ideas/, personal/ o voice-notes/ según un árbol de decisión.

Estrellas
28.9k

en todo el repo

Actividad
59

0–100, la ruta de este skill

Actualizado
hace 21 días

último commit aquí

Commits
2

últimos 90 días

Contexto
1.7k tok

59 tok en reposo

Paquete
2 archivos

7 KB

Instalar

Funciona con cualquier agente que lea SKILL.md

npx -y skills add garrytan/gbrain --skill voice-note-ingest --agent claude-code

Se instala solo en este repositorio.

Qué hace

  • Ingesta notas de voz preservando el fraseo exacto sin parafrasear
  • Sube el audio original al backend de almacenamiento de gbrain
  • Transcribe el audio si no se proporciona transcripción
  • Aplica un árbol de decisión para enrutar el contenido a la carpeta correcta
  • Crea back-links de línea de tiempo en las páginas de personas/empresas mencionadas

Úsalo cuando

  • El usuario envía un mensaje de audio o nota de voz por cualquier canal (Telegram, memo de voz, adjunto de openclaw)
  • Se necesita archivar una idea original, concepto, reflexión personal o pensamiento aleatorio capturado en audio

No lo uses cuando

  • Maneja una sola nota de voz por ciclo de ingesta; no hace procesamiento por lotes

Qué lo activa

Di cualquiera de estas frases y el agente debería cargar este skill.

  • Ingiere esta nota de voz
  • Transcribe y archiva este memo de voz
  • Guarda esta nota de audio

SKILL.md

En inglés

voice-note-ingest — Exact-Phrasing Voice Capture

Convention: see conventions/quality.md for citation rules, back-link enforcement, and exact-phrasing requirements.

Convention: see _brain-filing-rules.md for the filing decision protocol.

Iron Law

The user's exact words are the insight. Never paraphrase. Never clean up. The vivid, unpolished, stream-of-consciousness phrasing captures something that cleaned-up prose does not. Preserve it in block quotes. The Analysis section can interpret; the transcript section is sacred.

  • "The ambition-to-lifespan ratio has never been more fucked"
  • User noted the tension between ambition and mortality

When to invoke

The user sends an audio or voice message via any channel (Telegram, voice memo upload, openclaw audio attachment). The host agent typically provides the transcript text. If not, transcribe it with your host's transcription tool (Groq Whisper is fast and cheap; OpenAI Whisper works too — segment audio > 25MB via ffmpeg first).

The pipeline

1. STORE       → Upload original audio to gbrain storage backend
                 (S3 / Supabase Storage / local — pluggable per
                 src/core/storage.ts).
2. TRANSCRIBE  → Use the agent-provided transcript verbatim, OR
                 transcribe the audio yourself (see "When to invoke")
                 if no transcript was supplied.
3. ROUTE       → Apply the decision tree (below) to find the right
                 destination directory.
4. WRITE       → Create / update the destination brain page; preserve the
                 verbatim transcript in a block-quoted "User's Words"
                 section.
5. CROSS-LINK  → For every entity mentioned (person, company), add a
                 timeline back-link from THEIR brain page to THIS one
                 (Iron Law per conventions/quality.md).

Decision tree (where the content goes)

Apply in order. First match wins. If multiple categories apply, file to the primary directory and cross-link to the others.

  1. Original idea, observation, or thesis — the user is expressing a novel thought, framework, or connection THEY generated. → originals/<slug>.md. Use the user's vivid language for the slug.

  2. About a world concept they encountered — a framework or model someone else created that the user is referencing. → concepts/<slug>.md.

  3. About a specific person — new information, opinion, or observation about someone. → Update people/<person>.md timeline.

  4. About a specific company — new info about a company. → Update companies/<company>.md timeline.

  5. A product or business idea — something that could be built. → ideas/<slug>.md.

  6. A personal reflection — therapy-adjacent, emotional, identity. → Append to appropriate personal/<slug>.md.

  7. None of the above / random thought / doesn't fit cleanly — → voice-notes/YYYY-MM-DD-<slug>.md (catch-all).

Multiple categories? Create the primary page, then cross-link to all others. If the voice note covers a person AND a novel idea, create the originals/ page AND update the person's timeline.

Brain page format

For ALL voice-note-derived pages, include this skeleton:

---
title: "[Title derived from content]"
type: [original | concept | voice-note | ...]
created: YYYY-MM-DD
updated: YYYY-MM-DD
tags: [voice-note, relevant-tags]
sources:
  voice-note:
    type: voice_note
    storage_path: "[gbrain storage URL or relative path]"
    acquired: YYYY-MM-DD
    acquired_via: "voice note from <channel>"
---

# Title

> Executive summary of what was said and why it matters.

## User's Words

> "Exact transcript, verbatim, preserving every word, hesitation, and verbal
> tic. This is the primary source material. Do not edit."

🔊 [Audio]([gbrain storage URL or relative path])

## Analysis

[What this means, why it matters, connections to other thinking. The
analysis is the agent's interpretation; the transcript above is sacred.]

## See Also

- [Related brain pages with relative links]

---

## Timeline

- **YYYY-MM-DD** | voice note from <channel> — [Brief description]

Citation format

[Source: voice note, <channel>, YYYY-MM-DD]

Include timestamps when available:

[Source: voice note, <channel>, YYYY-MM-DD HH:MM PT]

Naming convention

  • Audio files: YYYY-MM-DD-<brief-slug>.<ext> (e.g., 2026-04-13-rick-rubin-creative-philosophy.ogg)
  • Brain pages: match the slug of the destination directory.

Bulk vs. single

This skill handles ONE voice note at a time. Each is its own ingest cycle. No batching.

Anti-Patterns

  • Paraphrasing the transcript. The exact words are the signal.
  • Cleaning up hesitations or filler words ("um", "like", "you know"). The texture matters.
  • Creating a page with no entity cross-links when people/companies were mentioned. Iron Law fail.
  • Skipping the audio storage step. Always upload the original; the brain page has a 🔊 [Audio] link back to it.

Related skills

  • skills/signal-detector/SKILL.md — same exact-phrasing pattern for text-channel idea capture
  • skills/idea-ingest/SKILL.md — for typed-text idea ingestion
  • skills/conventions/quality.md — citation + back-link rules

Contract

This skill guarantees:

  • Routing matches the canonical triggers in the frontmatter.
  • Output written under the directories listed in writes_to: (when applicable).
  • Conventions referenced (quality.md, brain-first.md, _brain-filing-rules.md) are followed.
  • Privacy contract preserved: no real names, no fork-specific filesystem path literals, no upstream-fork references.

The full behavior contract is documented in the body sections above; this section exists for the conformance test.

Output Format

The skill's output shape is documented inline in the body sections above (see "Output", "Brain page format", or equivalent). The literal section header here exists for the conformance test (test/skills-conformance.test.ts).

Reproducido de garrytan/gbrain bajo licencia MIT. Leer esta página en markdown.

Archivos

2 archivos en el paquete. Solo se lee SKILL.md al activarse — las referencias se cargan si el skill decide que las necesita.

Antes de instalar

Requiere un backend de almacenamiento gbrain (S3/Supabase/local) y, si no hay transcripción, una herramienta de transcripción como Groq Whisper u OpenAI Whisper.

Detalles

Creador
garrytan
Categoría
Productividad
Licencia
MIT
Recursos incluidos
Incluye scripts o referencias
Repositorio
garrytan/gbrain
Código fuente
Ver SKILL.md

Etiquetas

Más de garrytan/gbrain

Este repo incluye 75 skills. Si instalas uno, normalmente ya tienes los demás.

Setup

28.9k

Configura GBrain con auto-aprovisionamiento de Supabase o PGLite, inyección en AGENTS.md y primera importación.

Costo de contexto al activarse
7.4k tok
Tamaño del paquete
1 archivo
Última actualización
hace 4 días
bases de datos

Chequeos de salud del brain: aplicación de back-links, auditoría de citas, validación de filing, detección de info obsoleta, páginas huérfanas y benchmarks.

Costo de contexto al activarse
5k tok
Tamaño del paquete
1 archivo
Última actualización
hace 4 días
productividad

Migra un brain de gbrain-base a la taxonomía de 14 tipos canónicos de gbrain-base-v2 usando gbrain onboard --check y el handler Minion unify-types.

Costo de contexto al activarse
3.2k tok
Tamaño del paquete
1 archivo
Última actualización
hace 5 días
bases de datos

Cuándo y qué recuperar: abre la página del brain de una entidad relevante antes de responder desde memoria.

Costo de contexto al activarse
740 tok
Tamaño del paquete
1 archivo
Última actualización
hace 1 hora
productividad

Operaciones del brain: búsqueda primero, ciclo leer-enriquecer-escribir, atribución de fuentes, enriquecimiento ambiental y back-linking. Leer antes de cualquier interacción con el brain.

Costo de contexto al activarse
2.6k tok
Tamaño del paquete
1 archivo
Última actualización
hace 3 días
productividad

Importa exports de ChatGPT, Claude y Perplexity y transcripciones de sesiones como páginas fechadas en conversations/, valida y extrae hechos, y mantiene el archivo sin huecos con detección y backfill.

Costo de contexto al activarse
5k tok
Tamaño del paquete
2 archivos
Última actualización
hace 4 días
productividad

Skills relacionados

Archivista universal para archivos personales (Dropbox/B2/Gmail-takeout/disco local). Filtra contenido de alto valor y lo muestra de forma interactiva; exige un allow-list scan_paths explícito en gbrain.yml.

Costo de contexto al activarse
2.7k tok
Tamaño del paquete
2 archivos
Última actualización
hace 3 meses
productividad

Transforma volcados de texto crudo de artículos en el brain en páginas estructuradas con resumen ejecutivo, citas textuales, insights clave, por qué importa y referencias cruzadas.

Costo de contexto al activarse
1.5k tok
Tamaño del paquete
2 archivos
Última actualización
hace 3 meses
productividad

Filtro de calidad previo a la escritura para todo lo que entra al brain: nada de cp/mv en crudo. Resuelve entidades con nombre por registro y aplica el árbol de decisión de dedup leyendo el primer resultado.

Costo de contexto al activarse
3.7k tok
Tamaño del paquete
2 archivos
Última actualización
hace 9 días
productividad