Memory Ingest

basicmachines-co/basic-memory-skills/.agents/skills/memory-ingest

by basicmachines-co6d2b1d426d0dacf020aef45f029768c9d8c1e5e5No license24 starsListed Oct 9, 2026Updated Oct 9, 2026Repository updated 5 months ago

Process unstructured external input (meeting transcripts, conversation logs, pasted documents) into structured Basic Memory entities. Extracts entities, searches for existing matches, proposes new entities with approval, creates notes with observations and relations, and captures action items.

AI-generated overview

Turns raw transcripts, logs, and pasted documents into structured Basic Memory notes with entities, observations, and relations.

What it does
Parses unstructured input such as meeting transcripts, conversation logs, email threads, or pasted documents and extracts people, organizations, topics, and action items. It searches Basic Memory for existing entities, proposes new ones for user approval, and writes a source note preserving the original content verbatim alongside categorized observations and relations. It then creates approved entity notes and captures follow-ups and commitments.
When to use it
Use it when a user pastes a meeting transcript, conversation log, article, or email and wants it turned into structured knowledge. It fits requests to process notes or add material to Basic Memory, and any case where raw external text should become linked, searchable notes.
Requirements
Requires access to a Basic Memory knowledge graph with note search and write capabilities. Optional web research may use network access. Ships no scripts; instructions only.

Memory Ingest

Turn raw, unstructured input into structured Basic Memory entities. Meeting transcripts, conversation logs, pasted documents, email threads — anything with information worth preserving gets parsed, cross-referenced against existing knowledge, and written as proper notes.

When to Use

  • User pastes a meeting transcript or conversation log
  • User says "process these notes" or "add this to Basic Memory"
  • User pastes a document, article, or email for knowledge extraction
  • Any time raw external text needs to become structured knowledge

Workflow Overview

1. Parse raw input           → identify structure, extract key info2. Extract entities          → people, orgs, topics, action items3. Search existing entities  → multi-variation queries4. Research new entities     → optional web research (see memory-research)5. Present entity proposal   → get approval before creating6. Create source note        → verbatim content + observations + relations7. Create approved entities  → structured notes for each new entity8. Extract action items      → follow-ups and commitments

Step 1: Parse Raw Input

Read the pasted content and identify its structure:

  • Format: Meeting transcript, email thread, conversation log, article, freeform notes
  • Date: When this happened (extract from content or ask)
  • Participants: Who was involved (names, roles, organizations)
  • Sections: Any existing structure (headings, speaker labels, timestamps)

Don't rewrite or summarize the source content. Preserve it verbatim in the note — you'll add structured observations alongside it.

Step 2: Extract Entities

Scan the content for entities worth tracking in the knowledge graph:

Entity TypeSignals
PersonNames with roles, titles, or affiliations mentioned
OrganizationCompany names, agencies, institutions
Topic/ConceptTechnical domains, methodologies, standards discussed substantively
Action ItemCommitments, deadlines, "I'll do X by Y" statements

Infer type from context. If someone is introduced as "CTO of Acme Corp", that's both a Person and an Organization entity. If a technology is discussed in depth, it might warrant a Concept entity.

Exclude noise. Not every name mentioned is worth an entity. Filter for:

  • People with substantive roles or interactions (not passing mentions)
  • Organizations discussed in business/technical context
  • Topics with enough detail to warrant their own note

Step 3: Search Existing Entities

For each extracted entity, search Basic Memory with multiple query variations:

python
# Person — try full name, last namesearch_notes(query="Sarah Chen")search_notes(query="Chen")
# Organization — try full name, abbreviation, acronymsearch_notes(query="National Renewable Energy Laboratory")search_notes(query="NREL")
# Topic — try the full term and keywordssearch_notes(query="edge computing")search_notes(query="edge inference")

Classify each entity as:

  • Existing — found in Basic Memory. Will link to it with [[wiki-link]].
  • Proposed — not found. Will propose creation pending approval.

Step 4: Research New Entities (Optional)

For proposed entities where more context would be valuable, do a brief web search (2-3 queries max per entity):

  • Organizations: What they do, size, public/private, key products
  • People: Current role, background, expertise
  • Topics: Brief definition, relevance

Use hedging language ("appears to be", "estimated", "based on public information"). Never fabricate details.

This step is optional — skip it if the source material provides enough context, or if the user is in a hurry. See the memory-research skill for deeper research workflows.

Step 5: Present Entity Proposal

Before creating anything, present what you found and what you'd like to create:

Entities found in Basic Memory:  - [[Sarah Chen]] (Person — existing)  - [[Acme Corp]] (Organization — existing)
Proposed new entities:  - Jordan Rivera (Person — VP Engineering at NovaTech, mentioned as project lead)  - NovaTech (Organization — SaaS platform, Series B, discussed as integration partner)  - Federated Learning (Concept — core technical topic of the discussion)
Approve all / select individually / skip entity creation?

Include enough context with each proposed entity for the user to make a quick decision.

Step 6: Create the Source Note

Create the primary note for the ingested content. This is the "record of what happened" — it preserves the raw material and adds structured metadata.

Meeting / Conversation Note

python
write_note(  title="NovaTech Meeting - Jordan Rivera - Feb 22, 2026",  directory="meetings/2026",  note_type="meeting",  tags=["meeting", "novatech", "federated-learning"],  metadata={"date": "2026-02-22"},  content="""# NovaTech Meeting - Jordan Rivera - Feb 22, 2026
Brief one-sentence summary of what this meeting was about.
## Transcript[Preserve all source content verbatim — do not summarize or rewrite]
## Observations- [opportunity] NovaTech interested in integration partnership- [insight] Their platform handles 10K concurrent sessions, relevant to our scale needs- [next_step] Send technical spec document by Friday- [sentiment] Strong enthusiasm from their engineering team- [decision] Agreed to start with a proof-of-concept integration
## Relations- attended [[Jordan Rivera]]- with [[NovaTech]]- discussed [[Federated Learning]]- follow_up [[Send NovaTech Technical Spec]]""")

Document / Article Note

python
write_note(  title="Edge Computing Architecture Whitepaper",  directory="references",  note_type="reference",  tags=["edge-computing", "architecture", "reference"],  metadata={"source": "https://example.com/whitepaper.pdf", "date_ingested": "2026-02-22"},  content="""# Edge Computing Architecture Whitepaper
## Source Content[Preserve relevant content — for long documents, include key sections rather than the entire text]
## Observations- [key_finding] Latency drops 40% with edge inference vs cloud-only- [technique] Model sharding across heterogeneous edge nodes- [limitation] Requires minimum 8GB RAM per edge node
## Relations- relates_to [[Edge Computing]]- relates_to [[Model Optimization]]""")

Observation Categories

Use categories that capture the nature of the information. Common categories for ingested content:

CategoryUse For
opportunityBusiness or collaboration opportunities identified
decisionDecisions made or agreed upon
insightNon-obvious understanding gained
next_stepConcrete action items or follow-ups
sentimentEnthusiasm, concerns, hesitations expressed
riskRisks or concerns identified
requirementRequirements or constraints discovered
key_findingImportant facts from reference material
techniqueMethods, approaches, or patterns described
contextBackground information that may be useful later

Invent categories as needed — these are suggestions, not a fixed list.

Step 7: Create Approved Entities

For each entity the user approved, create a structured note. Match the entity type to an appropriate template.

Person

python
write_note(  title="Jordan Rivera",  directory="people",  note_type="person",  tags=["person", "novatech", "engineering"],  content="""# Jordan Rivera
## OverviewVP of Engineering at NovaTech. Met during integration partnership discussion.
## Background[Role, expertise, context from meeting + any web research]
## Observations- [role] VP Engineering at NovaTech- [expertise] Distributed systems, federated learning- [met] 2026-02-22 during integration discussion
## Relations- works_at [[NovaTech]]- discussed_in [[NovaTech Meeting - Jordan Rivera - Feb 22, 2026]]""")

Organization

python
write_note(  title="NovaTech",  directory="organizations",  note_type="organization",  tags=["organization", "saas", "integration-partner"],  content="""# NovaTech
## OverviewSaaS platform company. Series B stage.[Additional context from meeting + web research]
## Products & Services[What they offer, if discussed or researched]
## Observations- [stage] Series B, ~200 employees- [relevance] Potential integration partner for our platform- [first_contact] 2026-02-22
## Relations- employs [[Jordan Rivera]]- discussed_in [[NovaTech Meeting - Jordan Rivera - Feb 22, 2026]]""")

Concept / Topic

python
write_note(  title="Federated Learning",  directory="concepts",  note_type="concept",  tags=["concept", "machine-learning", "distributed-systems"],  content="""# Federated Learning
## Overview[Brief description of the concept from the discussion context]
## Observations- [definition] Machine learning approach where models train across decentralized data sources- [relevance] Core technique discussed in NovaTech integration
## Relations- discussed_in [[NovaTech Meeting - Jordan Rivera - Feb 22, 2026]]""")

Adapt templates to your domain. The key elements are: type and tags as parameters, an overview section, observations with categories, and relations linking back to the source.

Step 8: Extract Action Items

Review the source content for commitments and follow-ups:

Action Items:  - Send NovaTech technical spec document by Friday (your commitment)  - Jordan will share their API documentation by next week (their commitment)
Follow-Up Reminders:  - 1 week: Check if Jordan sent API docs  - 2 weeks: Schedule follow-up call to discuss POC scope

If using the memory-tasks skill, create Task notes for your action items. Otherwise, capture them as observations in the source note.

Guidelines

  • Preserve source content verbatim. The original text is the ground truth. Structure and observations are your interpretation layered on top.
  • Search before creating. Always check if entities already exist (see memory-notes search-before-create pattern). Update existing entities with new information rather than creating duplicates.
  • Get approval for new entities. Present proposed entities and let the user decide which to create. Don't silently populate the knowledge graph.
  • Infer, don't interrogate. Extract entity types and relationships from context. Only ask the user when genuinely ambiguous.
  • Be selective about entities. Not every name mentioned deserves its own note. Focus on entities the user will want to reference again.
  • Use hedging for researched info. Web research supplements — don't present it as fact. "Appears to be", "estimated", "based on public information".
  • Link everything back. Every created entity should relate back to the source note. The source note should link to all entities discussed.
  • Prose and observations together. Notes work best with both narrative context and structured observations. Prose gives meaning and tells the story; observations make individual facts searchable. Use the body for context, then distill key facts into categorized observations.

Source and attribution

Source:basicmachines-co/basic-memory-skillsin.agents/skills/memory-ingestat commit6d2b1d4

License: No license

Content belongs to its original authors. SourceWeft indexes it from a public repository.

Report or request removal