The ~/.claude/ directory was previously local-only — a machine wipe
would have lost the accumulated memory, custom skills, and settings.
This commit moves the durable parts into dotfiles with the same
symlink-to-home pattern used for CLAUDE.md, PENDING.md, REVIEWED.md,
and L2-BOOTSTRAP.md.
Preserved (symlinked from ~/.claude/* into here):
skills/audit/ — thinking-folder drift scanner
skills/symmetria/ — practice-of-return discipline
skills/vault-update-people/ — Obsidian People-file maintainer
skills/wake-up/ — session restoration
skills/wrap-up/ — session state capture
memory/ — 55+ memory files (MEMORY.md, sessions,
ledgers, project state, feedback, etc.)
settings/settings.json — user preferences (hooks, flags, no secrets)
Deliberately NOT backed up:
settings.local.json — contains operational secrets (HF_TOKEN,
SSH password in expect scripts); by naming
convention, *.local.* is not synced.
Needs separate review and probable rotation.
sessions/, history.jsonl, caches, telemetry — ephemeral
plugins/, marketplace skills and agents — reinstallable
The working copies at ~/.claude/skills/* and
~/.claude/projects/-Users-davidglidden/memory are symlinks into this
directory, so every write flows here automatically. install.sh
recreates the symlinks on a fresh machine.
FOLLOW-ON (flagged, not in this commit):
settings.local.json contains a HuggingFace token and an SSH password
as plaintext strings inside allowed Bash command patterns. These
should be rotated and moved to secure storage (keychain / pass /
env file outside the settings file).
Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>
12 lines
877 B
Markdown
12 lines
877 B
Markdown
---
|
|
name: PDF text extraction approach
|
|
description: AI agents cannot extract copyrighted book text — use pdftotext instead
|
|
type: feedback
|
|
---
|
|
|
|
AI extraction agents (Claude subagents) are blocked by content filtering when asked to reproduce copyrighted book text verbatim. This applies even for personal-use format-shifting.
|
|
|
|
**Why:** Content filtering policy blocks large-scale reproduction of copyrighted translations, regardless of fair-use context.
|
|
|
|
**How to apply:** For PDF-to-markdown extraction of copyrighted texts, use `pdftotext -layout` (from poppler, already installed via Homebrew) for mechanical extraction. The raw output is readable and citable, with minor OCR artifacts (spaced headings, split chapter numbers) that can be cleaned up with a script. AI agents can handle frontmatter, cataloguing, and cleanup — just not the verbatim text extraction itself.
|