Ingesting 100k papers into an LLM wiki: user seeks editorial policy for AI knowledge synthesis
gintrux · reddit · 2026-09-21
A Reddit user proposes using an "LLM wiki" to distill 100k+ IBS papers into a compressed knowledge base for finding overlooked interventions. Key challenge: ingestion requires constant editorial decisions (create/edit/rename pages, record relations), demanding an explicit editorial policy rather than relying on raw model intelligence. Their pipeline: export PubMed lists, download open-access PDFs, OCR to Markdown, then ingest in batches of 25 — one agent per paper in its own worktree, with a reconciliation agent merging diffs at batch end. They're asking the community for shared editorial policies.
More from coding & agent
- Moonshot's Kimi launches Code Desktop with multi-agent parallel coding on Mac and Windows — KimiDevs · 2026-09-21
- Rodin-generated 3D instruments + GPT-6 build a playable interactive instrument website in 4 steps — CurieuxExplorer · 2026-09-21
- Agent management may be the next IAM problem as AI agents become coworkers — ingliguori · 2026-09-21
- Pragmatic Engineer survey of 100+ companies: Scrum is conspicuously absent from Big Tech — blaizedsouza · 2026-09-21
- Cobro MCP lets you point at a UI element and tell Claude Code what to change — boonblade · 2026-09-21
- AutoClip: AI video clipping tool hits ~8K stars on GitHub — zhouxiaoka · 2026-09-21