Stuck at 70% Accuracy: Preventing LLMs from Altering Numbers in PDF Translation
Nervous_Classroom714 · reddit · 2026-08-31
A developer discusses a bottleneck in their PDF translation tool where the LLM occasionally reformats numbers embedded in sentences (e.g., decimal flipping, localization), despite pinning tokens. They tried extracting from source PDFs, small batches, and glossary passes. The post seeks advice on prompting patterns for in-sentence numbers, balancing batch size vs. context, and handling echoed segments.
More from coding & agent
- Anyone tried fine-tuning Minimax H3 for a character? — MoneyKenny · 2026-08-31
- AgentHQ Launches: Visualizing Multi-Agent Orchestration as a Pixel Art Office — KyeGomezB · 2026-08-31
- Workato launches Otto, an AI teammate for cross-app automation — adamse · 2026-08-31
- Tencent's ContextPilot Teaches Agents Proactive Context Management via Fine-grained RL — tencent · 2026-08-31
- Agent workflow quiz: Handling stale decisions — Darkcraft00 · 2026-08-31
- WebMCP provides a 'VIP lane' for AI agents to interact with sites — thisiskp_ · 2026-08-31