BAAI Proposes Wnuan: 3-Stage Post-Training for Enterprise QA
BorisMPower · x · 2026-08-05
BAAI introduced Wnuan, a post-training pipeline for question answering over proprietary enterprise knowledge, aiming to acquire enterprise knowledge without discarding general capabilities.
Methodology
Wnuan utilizes a three-stage pipeline:
- Constructs task-oriented supervision from documents.
- Performs supervised fine-tuning (SFT) with general-data replay.
- Applies reinforcement learning (RL) to residual errors.
Results
- On the 707-question WnuanBench, the 32B route raised the acceptable-answer rate from 52.76% (pre-adaptation) to 80.06% (post-SFT), reaching 91.51% after RL.
- Under a matched 100-update protocol, residual-error sampling outperformed full-pool and size-matched random sampling.
- The trade-off is a 5.17-point decrease in the general-benchmark average, primarily concentrated in instruction following.
More from Research
- MIT Team Uses AI to Design Novel Solvents, Boosting Sodium-Metal Battery Fast Charging — nordicinst · 2026-08-05
- Agent Harnesses and Prompting Drive Up to 30x Cost Swings, Benchmark Reveals — omarsar0 · 2026-08-05
- MIT Uses AI to Screen 100k Molecules in a Day, Solving Sodium Battery Fast-Charging Bottleneck — MIT News AI · 2026-08-05
- Measuring the Road to AGI: Insights from ARC Prize President — AnneliesGamble · 2026-08-05
- Training Models on Office Work Improves Coding by 5.8% — echen · 2026-08-05
- Pain Point in AI Academia: Difficulty Adding Authors After Conference Acceptance — liuzhuang1234 · 2026-08-05