
Distills multi-agent debate into one LLM through reasoning-enhanced fine-tuning, trajectory-based augmentation, and process-aware distillation. NeurIPS 2026.
Sep 24, 2026

Distills multi-agent debate into one LLM through reasoning-enhanced fine-tuning, trajectory-based augmentation, and process-aware distillation, moving computation from inference to training.
Feb 3, 2026