AgentForesight introduces online trajectory auditing that predicts multi-agent failures during execution, with a 7B model outperforming GPT-4.1 and DeepSeek-V4-Pro by up to 19.9% with 3x lower step error.
Terminator learns optimal early-exit points for chain-of-thought reasoning to cut token lengths by 14%-55% and boost inference speed over 2x with minimal accuracy loss.
An online algorithm achieves optimal ε-recalibration with ε² excess error in ε⁻³ rounds via Blackwell approachability, and yields simultaneous calibration and calibeating for smooth losses.
A framework generates population-aligned personas from social media via quality filtering, importance sampling, and task-specific adaptation, reducing bias in LLM social simulations.
CoreQ proposes a learning-free post-training quantization framework using a geometric closed-form layer-adaptive mismatch correction coefficient and successive rounding to improve LLM quantization accuracy without hyperparameter tuning.