Perplexity says a later Computer checkpoint reduced tool-call failures 21.2% in a live A/B test
Perplexity post-trained a Computer model to learn from its own errors using hint-guided self-distillation. In a live A/B test, Perplexity said, a later trained checkpoint reduced tool-call failures by 21.2% relative to an earlier checkpoint.