AINewsnow

Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation

Perplexity Research published a new post-training study. It trains a model inside Perplexity Computer on real user sessions, including failed ones. The method pairs rejection sampling fine-tuning with hint-guided self-distillation. In a live A/B test, tool-call failures fell from 2.24% to 1.77% bet…

Read the full story at MarkTechPost ↗

Timeline · 1 report

  1. 2026-09-25 14:30 · MarkTechPost
    Perplexity Trains Its Computer Agent on Real Mistakes With Hint-Guided Self-Distillation

More stories

  1. Need some help — r/AI_Agents
  2. DeepSeek-V4-Flash-0731 at ~40–50 tok/s on 2× Radeon AI PRO R9700 with the affinity engine (prebuilt quant + fixes) — r/LocalLLaMA
  3. Qwengram-0.8B: I transferred Qwen3.8 Flash-Next’s n-gram memory into Qwen3.5-0.8B — 5.05% lower validation perplexity — r/LocalLLaMA
  4. Perplexity's Photon Slashes Search Latency by 12x for AI Agents — AlphaSignal
  5. Perplexity Brings Portable Computer to AMD Ryzen AI Max PCs — AlphaSignal
  6. asked gemini and chatgpt the same question about local businesses and they recommended almost completely different ones. only 11% of the sources they cite overlap — r/GeminiAI
  7. open an incognito tab and ask chatgpt for the best [what you do] in your city. most owners have never checked whether they come up, and the numbers are worse than you'd think — r/PromptEngineering
  8. Bros benchmark is logo design — r/ArtificialInteligence

Get the daily brief of stories like this at 6:30 every morning →