Skip to main content
Back to News
Perplexity AI launched DRACO, an open-source benchmark evaluating research agents via 100 tasks from...
Technology
1 min read
US

Perplexity AI launched DRACO, an open-source benchmark evaluating research agents via 100 tasks from...

The AMW Read

Perplexity (a key player in research/search) is updating the agentic evaluation landscape with a production-grounded benchmark, but this is an incremental tool release rather than a structural shift.
NoveltySignificance
AI Agents · Player Map

Perplexity AI launched DRACO, an open-source benchmark evaluating research agents via 100 tasks from real user queries. Spanning 10 domains, Perplexity leads with 89.4 percent accuracy in Law and 82.4 percent in Academic research. Shifting from synthetic puzzles to production-grounded data creates a rigorous standard for multi-step reasoning. This systemic evolution forces the AI industry to prioritize factual depth over conversational fluency. 🚀

#AIResearch #DRACO #PerplexityAI #LLM #Technology

How This Connects

Based on AI Agents · Player Map

  1. 1d agoTencent is set to become the largest shareholder of AI developer Manus, as Meta unwinds its acquisit...Manus
  2. 4d agoMeta will unwind its $2 billion acquisition of Manus AI after Beijing ordered the deal to be reverse...Manus
  3. 1w agoHugging Face hack marks start of agentic AI cyber era, execs warn firms 'don't even know it'Hugging Face
  4. 1mo agoAlibaba Releases SkillWeaver Framework, Cutting Agent Token Consumption 99%
  5. 6mo agoPerplexity AI launched DRACO, an open-source benchmark evaluating research agents via 100 tasks from... · THIS ARTICLE

Related News

More news from Perplexity

Stay updated with the latest news and announcements from Perplexity.

View all Perplexity news

Discover AI Startups

Explore 5,000+ AI companies with VC-grade analysis, funding data, and investment insights.

Explore Dashboard