Elicit: AI for scientific research
Evaluating Elicit’s Systematic Literature Review Capabilities
Elicit hit 95% search recall, 97% abstract screening, 99% full-text screening, and 96% extraction across 994 Cochrane reviews.
Evaluating Claude Opus 4.5
Claude Opus 4.5 is better than Sonnet 4.5, Google Gemini 3 Pro, and OpenAI GPT 5 at data extraction and writing reports with fewer hallucinations.