
Beyond RDKit: Benchmarking the Rowan Agent Skill Against Experiment
A reproducible study of pKa, logD, tautomers, ADME, and docking run by an AI agent with the Rowan skill, measured against RDKit and experimental ground truth.
Blog
Explore tutorials, product updates, and insights on AI-powered research automation from the K-Dense team.
Showing 11 posts tagged "Benchmarks".

A reproducible study of pKa, logD, tautomers, ADME, and docking run by an AI agent with the Rowan skill, measured against RDKit and experimental ground truth.

K-Dense Web scored 45/50 on BixBench-Verified-50, a cleaned biology-agent benchmark designed to separate real model mistakes from benchmark noise.