
The Millimetre Problem: Introducing the lab-hardware-cad Agent Skill
Introducing lab-hardware-cad, an open-source Agent Skill that turns any AI agent into a CAD engineer for lab hardware, benchmarked over 98 geometry-scored runs.
Blog
Explore tutorials, product updates, and insights on AI-powered research automation from the K-Dense team.

Introducing lab-hardware-cad, an open-source Agent Skill that turns any AI agent into a CAD engineer for lab hardware, benchmarked over 98 geometry-scored runs.

K-Bench 01 is our internal benchmark built from 178 real scientific tasks. Nine frontier models ran them end to end; only one clears the acceptability bar.

K-Dense and Litmus Science are connecting AI-powered research with on-demand lab execution, helping scientists move from questions to evidence faster.

Twenty questions from a live session with a university multiomics center on what K-Dense does, where AI for science helps today, and where it still falls short.

A 20-task benchmark found K-Dense Web stronger at producing auditable research bundles, while Claude Science held a narrow scientific-quality lead.

A 35-case K-Dense benchmark of Google's Omni Flash shows stunning scientific video quality, but unreliable text, physics, and factual correctness.

A 240-image benchmark of four scientific image models shows Nano Banana 2 Lite is fastest, while GPT Image 2 still leads on raw figure quality.

The next AI scientist benchmark should be a lab escape room: a fresh hidden world, limited probes, real evidence, and no place for science theater to hide.

A controlled benchmark of 10 NVIDIA BioNeMo Agent Toolkit skills for NIM microservices shows where skills win: routing, cost, scale, and weak-model reliability.