Li Bearden

AI Safety Research Engineer

Li Bearden

I build measurement infrastructure for AI safety — and the tools that pressure-test it.

LLM evaluation methodology and sycophancy reproducibility, grounded in five years of applied ML at Deepgram — where I owned the eval and custom-training infrastructure and cut model delivery from 14 days to 1.

CURRENTLY Reproducing SycEval as the open falsification test · arXiv preprint targeting August 2026.
Li Bearden

The work

Four fronts · one practice