Hi, I'm Tara! I'm a CS student at Stanford interested in security, especially where the security we expect on paper breaks down in practice. Recently, I've been working on malware-model reliability, vulnerability intelligence, enterprise security, and security for autonomous AI systems.

Selected Results

Malware-model reliability

98.0% overall accuracy hid an 85.2% false-negative rate for one malware family.

Aggregate calibration looked excellent. Family-level reliability told a very different story.

Overall ECE
0.0031
Malicord ECE
0.602

CVE enrichment

Traditional text models scored higher than zero-shot 70B+ models on severity and CWE.

Rare classes and vulnerability details that had to be inferred remained much harder.

Government security guidance

Only 2 of 166 observed controls were universal across 10 deeply analyzed frameworks.

Even close security allies differed substantially in what they recommended enterprises prioritize.