Full Width [alt+shift+f] Shortcuts [alt+shift+k]
Sign Up [alt+shift+s] Log In [alt+shift+l]
1

Patterns for Building Cybersecurity Evals

from Eugene Yan [alt+shift+b] in AI

A sandboxed target, inputs that influence task difficulty, tools, and a grader.
21st Jun 2026

Stay updated

Get a weekly newsletter with the top 5 articles worth reading every week.

More from Eugene Yan

How to Work and Compound with AI

Context as infra, taste as config, verification for autonomy, scale via delegation, closing the loop.

3rd May 2026 • 1 votes
2025 Year in Review

An eventful year of progress in health and career, while making time for travel and reflection.

14th Dec 2025 • 1 votes
Product Evals in Three Simple Steps

Label some data, align LLM-evaluators, and run the eval harness with each change.

23rd Nov 2025 • 1 votes
Advice for New Principal Tech ICs (i.e., Notes to Myself)

Based on what I've learned from role models and mentors in Amazon

19th Oct 2025 • 1 votes

More in AI

The Business of Building God

a look at the changing economics of AI labs

a week ago • 1 votes
It’s Been a Minute

I’ve been meaning to write about *all of this*.

a week ago • 1 votes
The Overhang

Using your deep knowledge, wide knowledge, taste, and agency

a week ago
📚 BoredReading

You seem to be enjoying this.

Join free to unlock everything.

Create free account

Already have an account? Sign in