Full Width [alt+shift+f] Shortcuts [alt+shift+k]
Sign Up [alt+shift+s] Log In [alt+shift+l]

Eugene Yan

Sort By

Recent [alt+2]
Patterns for Building Cybersecurity Evals A sandboxed target, inputs that influence task difficulty, tools, and a grader.
21st Jun 2026
1

Patterns for Building Cybersecurity Evals

from Eugene Yan [alt+shift+b] in AI

21st Jun 2026
A sandboxed target, inputs that influence task difficulty, tools, and a grader.
How to Work and Compound with AI Context as infra, taste as config, verification for autonomy, scale via delegation, closing the...
3rd May 2026
1

How to Work and Compound with AI

from Eugene Yan [alt+shift+b] in AI

3rd May 2026
Context as infra, taste as config, verification for autonomy, scale via delegation, closing the loop.
2025 Year in Review An eventful year of progress in health and career, while making time for travel and reflection.
14th Dec 2025
1

2025 Year in Review

from Eugene Yan [alt+shift+b] in AI

14th Dec 2025
An eventful year of progress in health and career, while making time for travel and reflection.
Product Evals in Three Simple Steps Label some data, align LLM-evaluators, and run the eval harness with each change.
23rd Nov 2025
1

Product Evals in Three Simple Steps

from Eugene Yan [alt+shift+b] in AI

23rd Nov 2025
Label some data, align LLM-evaluators, and run the eval harness with each change.
Advice for New Principal Tech ICs (i.e., Notes to Myself) Based on what I've learned from role models and mentors in Amazon
19th Oct 2025
1

Advice for New Principal Tech ICs (i.e., Notes to Myself)

from Eugene Yan [alt+shift+b] in AI

19th Oct 2025
Based on what I've learned from role models and mentors in Amazon
Training an LLM-RecSys Hybrid for Steerable Recs with Semantic IDs An LLM that can converse in English & item IDs, and make recommendations w/o retrieval or tools.
14th Sep 2025
1
14th Sep 2025
An LLM that can converse in English & item IDs, and make recommendations w/o retrieval or tools.
Building News Agents for Daily News Recaps with MCP, Q, and tmux Learning to automate simple agentic workflows with Amazon Q CLI, Anthropic MCP, and tmux.
4th May 2025
1
4th May 2025
Learning to automate simple agentic workflows with Amazon Q CLI, Anthropic MCP, and tmux.
An LLM-as-Judge Won't Save The Product—Fixing Your Process Will Applying the scientific method, building via eval-driven development, and monitoring AI output.
20th Apr 2025
1
20th Apr 2025
Applying the scientific method, building via eval-driven development, and monitoring AI output.
Improving Recommendation Systems & Search in the Age of LLMs Model architectures, data generation, training paradigms, and unified frameworks inspired by LLMs.
16th Mar 2025
1
16th Mar 2025
Model architectures, data generation, training paradigms, and unified frameworks inspired by LLMs.
Building AI Reading Club: Features & Behind the Scenes Exploring how an AI-powered reading experience could look like.
12th Jan 2025
1

Building AI Reading Club: Features & Behind the Scenes

from Eugene Yan [alt+shift+b] in AI

12th Jan 2025
Exploring how an AI-powered reading experience could look like.
📚 BoredReading

You seem to be enjoying this.

Join free to unlock everything.

Create free account

Already have an account? Sign in