LM

Luis Mori Guerra

Engineering Manager / Technical Lead

HomeAbout MeTopics

Recent Articles

Anticipate the Path: The Simple Reason Rules, Skills, and Agents WorkHow Long Should an LLM Response Be? Deriving a Length Budget From Human Reading CapacityModern Algorithms Every Programmer Should Know: From Matching Markets to HNSW and ARCOh My Pi (omp.sh): A Quick Look at the Hardened Coding Agent CLIHow to Build the Fastest, Lightest Autonomous Agent Platform with Kubernetes, ArgoCD, Ante, and Frontier LLMs

Topics

AI Agents77AI30Architecture27LLM23Claude Code22Cursor21AI Engineering20Loop Engineering20
luismori.
ArticlesAboutTopicsGitHubConnect
← All topics

Langfuse

A topic hub collecting every article tagged Langfuse. Use it to explore related posts and follow this theme across the site.

1 article

Explore More Topics

AI Agents AI Architecture LLM Claude Code Cursor
AI Agents Evaluation Observability LangWatch LangSmith Langfuse LLMOps Testing

Trace Grading vs Scenario Testing: How to Evaluate Agents in Production

Why production agent evaluation is moving beyond output-only checks, how trace-aware grading complements scenario testing, and how LangWatch, LangSmith, and Langfuse compare.

Mar 30, 2026 12 min read

Quick find

Search the blog

Search by topic, title, framework, or pattern.

Interactive Diagram Viewer 100%
Tip: Drag to pan • Scroll to zoom • Double-click to reset • Press Esc to exit