Available today in the langchain-ai/langchain-skills repo: https://github.com/langchain-ai/langchain-skills
💻 Install in Codex or Claude Code, open the repo w/ the agent you want to evaluate, start with a prompt
✨ Get a @Harborframework task under evals/, an actual target…
building evals is hard! we're working on some skills to try to automate as much as possible. still requires human in the loop, but should help bootstrap
overall flow is:
- give coding agent the codebase + actual traces
- iterate on eval direction with user
- build evals (using…