Search

LangGraph · LLM evaluation

5 skills found.
Search results
#SkillRepositoryStarsUsed inTokensAuto-checkLicenceUpdated
1

Helps instrument a custom Python or TypeScript agent to record events for Failproof AI, verify what gets written, and run an evaluator worker that scores the runs.

FailproofAI/failproofai5.3k—~6kAutomated safety check: PassUnknown4 days ago
2

Decide which AI agent behaviors are worth an eval case, then write those cases — harness-, framework-, and language-agnostic.

agentailor/fullstack-langgraph-nextjs-agent132—~5.3kAutomated safety check: PassMIT1 mo ago
3

INVOKE THIS SKILL when building, testing, or deploying Managed Deep Agents in LangSmith.

langchain-ai/langchain-skills1.3k—~8.7kAutomated safety check: NotesMIT2 days ago
4

Build reproducible evaluation pipelines for LangChain 1.0 chains and LangGraph 1.0 agents — golden datasets, LangSmith evaluate(), ragas RAG metrics, deepeval LLM-as-judge, agent trajectory…

jeremylongshore/tons-of-skills-marketplace2.8k—~3.7kAutomated safety check: PassMITyesterday
5

Build type-safe AI agents and graph-based workflows with PydanticAI and PydanticGraph.

magnus919/agent-skills115—~4.3kAutomated safety check: PassMITyesterday