You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
PromptPotter: AI optimization engine for automated prompt engineering (APE). Searches and tunes prompts and parameters against your data using a generate-evaluate-critique cycle. Improve AI outputs, automate prompt optimization, and fine-tune your pipeline.
"EvArgs" is a Python module designed for value assignment, easy expression parsing, and type casting. It validates values based on defined rules and offers flexible configuration along with custom validation methods.