Multi-step prompt sequences for complex AI workflows.
19 tools found
Import AI evaluation test cases from SharePoint CSV files using Azure AD certificate authentication for enterprise test management.
A Promptfoo example evaluating Claude Opus 4.6's coding, bug-diagnosis, and tradeoff-reasoning skills with a 128K extended-thinking budget.
A Promptfoo example demonstrating Anthropic's web search and web fetch tools with domain filtering and source citations.
A Promptfoo example evaluating Claude Opus, Sonnet, and Haiku models deployed on Azure AI Foundry on explanation tasks.
Promptfoo example comparing GPT, Claude, Llama, and Mistral models side by side on Azure AI Foundry.
A promptfoo example for evaluating Azure AI Foundry agents through the newer v2 Responses runtime.
Promptfoo example testing Claude's thinking feature, comparing reasoning quality between the Anthropic API and AWS Bedrock on logic puzzles.