Berkeley Function Calling Leaderboard
Benchmark reference for comparing model performance on function calling and tool-use tasks.
DropTicks-made scripts, skill files, workflow docs, prompt packs, and curated external references that help builders ship stronger AI systems.
Benchmark reference for comparing model performance on function calling and tool-use tasks.
Decision checklist for comparing this technology against alternatives. Knowledge base of adversary tactics and techniques targeting AI systems.
Decision checklist for comparing this technology against alternatives. Open protocol for connecting AI clients to tools, data sources, prompts, and workflows.
Operational checklist for running this technology in real projects. Knowledge base of adversary tactics and techniques targeting AI systems.
Operational checklist for running this technology in real projects. Open protocol for connecting AI clients to tools, data sources, prompts, and workflows.
Practical implementation reference for product and engineering teams. Knowledge base of adversary tactics and techniques targeting AI systems.
Practical implementation reference for product and engineering teams. Open protocol for connecting AI clients to tools, data sources, prompts, and workflows.
Primary source documentation for evaluating and using this technology. Knowledge base of adversary tactics and techniques targeting AI systems.
Primary source documentation for evaluating and using this technology. Open protocol for connecting AI clients to tools, data sources, prompts, and workflows.
Decision checklist for comparing this technology against alternatives. NIST framework for managing AI risks, governance, measurement, and trustworthiness.
Decision checklist for comparing this technology against alternatives. AWS documentation for foundation models, agents, knowledge bases, guardrails, and inference.
Operational checklist for running this technology in real projects. NIST framework for managing AI risks, governance, measurement, and trustworthiness.