The Berkeley function calling leaderboard is a live leaderboard to evaluate the ability of different LLMs to call functions (also referred to as tools). We built this dataset from our learnings to be representative of most users' function calling use-cases, for example, in agents, as a part of enterprise workflows, etc. To this end, our evaluation dataset spans diverse categories, and across multiple languages. Checkout the Leaderboard at gorilla.cs.berkeley.edu/leaderboard.html BFCL V1: Our initial BFCL release BFCL V2: Our second release, employing enterprise and OSS-contributed live data BFCL V3: Introduces multi-turn and multi-step function calling scenarios Latest Version Release Date…
Organization
Gorilla LLM (UC Berkeley)
gorilla-llm
Large Language Models (LLMs), Teaching LLMs APIs
Models in Library0
Datasets in Library1
Models on Hugging Face17
Followers220