Add Tool Calling Evaluation section with BFCLv4 results

#1
Red Hat AI org

Added Tool Calling Evaluation section to the model card with Berkeley Function-Calling Leaderboard v4 (BFCLv4) results comparing this quantized model against the base nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55B-BF16 model. Includes Overall, Single Turn (Non-Live and Live), Multi-Turn, and Agentic accuracy with recovery percentages.

krishnateja95 changed pull request status to merged
This comment has been hidden (marked as Spam)

Sign up or log in to comment