Bengaluru's Kinetic-4B Beats Claude Haiku 4.5 on Tool Calling Accuracy and Speed
Conscious Engines, a Bengaluru-based AI lab, published benchmark results on 1 April 2026 showing its 4-billion-parameter Kinetic-4B model outperformed Anthropic's Claude Haiku 4.5 on a 300-sample tool-calling evaluation sourced from Composio. Kinetic-4B achieved 82.33% accuracy at 1.61 seconds p95 latency, compared to Haiku 4.5's 80% accuracy at 4.02 seconds and OpenAI's 120-billion-parameter GPT-OSS-120B's 76.33% at 7.99 seconds. Notably, Kinetic-4B's failed call rate of 4.67% was roughly half that of Haiku 4.5's 9.67%, which has practical implications for agent reliability and error-handling overhead. The model was fine-tuned from the open Qwen3-4B base using LoRA on a single rented GPU over approximately 4.5 hours, and the adapter has been released publicly on Hugging Face. Analysts caution that the benchmark is self-published and covers only a narrow task, and that general-purpose models like Haiku remain preferable for broader reasoning, coding, and summarisation workloads.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in