Compare two or more models against a reference result or each other.
Import of 65-marker blood work report tested with native Qwen3-6-27b ) vs. BottleCapAI’s fine-tuned Thinkingcap model.
Thinkingcap was 36% faster with the same results.
You can also compare cloud providers and see how much you are paying for what speeds. Some dare to charge much more and are much slower as well 😎
