Model comparison
Same task, different trade-offs.
We are not looking for an absolute winner. The comparison shows which model better suits your work, budget and required speed.
Models being compared
ThinkingCap-Qwen3.8-27B
Data from September 27, 2026
Release date unverified
A modification of Qwen3.8-27B focused on shorter internal task processing.
Gemini 3.8 Flash
Data from September 27, 2026
Published September 2, 2026
Processes text, images, audio and video; the output is text.
Price, speed and usage
Self-hosting costs. Commercial licensing under the terms of BottleCap AI; no uniform public price is listed.
$0.75 input / $3.75 output per million tokens through December 31, 2026. From January 1, 2027: $1.50 / $7.50.
We have not independently measured speed on a comparable basis.
We have not independently measured speed on a comparable basis.
Hugging Face after accepting the access terms. PolyForm Small Business 1.0.0 license and additional permission for personal use.
Gemini API and Google AI Studio
Not independently verified in this profile.
1,048,576 input tokens; output capped at 65,536 tokens.
Text and images → text
Text, images, video, audio and PDF → text
ThinkingCap-Qwen3.8-27B
Use when…
- Testing the model on your own hardware
- Token usage comparison with the default model
Choose another model when…
- This is not an unrestricted Apache 2.0 license; check the terms for your use.
- Lower token usage does not mean higher accuracy on every task.
Gemini 3.8 Flash
Use when…
- Analysis of documents, audio and video
- Programming and multi-step business tasks
Choose another model when…
- A different model is needed to generate images or audio directly.
Original sources