Skip to content

Three model cards, three evaluation questions

An initial evaluation shortlist across text, speech, and multimodal tasks. These are selected models, not a ranking or a list of new September releases.

Start with the task

A model card establishes what the publisher describes. Your evaluation establishes how the candidate handles your work.

Source review: 2026-09-26. This issue contains no Sophrono benchmark results.

ModelTaskLicense / termsHardware evaluation
Qwen3-8BText generationApache 2.0Measure memory and latency at the selected precision and context length.
Qwen3-ASR-1.7BSpeech recognitionApache 2.0Benchmark the supported inference configuration on representative audio.
Gemma 3nMultimodal input; text outputGemma Terms of UseDesigned for low-resource devices; validate the chosen configuration.

Read the primary sources

Qwen3-8B

Does the chosen reasoning mode meet the workload quality and latency criteria?

Publisher model card for Qwen3-8B

Qwen3-ASR-1.7B

How does transcription perform on your audio, language, and terminology?

Publisher model card for Qwen3-ASR-1.7B

Gemma 3n

Which input modalities and quality thresholds does your workflow need?

Publisher model card for Gemma 3n

License review

The linked Qwen cards list Apache 2.0. The Gemma card links its Terms of Use. Record the applicable model version and terms before deployment.

All issues · AI Workload Evaluation

Super Intelligence Newsletter

The frontier of AI, once a month.

The models, research, and releases that change what a business can build, with the sources and the question worth testing.

Read recent issues

One email a month. Unsubscribe anytime.

Build the system that compounds.

Book a time