AI Model Timeline

Tracking the accelerating release frequency of frontier AI models.

QwQ-32B-Preview

Alibaba·

LLMsreasoningcloud + localEdge / Consumer

Overview

Reasoning-focused model with chain-of-thought capabilities.

Capabilities and innovations

32B ParametersChain-of-thoughtMath reasoningSelf-reflectionDeliberative ReasoningExtended ThinkingSelf-Correction

Benchmarks

BenchmarkScoreSource
GPQA Diamond54.5%unsourced
  • Humanity’s Last Exam (no tools): not reported by the vendor
  • SWE-bench Verified: not reported by the vendor
  • SWE-bench Pro: benchmark did not exist at release
  • MMLU: not reported by the vendor
  • MMLU-Pro: benchmark did not exist at release
  • Terminal-Bench 2.x: benchmark did not exist at release
  • MMMU-Pro: benchmark did not exist at release

Architecture and hardware

Parameters
32B total (Dense)
Estimated VRAM at Q4
~19 GB, Edge / Consumer class

Links

More from Alibaba

Data curated by AI Model Timeline. See the methodology for admission criteria, benchmark eras and source priority.