models · stepfun-ai

Step-3.7-Flash

image-text-to-text model harvested from Hugging Face (trending). Catalogued at Quarantine — unverified, pending attestation.

Compare Verify

What this registry has checked

Quarantine. We hold the publisher’s own published hash for these weights and have recorded it against this entry. We have not downloaded and hashed the weights ourselves, so the hash is the publisher’s claim, not our measurement.
Methodhf-published-sha256

What is recorded

Publisherstepfun-ai
OriginChina
Licenceapache-2.0
Parameters201,365,316,160 (201.4B)
Context window262,144 tokens
Modalitytext+image+video->text
Native precisionBF16
Pulls recorded here50,207
Downloads reported upstream19,350

Reported scores

GPQA Diamond80.9
IFBench67.3
Humanity's Last Exam21.4

Reported by a third party and recorded here, not re-run by us. Source: artificialanalysis.ai/api/v2.

Measured on real hardware

None of these runs are ours. Each figure is the reporter’s, linked to their own evidence, under their own trust label. The number shown is single-stream decode — one user, concurrency 1 — because aggregate throughput across many users is a different quantity and reads as far faster than anyone will see.

Reported to local-ai-registry

HardwareEngineBuild measuredDecode, 1 streamLabelReported by
Apple M5 Max 128GBllama.cppGGUF IQ3_XXS60.3 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppGGUF iq4xs59.6 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppGGUF IQ4_XS59.6 tok/s @ 8,256candidateexo-postgres
Apple M3 Ultra 96GB 80-core GPUllama.cppGGUF IQ4_XS59.0 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppGGUF IQ3_XXS58.2 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppQ4_K_S58.2 tok/s @ 8,256candidateexo-postgres
Apple M3 Ultra 96GB 80-core GPUllama.cppGGUF IQ3_XXS57.7 tok/s @ 8,256candidateexo-postgres
Apple M3 Ultra 96GB 80-core GPUllama.cppQ4_K_S57.7 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppGGUF UD-IQ2_XXS56.5 tok/s @ 8,256candidateexo-postgres
Apple M5 Max 128GBllama.cppGGUF Q3_K_M56.3 tok/s @ 8,256candidateexo-postgres

From local-ai-registry (MIT), commit 124959d: 44 runs across 8 machines, 0 validated — model revision and runtime pinned, launch accepted — and 44 candidate, their label for useful evidence without a reproducible-launch promise. 28 more not listed.

Measured by local.ai

HardwareBuild measuredDecode @ 8KMemory @ 8KDecode mode
NVIDIA RTX PRO 6000 Blackwell 96GBIQ3_XXS158 tok/s47.5 GBordinary
NVIDIA RTX PRO 6000 Blackwell 96GBUD-IQ2_XXS144 tok/s75.6 GBordinary
NVIDIA RTX PRO 6000 Blackwell 96GBUD-IQ2_M142 tok/s75.7 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUIQ3_XXS66.5 tok/s48.2 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUQ3_K_M62.5 tok/s106.0 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUUD-IQ2_XXS61.9 tok/s74.6 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUUD-IQ2_M60.1 tok/s74.7 GBordinary
MacBook Pro M5 Max 128GB · 40-core GPUIQ4_XS59.8 tok/s117.2 GBordinary

Measured by local.ai, not by us: 33 runs, 25 more not listed. Each names its engine pinned by image digest and the exact serve command, in the full record.

Elsewhere in the registry

Registry record as of 2026-09-01. Identity, licence and parameter count are properties of a release and do not move; tier can, and the verify link above re-checks it against the live registry rather than this page.