Apertus 1.5 is a 70B European foundation model with native image and speech support

AxSaucedo · x · 2026-07-27

Apertus 1.5, a 70B European foundation model, adds native image and speech capabilities

ETH Zurich, EPFL, and the Swiss National Science Foundation released Apertus 1.5, a 70B foundation model from Europe. The post says it includes native image understanding and speech processing.

The chart in the image compares its visual performance on 33 image benchmarks:

The post frames Apertus as part of a broader wave of European foundation model activity.

Original post →

More from Multimodal

Multimodal channel →