StepFun Releases StepEdge On-Device Models
新智元 · wechat · 2026-07-12
StepFun has released the StepEdge on-device model family, targeting local execution scenarios for devices like smartphones and car infotainment systems. It covers capabilities such as text and visual understanding, voice/audio interaction, screen and GUI operations, and image generation/editing.
The article emphasizes that this is not a single model, but part of a device-cloud collaborative system:
- StepEdge: Text and visual understanding
- StepEdgeAudio: Voice and audio interaction
- StepEdgeGUI: Screen understanding, widget localization, and UI operation
- StepEdgeGen: On-device image generation and editing
The text provides multiple benchmark results, claiming top rankings among its peers in comprehensive benchmarks, OSWorld, and audio understanding metrics. It is paired with the self-developed StepInferenceNPU inference engine, highlighting local low latency, offline availability, and privacy protection. The article positions this as the final puzzle piece in StepFun's comprehensive "Pro + Flash + Edge" model layout.
Related event: StepFun Releases Step Edge On-Device Models(3 posts)→
More from Models
- NVIDIA says Nemotron 3 Ultra scored 30/42 on the 2026 IMO problems — NVIDIAAI · 2026-07-22
- OpenAI is reportedly briefing U.S. lawmakers on its next model family — kimmonismus · 2026-07-22
- Muse Spark 1.1 lands at 1495 on Text Arena with standout agentic-coding price performance — ycombinator · 2026-07-22
- Advanced AI Models Are Becoming Impossible to Plug and Play — emollick · 2026-07-22
- Google Gemini's AI Problem: No Leading Model for Core Workloads — bindureddy · 2026-07-22
- Model Offers 1M Token Context Window at Just $0.33/1M Tokens — MickeySteamboat · 2026-07-22