6 Million Images Reveal AI Art Ecosystem
rao2z · x · 2026-07-14
This research, "Navigating the Open-Source Model Ecosystem," analyzed **6 million Pixiv AI-tagged images** to understand how creators actually use open-source image models and LoRA in the real world. Key findings include: - The data covers over **22,400** base models and **154,000+** LoRAs - Despite the massive number of models, usage is highly concentrated, with **560 base models generating 80%** of the samples - By 2025, **75%** of images used at least one LoRA, making LoRA the de facto standard - Works using LoRA generally receive more views and bookmarks - Returns diminish after stacking **3 to 5 or more LoRAs**, which can even hurt the artwork's performance - **Character LoRA** popularity is declining, while **style/concept LoRAs** are more favored because new base models can better handle direct character prompting - Creators are cautious about upgrading to new model versions: even 20 weeks after release, only **40%** of images use the latest version, partly because compatibility issues can break existing LoRAs or alter aesthetic styles The authors have also made the dataset public on HuggingFace for others to continue analyzing prompts and metadata.
Related event: Study Analyzes 6M Pixiv AI Images to Reveal Ecosystem(2 posts)→
More from Multimodal
- LTX 2.3 LoRA demo changes a video’s camera angle — CQDSN · 2026-07-21
- Reddit user shares a surreal ChatGPT-generated poster — Creamy-Sundae-9991 · 2026-07-21
- A cinematic SEEDANCE 2 prompt turns an empty sunrise city into a memory-driven video — LudovicCreator · 2026-07-21
- Krea 2 Identity Edit transfers a karate pose from a line sketch without ControlNet — NatalieCrypto · 2026-07-21
- SenseTime unveils U1 Pro and open-sources a 50M-sample vision dataset at WAIC 2026 — 机器之心 · 2026-07-21
- Anatomy of Dynamic AI Images: Subject, Environment, and Camera — GPU_FieldNotes · 2026-07-21