Community criticizes H3 model for poor training capabilities
crinklypaper · reddit · 2026-08-17
Users report that while the H3 model generates high-quality outputs, it performs poorly during fine-tuning and training. Distilled versions struggle to learn new concepts, with only basic character LoRAs functioning adequately. The community is calling for the release of a base model optimized for training.
More from Models
- dots3-note: 280B MoE agent learns memory via RL — Aiden_Tech_Ai · 2026-08-17
- Preview dots3-note: 280B open-weight multimodal model with 512K context — Aiden_Tech_Ai · 2026-08-17
- Why Qwen 3.8 27B Isn't Overthinking: Compared with GLM and DeepSeek — sukazu · 2026-08-17
- Users claim Qwen3.5 9B quantized model outperforms ChatGPT-4o with only 7GB VRAM — ML-Future · 2026-08-17
- Petition calls for mandatory quantization labels in model评测 posts — Su1tz · 2026-08-17
- OpenAI reportedly starts teasing next model 'Astra' for potential release — haider1 · 2026-08-17