HuggingHack adds S3, MinIO, Ollama and vLLM dispatch in a self-hosted layer
TyedalWaves · reddit · 2026-07-27
The author shares major updates to HuggingHack, a self-hosted control layer for Hugging Face, NAS/object storage, and machines running Ollama or vLLM.
New features include local accounts, private or shared model repos, resumable browser uploads, S3/MinIO/Ceph support with local NAS caching, model restore/eviction, dispatch to remote Ollama or vLLM rigs, GGUF inspection, optional PostgreSQL, and stronger security. The project still intentionally avoids being a chat UI or a model-serving replacement.
More from coding & agent
- A new workflow converts ChatGPT web sessions into local Codex sessions — georgemillo · 2026-07-27
- AI still can’t one-shot real SaaS, says builder who starts with data model first — doooyle · 2026-07-27
- Open-source profiler tracks every STT, LLM, and TTS call in self-hosted voice agents — mahimairaja · 2026-07-27
- GPT-5.6 can use about 8× more tokens than 5.5 because its harness splits batch jobs into repeated tool calls — breath_mirror · 2026-07-27
- User tries to pair external audio with an LTX LoRA video workflow — itchplease · 2026-07-27
- A tiny ESP32 board becomes a tongue-in-cheek “coding agent” setup — threepointone · 2026-07-27