CHANNEL
Infra
"Infra" is a topic channel on AGI Hunt, an AI news site updated around the clock every 30 minutes. Coverage: Compute, GPUs and chips, data centers, inference and serving stacks, training costs, supply chain and compute economics.
- AI labs are now talking in gigawatts, and Europe is still planning in megawatts — AymericRoucher · 2026-07-23
- COLMAP merges DEGENSAC into main RANSAC, boosting plane-heavy scenes — ducha_aiki · 2026-07-23
- xyOps says it will auto-file incident tickets with full context when jobs fail — tom_doerr · 2026-07-23
- CameoDB embeds MCP directly in a Rust query engine and handles 80M writes a day — Negative_Ad5847 · 2026-07-23
- A 35B MoE ran fully on a phone and shipped a Snake game in 10 minutes — dai_app · 2026-07-23
- Pi exposes cache behavior as a debate over agent harnesses burning KV caches heats up — mitsuhiko · 2026-07-23
- Developer says Sol burns through usage in hours, asks how teams afford $10K-$20K/month — Dan_Jeffries1 · 2026-07-23
- An agent hit GitHub Actions limits, spun up a VM, and kept the build running — davidcrawshaw · 2026-07-23
- Apple’s Mac roadmap points to a new split: devices for people and hosts for agents — APPSO · 2026-07-23
- Progress to buy Domo’s AI and data platform for $400M in cash — shashib · 2026-07-23
- Open-source coding agent octomind chooses persistent cloud machines over per-session sandboxes — donk8r · 2026-07-23
- Farmers say data centers are killing bees, raising a food-system alarm — eyishazyer · 2026-07-23
- Orchestration and Fault Tolerance Are the Real Challenges in Multi-Agent Systems — njanChe1 · 2026-07-23(2 related)
- AI video dubbing costs about $5–7 per finished minute once lip sync is included — Madmahi25 · 2026-07-23
- Nebius shows its first NVIDIA Vera Rubin NVL72 rack in Finland — demian_ai · 2026-07-23
- A Bittensor subnet launches inference at roughly half the usual price — markjeffrey · 2026-07-23
- Ascend SuperPOD Optimization Boosts DeepSeek-V4 Training MFU to 34.22% — pmttyji · 2026-07-23(2 related)
- Turbopuffer halves queue time after fixing deceptively hard autoscaling — DanielLockyer · 2026-07-23
- For a 1 GW data center, build 2 GW into the grid — anderssandberg · 2026-07-23
- Google is spending $200B+ on cloud and compute, Beff Jezos says — beffjezos · 2026-07-23
- Alibaba Cloud says its Zhenwu M890 supernode now runs Qwen3.8 for inference — zephyr_z9 · 2026-07-23
- Alibaba listing shows a $5,000 container data center in the shopping cart — bronzeagepapi · 2026-07-23
- 20VC maps the AI market debates around Kimi, OpenRouter, Fireworks and Nvidia — 20VC · 2026-07-23
- Open-Sourced DSpark Speculator Boosts Inkling Throughput by 1.89x — ying11231 · 2026-07-23(3 related)
- Neo4j workshop shows how graph shapes can give agents better lakehouse context — AI Engineer · 2026-07-23
- Cerebras-style knowledge base drew 2.4 million views and 15,000 questions in 3 months — femke_plantinga · 2026-07-23
- SenseTime's DaJiangZao Achieves Profitable Domestic Compute Scale, Processing 2.42T Daily Tokens — 智东西 · 2026-07-23
- Builder says NVIDIA DGX Spark feels CPU-starved for agentic swarms in 2026 — prasanna_says · 2026-07-23
- Google may be winning enterprise AI deals by pairing algorithms with cloud infra — huangyun_122 · 2026-07-23
- Optical phase-change memory shows 40 dB loss but a 2D VCSEL array — jwt0625 · 2026-07-23
- GPU racks are stalling on cold-plate and CDU capacity, not chip supply — tengyanAI · 2026-07-23
- Celeris launches a lab to build an LLM with microsecond response times — timshi_ai · 2026-07-23
- Developers Discuss Preventing AI Agents from Burning Budgets — Designer_Power3691 · 2026-07-23(2 related)
- A curated guide to LLM cache management spans KV cache, batching, and decoding — gaganghotra_ · 2026-07-23
- RunPod users get a Chrome extension that notifies and auto-claims saved pods — Particular-Repair895 · 2026-07-23
- MiniMax says MI355X is nearing B200 for serving its 428B multimodal model — salykova_ · 2026-07-23
- Aurora launches an open-source Go gateway for routing and securing LLM traffic — Select-Medicine-9310 · 2026-07-23
- Investors betting against infrastructure spending are missing the current AI cycle — cgarciae88 · 2026-07-23
- TileLang debate says leaving CUDA could cut inference costs with only 1–2% loss — teortaxesTex · 2026-07-23
- Agentic RL Colocation Solution Sparks Discussion — willccbb · 2026-07-23(2 related)
- An AI engineer’s learning list covers caching, routing, RAG, and quantization — EmployerNegative5653 · 2026-07-23
- Huawei GPUs may only have a 3-year life, while NVIDIA chips keep a 5-year depreciation window — teortaxesTex · 2026-07-23
- Intel and AMD Lock in Long-Term China Orders Amid Server CPU Price Hikes — firstadopter · 2026-07-23(2 related)
- Kimi V4 is headed for native multimodality, but Moonshot says compute is the limit — teortaxesTex · 2026-07-23
- Couchbase's Capella iQ shows why the real AI product is the inference control plane — krishnan · 2026-07-23
- OpenAI’s planned Australian data center drops recycled-water cooling as Sydney grid nears capacity — Polymarket · 2026-07-23
- Four RTX 3080s hit 69 tok/s on Qwen3.6-27B for about $2,000 — starkruzr · 2026-07-23
- Free market-data API ships with llms.txt, OpenAPI and an Agent Skill — CaseLivid4116 · 2026-07-23
- YC Proposes Offshore Data Centers to Solve AI Power Bottlenecks — Y Combinator · 2026-07-23(3 related)
- Domestic DF1000 Chip Claimed to Rival Hopper Performance — teortaxesTex · 2026-07-23(3 related)