NEAR AI Cloud launches confidential inference on SayGm, both sides run in Intel TDX enclaves
bittingthembits · x · 2026-09-21
NEAR AI Cloud's confidential inference is now live on SayGm's confidential tier. SayGm reaches dozens of models through a single API key and runs its own routing inside an Intel TDX enclave instead of ordinary servers; NEAR AI Cloud runs the model inside a TDX enclave too. With two operators in the path, neither has visibility into the request, delivering end-to-end private inference.
Related event: NEAR AI launches confidential inference on SayGm(3 posts)→
More from Infra
- 99.7% cache hits: engineered DeepSeek Harness with self-hosted GLM-5.3 — burny_tech · 2026-09-21
- Solo dev open-sources 4 systems projects, asks engineers to roast them — Accomplished_Row1433 · 2026-09-21
- Andrew Chen: strong LLMs are far from running on phones, on-device AI faces bandwidth, heat and model-size hurdles — andrewchen · 2026-09-21
- Tobi Lütke: local Dell server runs DeepSeek 4.1 Flash at ~300 tok/s, a billion tokens a month — BLUECOW009 · 2026-09-21
- Running Qwen3.8-27B EXL3 on RTX 3060 + 5060 Ti: 50 tok/s with tensor parallelism and MTP — bring_back_the_v10s · 2026-09-21
- Baseten CEO says token volume grew 40x YoY while revenue grew ~10x in 12 months — rohanpaul_ai · 2026-09-21