Perplexity Model Hosted on B200
nvidia · x · 2026-07-10
This Perplexity model is hosted in the US and runs on Nvidia B200 GPUs. The team noted that it will continue to be improved during the research preview phase, with full benchmarks to be released in the coming weeks.
Related event: Perplexity Releases GLM 5.2 Orchestrator Model Preview(7 posts)→
More from Infra
- Tabul AI launches Metal TreeSHAP to speed up Shapley values on Apple silicon — Scobleizer · 2026-07-22
- DeepSeek-V4-Flash tops out at 770 tok/s on one B300 in a vLLM batch test — Moreh · 2026-07-22
- NVIDIA starts shipping 102.4 Tbps Spectrum-6 switches for Vera Rubin AI factories — nvidia · 2026-07-22
- Apple publishes SOC 3 audit reports for Private Cloud Compute — throwfaraway4 · 2026-07-22
- Reddit GPU renters say existing platforms only give you two of three: code, recovery, fair billing — legendpizzasenpai · 2026-07-22
- The Sandboxing Manifesto: Secure Execution Environments for Agents — spirosoik · 2026-07-22