Crash-testing ChatGPT plugin discovery with 3,000+ queries: small catalog, rarely invoked

Alpic-ai · reddit · 2026-10-01

Alpic ran over 3,000 queries against OpenAI's newly announced plugin search to stress-test it.

Findings: the catalog is small and skewed toward business software, and models rarely consult it unless the user explicitly asks. Their article details the methodology and implications.

Original post →

More from coding & agent

coding & agent channel →