Classifying 1,018 AI Papers for $4: A Two-Model Pipeline at 256ms Median Latency

nutlope · x · 2026-09-17

Developer Nutlope built a pipeline to classify 1,018 AI research papers for about $4 total: DeepSeek V4 Flash summarized each paper ($3.99 on Together AI), then Jev classified each into one of 24 topics ($0.08 on Typesafe AI), with 256ms median end-to-end latency per paper. Results power a live paper-explorer site. His takeaway: use different models for different workflow stages instead of one model for everything; evals are running before replacing current classifications.

Original post →

More from coding & agent

coding & agent channel →