Testing Gemini 3.1 Pro on Identifying Judo Throws

Hour-Wish8158 · reddit · 2026-09-01

The author is working on a project to benchmark how Vision-Language Models (VLMs) perform at classifying grappling techniques, specifically Judo throws. Initial results with vanilla (unfine-tuned) models are hit-or-miss. The author believes accuracy would improve significantly with fine-tuning and sufficient data. They are open to open-sourcing the tool for fellow grapplers and engineers interested in experimenting.

Original post →

More from Models

Models channel →