RL post-training Qwen 27B as a Cypher agent lifts graph-query accuracy 6.2 points for $119

sophiamyang · x · 2026-10-11

Adithya Giridharan RL post-trained Qwen 3.8 27B on Fireworks' serverless training API to answer questions over a 650K-node Neo4j graph by writing Cypher queries.

Key facts:

Original post →

More from coding & agent

coding & agent channel →