SimileAI is hiring for evaluation infrastructure
skairam · x · 2026-07-20
This is a hiring follow-up for an Evaluations Engineer at SimileAI.
The role is aimed at strong backend, data, ML-infra, research-tooling, or systems engineers who can own the end-to-end evaluation stack. Prior eval-infra experience is helpful but not required.
The post emphasizes working closely with modeling and research teams, and the attached context describes the infrastructure to be built: execution across models and datasets, provenance/versioning, customer-validation automation, expert review, and comparison/regression tooling.
Related event: SimileAI Treats Model Evaluation as a Systemic Challenge(4 posts)→
More from coding & agent
- Chaining dependent MCP tool calls: no rollback, duplicate risk — agentrsdg · 2026-09-11
- DeepMind-led paper makes design docs the source of truth, code disposable — SMART regenerates in 1.5-3h for ~$100 — Roger_M_Taylor · 2026-09-11
- Agent-built classifier labels 192k docs for $0.70 vs $13-26 with frontier LLMs — vanstriendaniel · 2026-09-11
- MathModelAgent gains traction: auto-solves math modeling and writes a submission-ready paper — jihe520 · 2026-09-11
- alphaXiv open-sources OpenResearch to run parallel research agents with any model — alphaXiv · 2026-09-11
- DeskcommCRM: open-source AI sales CRM with native agents and WhatsApp hits 1k stars — melgarafael · 2026-09-11