SimileAI is hiring for evaluation infrastructure
skairam · x · 2026-07-20
This is a hiring follow-up for an Evaluations Engineer at SimileAI.
The role is aimed at strong backend, data, ML-infra, research-tooling, or systems engineers who can own the end-to-end evaluation stack. Prior eval-infra experience is helpful but not required.
The post emphasizes working closely with modeling and research teams, and the attached context describes the infrastructure to be built: execution across models and datasets, provenance/versioning, customer-validation automation, expert review, and comparison/regression tooling.
Related event: SimileAI Treats Model Evaluation as a Systemic Challenge(4 posts)→
More from coding & agent
- GitHub review bot hits its PR limit and forces a 39-minute cooldown — DanielLockyer · 2026-07-22
- Max reasoning effort appears to be mobile-only in Codex Remote, not desktop — GabGarrett · 2026-07-22
- A Reddit demo argues online stores should expose carts and pricing through MCP — gelembjuk · 2026-07-22
- Open-source AI SDK provider routes Vercel apps through a local Codex subscription — lgrammel · 2026-07-22
- Codex vs Claude Code: Which Is More Popular? — jxnlco · 2026-07-22
- CodeRabbit uses layers, diagrams and a chat agent to rethink code review — _jaydeepkarale · 2026-07-22