You own LLM-powered features end to end, from the prompt to what the editor sees, and you care whether the output actually gets published. You get ideas into working shape quickly instead of waiting for the perfect spec, and you stay with them until they hold up in production.
- Build product-facing features across the stack, with a clear bias toward shipping and iterating.
- Work on our multi-LLM pipeline (Anthropic, OpenAI, Gemini, Fireworks/Kimi) and make deliberate calls on model selection, cost and latency.
- Build eval loops against real editorial output: define the baseline, iterate, and be able to show what measurably improved.
- Own traceability across our LLM systems, meaning per-step tracing and prompt provenance, so a bad output can be traced back to the call that produced it.
- Design human-in-the-loop agent workflows: agents propose, the newsroom decides. Nothing goes live autonomously.
- Prototype fast, harden afterwards, and ship your own work to production, including configuration, feature flags and scheduled jobs.
- Bring your own product ideas, rather than only working through tickets.
