Topic
Servicenow
2 episodes
-
Model Behavior: Week of September 14, 2026
We read this week as a control-layer week: the flashy model race kept moving, but the real competitive shift was toward owning where agents run, what data they can reach, and how enterprises actually deploy them. We still give Anthropic credit for raw capability, but OpenAI, ServiceNow, Nvidia, SSI, and the open-weight wave made the board feel less like a benchmark race and more like a runtime fight.
-
StarHarness: Evolving Harnesses with Stratified Search for Enterprise Environments
Edmund and Geffen dig into StarHarness, a ServiceNow and Mila paper that evolves agent harnesses — prompts, tool interfaces, skills, subagent structure — around a frozen model to close the gap between what an LLM can do and what a messy enterprise environment actually needs. Twenty to thirty-five percentage point gains across three benchmarks, and the harness transfers across GPT and Qwen model families without re-running the search.