Presets
Three presets cover most teams. Each card shows the model, its reasoning level, an estimated cost, and our take.
A workspace that has not chosen runs on Smart.
All models
Below the presets, All models lists the full catalog with a relative cost label for each (base rate, cheaper by a percentage, or a multiple of base). The catalog includes models from Anthropic, OpenAI, xAI, and Moonshot; some entries are served by US inference providers and are marked as such.Reasoning levels
Each model accepts a range of reasoning levels: None, Low, Medium, High, Extra high, and Max. Higher levels think longer and cost more per turn. Set the level alongside the model. A level a model does not support is rounded to the nearest one it does.Working style
Working style sets how deeply Henry works on each conversation across Slack, Teams, and the web app:- Quick: fast answers with a small tool budget. Best for simple lookups. Henry says when a question deserves a deeper pass.
- Standard: the balanced default, with enough tool calls and thinking for most day-to-day work.
- Deep: maximum thoroughness. Henry cross-checks, paginates, and uses many tool calls. Slower and costs more per turn.
Fast replies
Add!fast anywhere in a Slack message to answer that one message with a cheaper, lower-reasoning model. Choose which model handles fast replies under Fast replies, or leave it on the default.