What changed in the chat experience and how to get the best out of it.
The assistant now follows 12 imperatives on every analytical question:
Simple lookups ("what's the address of Acme?") still get a 1–2 sentence prose answer — the structured TL;DR / Details / Recommendation shape kicks in for analytical questions only.
After every substantive assistant message, 3–4 chips appear below the response. Each chip is a one-tap follow-up that drills deeper, drafts outreach, schedules a follow-up, or compares alternatives. Chips never appear on simple confirmations or clarifying questions.
When you open the chat panel on a CRM detail page (a company, contact, or deal), the opener chips reflect the entity's actual state — stall time, recent activity, last contact — instead of generic “summarize this”.
Generation is cached per entity for 5 minutes, so revisiting the same page in a short window does not re-run the AI call. Each cache miss is metered against your AI budget alongside other AI features — see Usage Limits for how this counts toward your monthly spend.
Laureo automatically picks how hard to think about each message. A quick lookup (“what's Acme's phone number?”) is answered by a fast, low-cost model, while a harder, multi-step question (“which of my deals are most at risk and what should I do this week?”) is routed to a stronger model. A short follow-up after a complex answer (“and the second one?”) drops back down to the cheaper model so you're not billed for heavy reasoning on a trivial turn.
You don't configure this for chat — it happens behind the scenes on every message and counts toward your AI budget the same as any other AI feature. For background AI Agents, you can pin the model quality yourself — see Model quality for AI Agents.
Preferences saved via /remember or the AI Preferences settings page apply on every future chat. See the AI Preferences doc for details.