Skip to content

Models & capabilities, read from an operations desk

Everything we have published under Models & capabilities, read from an operations desk: what it changes for a B2B company of 10-200 people.

  1. Latest

    OpenAI publishes GPT-5.6 builder guide with new tool-calling and context specs

    OpenAI released a builder guide for GPT-5.6 outlining updated tool-calling behavior, context handling, and recommended architecture patterns. For B2B operators running or evaluating AI agents in sales, support, or ops, this changes how reliably those agents call external systems (CRMs, ticketing, ERPs) and how much conversation history they can retain before context is dropped.

    What changes for operators — If you have an AI agent handling inbound support tickets, qualifying leads, or triaging ops requests, the guide's tool-calling recommendations matter more than the model's raw benchmark scores: unreliable function calls mean an agent that silently fails to update a CRM record or escalate a ticket, and nobody notices until a customer complains. Teams running 10-200 person operations should re-test any GPT-5.6-based agent against their actual tool schemas (not just chat prompts) before treating it as a drop-in upgrade, and check whether prompt or workflow changes recommended in the guide require updating existing automation logic to avoid regressions in accuracy or latency.

  1. OpenAI Ships GPT-5.6, Pitches It as Cheaper Per Task Than GPT-5

    If your sales or support automation runs on OpenAI's API — lead qualification bots, ticket triage, call summarization, CRM enrichment — this release is worth a look purely on cost grounds. Price-performance improvements in a new model version typically translate into lower per-call spend or faster throughput at the same spend, which matters when you're running thousands of automated interactions a month. The practical move is not to rush to adopt GPT-5.6 blindly, but to have whoever manages your model calls (in-house or your automation vendor) benchmark it against your current model on your actual prompts — support macros, sales scripts, whatever you've built — before switching. Model upgrades sometimes shift output tone or formatting slightly, which can break brittle prompt chains or downstream parsing. Treat this as a scheduled maintenance item: check cost, check quality, then migrate if it holds up.

  2. OpenAI Lays Out Vision for "Abundant Intelligence," Light on Product Specifics

    For a 10-200 person B2B company, this particular post changes nothing operationally this week — there is no new model, API, price, or SDK to evaluate. What it does signal is direction: OpenAI is publicly framing its roadmap around making high-quality AI cheap and ubiquitous, which historically has preceded price drops and capability jumps that make previously uneconomical automation (deeper support triage, multi-step sales research, ops reporting) suddenly viable. The sensible operator response is not to build anything new today, but to keep a running list of manual, judgment-heavy workflows currently deemed "too expensive to automate" — because the cost curve behind this kind of announcement tends to move faster than internal roadmaps expect.

  3. OpenAI Tunes GPT-5.6 Sol's Behavior, Opens Luna to Free ChatGPT Users

    For a 10-200 person B2B company, this is a low-drama update but worth a note to whoever owns your AI tooling stack: if staff use free-tier ChatGPT for drafting emails, summarizing calls, or triaging support tickets, their default model behavior just changed without any action on your part. That's the real risk with consumer AI tools embedded in business workflows — model updates roll out silently and can shift output tone, accuracy, or refusal patterns overnight. If any part of your sales or support process leans on ChatGPT outputs going to customers unreviewed, this is a good prompt to spot-check recent outputs against what you were getting last week, and to confirm whether your team is on a paid tier where model versioning is more predictable.

  1. OpenAI Previews Ultrafast Mode for GPT-5.6, Promising 14x Faster Responses

    For a B2B company running automated support chat, voice agents, or real-time sales qualification bots, latency is often the difference between a tool people actually use and one they abandon mid-task. A 14x speed claim, if it holds up in production and not just cherry-picked demos, could make agentic workflows — the kind that chain multiple model calls together for a single customer interaction — feel instant rather than sluggish. That matters most for voice-based support and live chat handoffs, where every second of "thinking" time costs trust. The caveat: speed previews from model labs frequently ship with caveats around cost multipliers or reduced context windows, so treat this as a signal to watch, not a reason to re-architect anything yet.

Next step

Free AI Diagnostic

Fifteen minutes, no email required. It maps where your work actually goes and ranks what is worth automating first.

Start the free diagnostic

Starts immediately in the browser.

Fee
Free
Length
15 minutes

You keep the ranked list of candidates either way.