Tag: AI Agents

  • Notion AI for Customer Success: QBRs, Health Scores, and Account Plans

    Notion AI for Customer Success: QBRs, Health Scores, and Account Plans

    Related on Tygart Media: Notion AI for sales · agency owners · client portals.

    The 60-second version

    CS work is constrained by CSM bandwidth. The bandwidth gets eaten by documentation: QBRs, account plans, health score updates, internal reporting. Custom Agents take that documentation work over so CSMs can spend their time on customer calls. The result is CS teams that cover more accounts at the same headcount or go deeper on the same accounts. Either way, the math improves.

    Four CS-specific agent patterns

    Four cards for content, ops, build, and knowledge work with Claude
    Four CS-specific agent patterns.

    1. The QBR draft agent. Triggered before QBR season. For each account: pulls usage data (via integration), product adoption metrics, support ticket trends, key milestones, prior QBR action items. Drafts the QBR deck content in the team’s template. CSM customizes for the specific customer instead of building from scratch.
    2. The health score maintenance agent. Daily or weekly. Reads usage data, support patterns, engagement signals, NPS responses. Updates each account’s health score in the customer database. Surfaces accounts that dropped a tier in the last week.
    3. The account plan agent. Monthly per account. Reviews account activity, identifies expansion opportunities, surfaces stalled adoption areas, drafts the updated account plan with specific next-quarter goals.
    4. The renewal risk agent. Continuous. Scans accounts approaching renewal. Cross-references health score, recent engagement, support ticket sentiment, and upcoming contract dates. Flags 60-90 days before renewal so CSM has runway to address issues.

    What stays CSM

    Floor versus ceiling cards for commoditized work and human-network premium
    What stays CSM.
    • Customer conversations
    • Expansion negotiations
    • Crisis response when accounts are unhappy
    • The judgment about which accounts deserve which level of investment
    • Reading the customer relationship temperature
      The agent surfaces signals; the CSM interprets them.

    The leverage math

    A typical CSM covers 25-40 accounts. Documentation work consumes 30-40% of their week. Custom Agents take that to 10-15%. The CSM either covers more accounts (50-60) or goes deeper on the same accounts (more strategic, more frequent touch).
    The strategic question: which path matches your business? Higher coverage favors expansion-led businesses. Deeper accounts favor retention-led businesses. Don’t let agents accidentally pick the path for you by default.

    Where CS teams go wrong

    Seven cards naming common AI chatbot failure modes
    Where CS teams go wrong.

    1. Letting agents update health scores autonomously into a “you’re red” customer-facing alert. Health scores have political weight inside the customer’s organization. Auto-flagging customers as red without human review can damage the relationship.
    2. Skipping the QBR review. The agent draft is starting material. The customization for that specific customer is what makes the QBR land. Don’t ship the agent draft as-is.
    3. Trusting renewal risk flags without context. A customer can look “at risk” by the data while being fine in the relationship. CSM context wins. Don’t escalate based on the agent flag alone.

    What to read next

    Notion AI for Sales Teams, Account Research, AI-Native Company Patterns.

  • Notion AI for Legal Ops: Contract Review Triage Without Replacing Counsel

    Notion AI for Legal Ops: Contract Review Triage Without Replacing Counsel

    Related on Tygart Media: Notion AI for HR · Claude for lawyers · Notion Command Center.

    The 60-second version

    Legal ops is constrained by counsel time. Custom Agents change which work counsel actually has to do. Routine NDAs that match the playbook? Triaged and approved. Contracts with non-standard clauses? Flagged with the specific deviations and counsel reviews only those. Vendor compliance trackers? Auto-updated. Meeting briefings? Drafted. Counsel reviews exceptions; agents handle volume. The split protects legal quality while massively expanding throughput.

    Four legal-ops-specific agent patterns

    Four cards for content, ops, build, and knowledge work with Claude
    Four legal-ops-specific agent patterns.

    1. The NDA triage agent. New NDA arrives. Agent compares it against the playbook (standard mutual NDA terms, acceptable carveouts, dealbreakers). Classifies as GREEN (auto-approve), YELLOW (counsel review), or RED (substantive negotiation). For GREEN, drafts the response. For YELLOW/RED, prepares a deviation report.
    2. The contract review preparation agent. Triggered for any contract not handled by NDA triage. Reads the contract, compares against playbook, marks every deviation, and produces a redline-ready summary for counsel. Counsel opens the document and starts reviewing the deviations directly, not the entire contract.
    3. The vendor compliance tracker. Maintains a database of vendor agreements, renewal dates, surviving obligations, and required documents (DPA, BAA, COI). Flags upcoming renewals 60 days out and missing documentation continuously.
    4. The meeting brief agent. Before any contract negotiation or compliance meeting, pulls relevant context: prior agreements with the counterparty, related correspondence, current playbook positions on the topics expected. Counsel walks in prepped without the prep work.

    What absolutely stays counsel

    Floor versus ceiling cards for commoditized work and human-network premium
    What absolutely stays counsel.

    The non-negotiable boundaries:
    – Legal advice (period — agents never deliver this)
    – Substantive contract negotiation strategy
    – Risk assessment on novel issues
    – Anything that gets sent to opposing counsel as the firm’s position
    – Privileged communications
    Agents prepare the inputs to counsel’s judgment. They never replace the judgment.

    The triage discipline

    Five security domains: identity, data, code governance, audit, agents
    The triage discipline.

    The triage agent only works if the playbook is explicit. “Standard NDA” is not a playbook; “12-month confidentiality, mutual obligation, no non-solicit, US jurisdiction acceptable, EU DPA required if data crosses border” is. The discipline of writing the playbook is what makes the agent reliable.
    Most legal ops teams underestimate how much playbook documentation they need. The first 90 days of a legal-ops agent rollout is mostly playbook work, not agent building.

    Where this goes wrong

    1. Treating the agent’s classification as final. GREEN means “agent thinks this matches playbook.” It doesn’t mean “approved without review.” A spot-check on 10% of GREEN classifications keeps the system honest.
    2. Letting the agent draft anything that goes to opposing counsel. Even a “thank you, attached is our standard NDA” response should have counsel eyes before send for high-stakes counterparties.
    3. Building too aggressive a YELLOW threshold. If too much routes to counsel, the agent isn’t saving time. Tighten YELLOW criteria. If too little routes, the agent is missing things — loosen YELLOW.

    What to read next

    Notion AI for Operations Managers, Notion AI for Finance, Vendor Check, Editorial Surface Area.

  • The Solo Operator’s Notion AI Stack: Running Multiple Businesses With One Agent Team

    The Solo Operator’s Notion AI Stack: Running Multiple Businesses With One Agent Team

    Related on Tygart Media: Notion AI for agency owners · AI operator’s stack · Notion Command Center.

    The 60-second version

    Running multiple businesses solo used to mean either hiring an assistant or accepting that things slipped through. Custom Agents change the math. A small agent team — three to seven specialized agents — handles the operational layer across all businesses simultaneously, leaving the operator to focus on relationships, strategy, and exception work. The cost is real (post-May 4, somewhere between a coffee budget and a low-end consultant invoice per month) but the leverage is dramatic. The skill isn’t building agents. It’s deciding what to delegate to them.

    The starter loadout

    Four cards for content, ops, build, and knowledge work with Claude
    The starter loadout.

    Seven agents that earn their keep for a multi-business solo operator:
    1. The morning briefing agent. Runs at 6 AM. Reads overnight emails, calendar for the day, project status changes across all businesses. Drops a one-page digest in your daily notes. You read it with coffee.
    2. The intake triage agent. Triggers on new inbound (form submissions, sales leads, partnership inquiries). Categorizes by business, urgency, and type. Drafts a first response. Routes for review.
    3. The calendar prep agent. Runs 30 minutes before each meeting. Pulls relevant project context, prior meeting notes, action items, and any open threads. Briefing arrives in your inbox before the meeting.
    4. The weekly status agent. Runs Friday 4 PM. For each business, summarizes what happened, what shipped, what’s at risk. Output: one digest per business plus a meta-digest across all of them.
    5. The follow-up watcher. Runs daily. Scans all open conversations, projects, and commitments. Flags anything that’s been waiting on you for more than 48 hours.
    6. The content production agent. Runs on schedule per business. Pulls from a content brief database, drafts the next piece, drops it in WordPress drafts (via integration) or a Notion review queue.
    7. The end-of-day capture agent. Runs at 6 PM. Prompts you for a quick voice note on what happened. Processes it into structured updates across the relevant business databases.

    What this stack costs

    Rough credit math at \$10/1000 (post-May 4):
    – Morning briefing: 30 days x ~15 credits = ~\$4.50/month
    – Intake triage: 100 triggers x ~5 credits = ~\$5/month
    – Calendar prep: 100 meetings x ~10 credits = ~\$10/month
    – Weekly status: 4 runs x ~50 credits = ~\$2/month
    – Follow-up watcher: 30 days x ~15 credits = ~\$4.50/month
    – Content production: 12 runs x ~80 credits = ~\$9.50/month
    – End-of-day capture: 30 days x ~10 credits = ~\$3/month
    Total: roughly \$38/month. Add Business plan seat fee. Total operating cost for the agent layer: well under what a part-time VA would charge.

    What this stack doesn’t do

    Seven cards naming common AI chatbot failure modes
    What this stack doesn’t do.

    Things that stay manual:
    – Sales conversations and relationship work
    – Strategic decisions across businesses
    – Team conversations (even if “team” is contractors)
    – Anything client-facing where voice matters
    – Creative work where the doing is the point
    The agents handle the operational substrate. You handle the layer above it.

    How to start

    Desk with laptop, checklist notebook, and billing card ready before creating an Anthropic API key
    How to start.

    Don’t build all seven on day one. Build the morning briefing first. Live with it for two weeks. Tighten the prompt. Then build the next one. Sequential beats parallel.

    What to read next

    What Notion AI Agents Are, How Skills Work, Custom Agents vs Basic, ROI Math.

  • Mobile AI in Notion: The Real Test of Whether Agents Are Ready for Daily Use

    Mobile AI in Notion: The Real Test of Whether Agents Are Ready for Daily Use

    Related on Tygart Media: solo operator Notion AI stack · Notion AI for knowledge workers · Notion second brain.

    The 60-second version

    The real test of any AI feature is whether it survives the move to mobile. Notion 3.2 made that move in January 2026 — agents on mobile, full Custom Agent support, the same auto-model selection across Claude, GPT, and Gemini. The honest assessment after a few months in the wild: it works, but mobile AI is best for consumption and quick interaction, not heavy production. Voice input for prompts is a desktop-only feature so far. Mobile is where you check on agent runs, approve drafts, and ask quick questions — not where you set up complex skills or build workflows.

    What works well on mobile

    Four cards for content, ops, build, and knowledge work with Claude
    What works well on mobile.

    Three patterns that genuinely shine on the phone:
    1. Quick agent queries during in-between moments. Walking between meetings, in line for coffee, on a train. “What’s the status of project X” or “summarize this thread for me.” Phone-sized interaction, phone-friendly output.
    2. Approving and editing agent output. Custom Agent runs overnight, drops a draft in your workspace, you wake up, you read on your phone, you tap-edit a few sentences, you send it. The mobile review pattern is solid.
    3. Quick capture into AI-enriched databases. Voice memo or quick note drops into a Notion database, Autofill fills in summary, tags, owner, date. The phone is the input device; the agent is the cleanup crew.

    What’s painful on mobile

    Seven cards naming common AI chatbot failure modes
    What’s painful on mobile.

    Equally important to name:
    Building skills. Notion Skills require defining instructions, scope, and triggers. The mobile UI for this is functional but slow. Build skills on desktop; run them everywhere.
    Long-context work. Mobile screens make it hard to verify whether the AI pulled from the right pages. If the task involves cross-referencing or fact-checking a synthesis, do it on desktop.
    Multi-step debugging. When an agent run goes sideways and you need to trace why, mobile makes it hard to inspect the trail. The fix is rarely on mobile.
    Voice input. Currently desktop-only on macOS and Windows. Even on those platforms, voice works only inside AI prompt fields, not for general document dictation. Mobile voice is on the roadmap but unannounced as of April 2026.

    How operators are actually using mobile AI

    Three panels showing one problem, three options, one recommendation
    How operators are actually using mobile AI.

    Patterns that have settled into real use:
    The morning check-in. Open Notion on mobile first thing. Read the overnight Custom Agent digest. Approve, edit, or escalate. Closes the inbox before the day starts.
    The drive-time capture. Voice memo into a quick capture database during a drive. Agent processes it later. The phone is the input; the desktop is where you act on it.
    The travel survival mode. When your only device is your phone for a few days, Notion AI on mobile is enough to keep workflows running. Not optimal, but operational.

    The honest limitation

    Mobile AI is good. Mobile AI isn’t a desktop replacement.
    If you’re trying to make your phone the primary tool for Notion AI work, you’ll feel friction. The screen is the bottleneck — not the AI capability, not the model selection, not the agent. Reading multi-paragraph synthesis on a 6-inch screen is what creates the strain.
    The right mental model: desktop is where you build, mobile is where you maintain. Skills, complex prompts, agent configurations, Worker setup — desktop. Daily interaction, approvals, quick captures, drive-time inputs — mobile.

    What to expect next

    Voice input on mobile is the obvious next shoe to drop. The desktop version exists; extending it to mobile is engineering, not strategy. Reasonable timeline: by end of 2026.
    Beyond voice, the more interesting mobile question is whether Custom Agent triggers can fire from mobile-specific events — location, motion, calendar proximity. Notion hasn’t announced anything here, but the “agent that wakes up when I land at the airport” workflow is a natural mobile pattern.

    What to read next

    Corpus follow-ups: Auto Model Selection (how mobile picks models), Custom Agents foundation piece (mobile inherits all the same Custom Agent capabilities), and the Solo Operator workflow article (the real-world mobile pattern).

  • What Notion AI Agents Actually Are (And What They Aren’t)

    What Notion AI Agents Actually Are (And What They Aren’t)

    Related on Tygart Media: how Notion Skills work · AI autofill databases · Notion second brain setup.

    The 60-second version

    A Notion AI Agent isn’t a chatbot. It’s a worker that lives inside your workspace and acts on it. The base version waits for prompts. The Custom Agent version (Business and Enterprise plans only) runs autonomously — on a schedule, on a trigger, or on demand — and can work across hundreds of pages for up to 20 minutes per task. Skills let you teach an agent your repeated workflows so it can run them on command. Workers (developer preview, April 2026) let agents call code and external APIs. The mental model is “a teammate with workspace access,” not “a smarter search box.”

    Why the distinction matters

    Three stacked layers: chat UI, tools, agent runtime
    Why the agent distinction matters.

    Most coverage treats “Notion AI” as one thing. It isn’t. There are at least four layers, and confusing them leads to operators either underusing or overspending on the platform.
    Layer 1: Notion AI in a doc. This is the inline AI you summon with the space bar or /. It rewrites, summarizes, and drafts inside the page you’re on. It’s a writing assistant. It doesn’t act outside the page.
    Layer 2: AI Autofill on databases. This populates or updates database properties based on row content. Basic Autofill is included on Business and Enterprise plans. Custom Agent Autofill uses Notion Credits for richer reasoning. It’s an enrichment layer, not an agent in the proactive sense.
    Layer 3: Standard Notion Agent. Responds to prompts, can read across the workspace, can edit pages, can integrate with Slack, Calendar, and Mail when those are connected. Reactive — it does what you ask, when you ask.
    Layer 4: Custom Agent. Proactive. Runs on schedule or trigger. Can work autonomously for up to 20 minutes. Can have skills attached. Can call Workers (in developer preview). This is the layer most people mean when they say “agents.” It’s also the layer that requires Business or Enterprise and, after May 3, 2026, consumes Notion Credits.
    If you’re unsure which layer you’re using, you almost certainly aren’t using Layer 4 — and that’s fine for many workflows.

    What agents are good at right now

    Four cards for content, ops, build, and knowledge work with Claude
    What agents are good at right now.

    Three categories where agents earn their keep without much fuss:
    1. Database hygiene. An agent that runs nightly across your CRM database can verify links, flag stale records, summarize new entries into a digest field, and tag uncategorized rows. This is dull, repetitive work and it stops being your problem.
    2. Recurring document production. Weekly status updates, daily standups, meeting prep briefs. Anything where the format is stable and the inputs change. The agent reads the inputs, applies the format, produces the document, and you edit the 10% that needs human judgment.
    3. Cross-source synthesis. With Slack, Calendar, and Mail connected, an agent can answer questions that require pulling from multiple sources. “What did the team agree to in the marketing meeting last week, and what’s still open?” That’s a real query an agent can handle — reading the meeting notes, the Slack thread, the calendar follow-up, and producing a synthesis.

    What agents are not good at yet

    Seven cards naming common AI chatbot failure modes
    What agents are not good at yet.

    Equally important to name the gaps.
    Anything requiring judgment about people. Performance review drafting, hiring decisions, conflict mediation. The agent can summarize and surface; it shouldn’t decide.
    Compliance-sensitive output. Legal language, regulated medical content, financial guidance. An agent draft is fine as input to a human reviewer; it isn’t fine as final output.
    Novel reasoning under uncertainty. Agents do well when the pattern is established. They do worse when the situation has no precedent in your workspace. “Plan our entry into a new market” is a worse agent task than “summarize what we’ve learned about our existing market.”
    Stateful work across long timelines. Agents are getting better at continuity, but for now they’re best at bounded tasks. A 20-minute autonomous run is an upper bound, not a target.

    How to think about which layer you need

    A simple decision tree:
    – Just want help drafting? → Layer 1 (inline Notion AI).
    – Want a database to maintain itself? → Layer 2 (Autofill). Use Custom Agent Autofill only when basic isn’t smart enough.
    – Want to ask questions across your workspace and get pulls and edits? → Layer 3 (standard agent).
    – Want recurring autonomous work on a schedule? → Layer 4 (Custom Agent). Be ready to budget Notion Credits after May 3, 2026.
    Most operators land on a mix of Layers 1, 2, and 3. Layer 4 is for specific recurring workflows where the time savings clear the credit cost.

    What to read next

    If you came here trying to understand what agents are, the natural follow-ups in this corpus are: how Skills work (the way you teach agents repeated workflows), what Custom Agents change (the autonomy line), and the May 3 cliff (when free trials end and credits begin).

  • Notion AI for Finance: Automate Close & Reconciliations

    Notion AI for Finance: Automate Close & Reconciliations

    Anchor fact: Custom Agents can manage close calendars, draft variance commentary, sequence reconciliations, and produce audit-ready documentation — but should never autonomously approve journal entries or sign off on financial statements.

    How does a finance team use Notion AI?

    Finance teams use Custom Agents to manage close calendars, draft variance commentary, surface reconciliation exceptions, and prepare audit documentation. The agents handle the documentation and synthesis layer; humans retain decision authority for journal entries, approvals, and any output that gets signed.

    The 60-second version

    Finance work is 60% documentation and synthesis, 40% judgment. Custom Agents handle the documentation and synthesis layer well. Close calendars, variance narratives, reconciliation status, period-over-period write-ups — agents produce these faster than humans and the audit trail is cleaner. The judgment layer — booking entries, approving reconciliations, signing financial statements — stays human. The split is clean and the leverage is real.

    Four finance-specific agent patterns

    Four cards for content, ops, build, and knowledge work with Claude
    Four finance-specific agent patterns.

    1. The close calendar agent. Manages the month-end close sequence. Reads the close database, identifies dependencies, sequences tasks, surfaces blockers daily. Produces the close standup in three sentences instead of a 30-minute meeting.

    2. The variance commentary agent. Reads actuals vs budget. Decomposes variances into drivers. Drafts narrative commentary in your team’s house format. Human reviews, tightens, signs.

    3. The reconciliation status agent. Reads the reconciliation database. Flags reconciliations that have stalled, items aging beyond threshold, balances that don’t tie. Surfaces priority queue for the controller’s morning review.

    4. The audit prep agent. Pulls evidence packages on demand. Given a control number, assembles the testing workpaper, the sample selections, the evidence references, and the deficiency log. Auditor asks for X; you have it in 15 minutes instead of a week.

    What absolutely stays human

    Floor versus ceiling cards for commoditized work and human-network premium
    What absolutely stays human.

    The lines that don’t move:

    • Booking journal entries (agent drafts, human posts)
    • Approving reconciliations (agent surfaces, human signs)
    • Signing off on financial statements (agent prepares; human owns)
    • Estimates and judgmental accruals (the judgment is the work)
    • Anything that goes to a regulator (period)

    The agents do the work that prepares the human to make these calls faster. They don’t replace the calls themselves.

    The audit posture shift

    Five security domains: identity, data, code governance, audit, agents
    The audit posture shift.

    For SOX-regulated entities, agent audit trails change the conversation with internal and external audit. Every agent action is logged. The reproducibility of evidence packages improves. Sample selections that used to take days assemble in hours. This isn’t theoretical — finance teams running this pattern in 2026 are reducing audit-prep cycle time meaningfully.

    The caveat: audit doesn’t accept “the agent did it” as substantiation. The human review at each gate has to be visible in the trail.

    Where finance teams go wrong

    1. Letting the agent draft commentary without source attribution. Every variance number needs to tie back to an underlying report or pull. Agents that produce commentary without citations are a control weakness.

    2. Skipping period-end re-runs. Agent output reflects the moment it ran. If data changes after the agent drafted commentary, the commentary is stale. Build re-run discipline into the close.

    3. Building one mega-agent for finance. Specialized agents (close, variance, recon, audit) outperform a single agent trying to do everything.

    Agent drafts, human posts. That line doesn’t move.

    Sources

    • Notion 3.3 release notes (February 24, 2026)
    • Tygart Media editorial line

    Continue the journey

    This article is part of the May 3 Cliff Decision journey-pack on Tygart Media. Here’s where to go next:

  • Scale Notion AI Output: Why Quality Gates Beat Volume

    Scale Notion AI Output: Why Quality Gates Beat Volume

    Anchor fact: AI amplifies whatever editorial infrastructure you have. Tighter inputs and clearer gates produce more reliable output at scale than adding more agents or more credits.

    What does “gates before volume” mean for AI workflows?

    Gates before volume is the principle that scaling AI output requires tightening quality controls before increasing throughput. Adding more agent runs without first improving inputs, prompts, and review checkpoints multiplies bad output, not good output.

    The 60-second version

    The temptation when AI starts working is to run more of it. Resist that. The order that works is gates first — the inputs the agent reads, the prompts it uses, the checkpoints that catch bad output — then volume. Operators who skip the gate-tightening phase end up with high-volume slop. Operators who tighten gates first end up with high-volume quality. Same agent, same model, same credits. The difference is the gates.

    What a gate actually is

    Five security domains: identity, data, code governance, audit, agents
    What a gate actually is.

    A gate is any checkpoint where output quality gets verified before it propagates downstream. In a Notion AI workflow, gates exist at five points:

    1. Input gate — the data the agent reads (database hygiene)
    2. Prompt gate — the instructions the agent receives (specificity)
    3. Output gate — the format and quality criteria the agent produces against (rubric)
    4. Review gate — the human checkpoint before downstream use
    5. Distribution gate — what triggers final propagation (publish, send, file)

    Each gate is a place where a small fix prevents large drift. Each missing gate is a place where bad output silently propagates.

    The volume trap

    Seven cards naming common AI chatbot failure modes
    The volume trap.

    Without gates, scaling looks like this: agent runs once, output is mediocre but acceptable. Operator runs it 10× per week. Now there’s 10× the mediocrity. By month three, the operator has built a content factory that produces volume but nobody trusts the output enough to skip review. The “scale” never actually shipped because everything still goes through human eyes anyway.

    With gates, scaling looks like this: tighten input substrate, write specific prompts, define a rubric, set a review checkpoint, then ramp volume. Each piece that ships clears the gates. Trust accrues. Eventually the review gate can be sampled rather than universal. That’s when the scale is real.

    Five gates worth installing this month

    Three panels showing one problem, three options, one recommendation
    Five gates worth installing this month.
    1. A controlled-vocabulary tag system on the databases your agent reads from
    2. A prompt template library so prompts are versioned, not improvised
    3. A quality rubric for the output type (the foundry article uses a 5-dimension rubric — same idea)
    4. A weekly review window where you sample 10% of agent output
    5. A failure log where caught drift gets recorded so prompts can be tightened

    Why this is hard

    Because gates are boring. Volume is exciting. Adding a new Custom Agent feels like progress. Tightening a tag taxonomy feels like procrastination. The operators who win at AI scale are the ones who can stay with the boring work long enough that the volume is actually trustworthy.

    Same agent, same model, same credits. The difference is the gates.

    Sources

    • Tygart Media editorial line
    • Notion 3.3 release notes (February 24, 2026)

    Continue the journey

    This article is part of the May 3 Cliff Decision journey-pack on Tygart Media. Here’s where to go next:

  • Notion Workers for Agents: Code Execution for Builders

    Notion Workers for Agents: Code Execution for Builders

    Anchor fact: Workers for Agents is in developer preview as of April 2026, accessible via the Notion API but not exposed through any consumer-facing UI yet. Workers run server-side JavaScript and TypeScript, sandboxed via Vercel Sandbox, with a 30-second execution timeout, 128MB memory limit, no persistent state, and outbound HTTP restricted to approved domains.

    What is Notion Workers for Agents?

    Workers for Agents is Notion’s code execution environment for AI agents, in developer preview as of April 2026. Workers run server-side JavaScript and TypeScript functions that an agent calls when it needs to compute, query a database, transform data, or call an approved external API. Workers are sandboxed (30-second timeout, 128MB memory, no persistent state) and run on Vercel Sandbox infrastructure.

    The 60-second version

    Workers turn Notion AI from a text layer into a compute layer. Before Workers, Notion AI could read pages and write text. It couldn’t run code, couldn’t transform data, couldn’t reliably call external APIs. With Workers, an agent can offload computational tasks to a sandboxed JavaScript or TypeScript function — running for up to 30 seconds in 128MB of memory, with outbound HTTP restricted to approved domains. It’s the upgrade that makes Notion agents capable of real workflow automation, not just document assistance.

    Why Workers matter

    Three stacked layers: chat UI, tools, agent runtime
    Why Workers matter.

    Three things change when agents can call code:

    1. Real database queries. Before Workers, an agent could read pages but couldn’t reliably do “give me all rows where date is in the next 7 days and owner is unassigned.” With Workers, that’s a one-line query that returns structured data the agent uses in its response.

    2. Approved external API calls. An agent can fetch live exchange rates, look up shipping status, query an internal CRM, or pull from any service exposed through an approved domain. The agent doesn’t make the call directly — it delegates to a Worker that does the call and returns the result.

    3. Multi-step transformation chains. Read CSV → transform → enrich → write back to a database. Each step is a Worker. The agent orchestrates the chain. This is the pattern that lets agents handle real ops workflows that previously required Zapier, n8n, or custom code.

    The technical constraints worth knowing

    Workers are not Lambda. They have intentional limits:

    • 30-second execution timeout. Anything longer needs to be split into smaller Workers or moved off-platform. No long-running batch jobs.
    • 128MB memory limit. Streams and chunked processing only for large data. No loading 500MB CSVs into memory.
    • No persistent state between calls. Each Worker invocation is fresh. State lives in Notion databases or external services, not in the Worker.
    • Outbound HTTP restricted to approved domains. You declare which domains a Worker can reach. This is a security feature, not a limitation to fight.
    • Sandboxed via Vercel Sandbox. Workers run on Vercel’s untrusted-code infrastructure. Performance is solid; cold starts exist.

    What you need to use Workers

    This is not a point-and-click feature. Requirements:

    • A Notion developer account
    • A Notion integration set up
    • Familiarity with the agent configuration format
    • API access — Workers are API-only as of April 2026

    If you’ve never built on the Notion API, Workers aren’t your starting point. Standard agents and skills are. Workers are the next step once those don’t go far enough.

    Three Worker patterns to start with

    Four cards for content, ops, build, and knowledge work with Claude
    Three Worker patterns to start with.

    1. The data-fetch Worker. Agent says “I need the current value of X.” Worker calls an approved external API, parses the response, returns a structured value. Common pattern: looking up live data the agent doesn’t have access to natively.

    2. The transform-and-write Worker. Agent passes structured input to a Worker. Worker reshapes the data — formatting dates, normalizing strings, computing derived fields — and writes the result to a Notion database row. Common pattern: cleaning incoming form submissions before they land in the CRM.

    3. The chain-orchestration Worker. A Worker that calls other Workers in sequence, collecting results and returning a synthesized output. Common pattern: a multi-step intake process where each step needs different logic.

    Why this is the more interesting story than May 3

    The May 3 credit cliff is the news story. Workers are the strategic story. Workers are why credits exist — Notion can’t ship “an agent that calls any code you want and any API you want” on a flat fee. Credits make Workers viable as a product. The pricing news is the boring infrastructure that supports the interesting capability.

    If you’re a developer or an agency building on Notion, Workers reshape what’s possible. A custom Notion deployment for a client used to mean “we set up databases and trained the team.” Now it can mean “we set up databases, trained the team, and built five Workers that handle their specific workflows.”

    What’s still missing

    Seven cards naming common AI chatbot failure modes
    What’s still missing.

    Three gaps in the current developer preview worth tracking:

    • No consumer UI. Workers are API-only. End users can’t build them in the Notion app. This will change.
    • Limited debugging. Errors in Workers surface as agent errors. Better tooling for inspecting Worker execution is on the roadmap.
    • Sandbox boundaries are evolving. Approved domain lists, memory limits, and timeout limits are likely to relax over time. Build with current limits; don’t bet on them staying fixed.

    Workers turn Notion AI from a text layer into a compute layer.

    Sources

    • Notion 3.4 part 2 release notes (April 14, 2026)
    • Vercel blog — How Notion Workers run untrusted code at scale with Vercel Sandbox
    • Notion API documentation — Workers for Agents (developer preview)

    Continue the journey

    This article is part of the May 3 Cliff Decision journey-pack on Tygart Media. Here’s where to go next:

  • When Not to Use a Notion AI Agent (What to Keep Manual)

    When Not to Use a Notion AI Agent (What to Keep Manual)

    Anchor fact: Custom Agents are powerful but inappropriate for tasks involving novel judgment, regulated content, sensitive personnel matters, or work where the cost of being wrong exceeds the cost of doing it manually.

    When should you not use a Notion AI agent?

    Don’t use Notion agents for tasks requiring novel judgment about people, compliance-sensitive output (legal, medical, financial guidance), one-off work that won’t repeat, or any decision where the cost of being wrong is higher than the cost of doing the work manually.

    The 60-second version

    Notion agents are a hammer. Not everything is a nail. The honest list of tasks that should stay manual is longer than most operators want to admit. Performance reviews. Hiring decisions. Compliance-sensitive drafting. Anything that gets sent to a regulator or a lawyer. One-off work. Anything where the value of doing it yourself is the thinking, not the output. The discipline of saying “not this one” is what separates operators who use AI from operators who use AI badly.

    Five categories that stay manual

    Four cards for content, ops, build, and knowledge work with Claude
    Five categories that stay manual.

    1. Decisions about specific humans. Performance reviews, hiring choices, conflict mediation, layoff decisions. The agent can summarize and surface evidence; it shouldn’t draft the decision. The risk isn’t that the output is wrong — it’s that the decision-maker outsources the moral weight of the call. Don’t.

    2. Regulated or compliance-sensitive output. Legal language, medical guidance, financial advice, anything that gets reviewed by a regulator. Use AI to draft inputs to a human reviewer. Never ship the AI output as final.

    3. Novel work without precedent. “Plan our entry into a new market.” “Write our crisis response if X happens.” Agents synthesize from existing patterns. They struggle when the situation has no analog in your workspace.

    4. One-off tasks. Building a Custom Agent for a task you’ll do once is more work than just doing the task. The investment in setup (prompt, scope, rubric, review) only pays back across many repetitions.

    5. Work where doing it is the point. Strategic thinking. Writing meant to clarify your own ideas. Reflection journals. The output isn’t the value; the doing is. AI shortcuts the doing, which destroys the value.

    The dangerous middle category

    Long paper tape measure unrolling across a desk beside a laptop, metaphor for context window length
    The dangerous middle category.

    Worse than tasks that obviously shouldn’t be agent work are tasks that look like agent work but aren’t. Examples:

    • “Draft client emails” — sounds like a clear agent task, but the relationship cost of off-tone email outweighs the time saved
    • “Summarize our team’s wins for the board” — looks easy, but framing matters and an agent’s framing is generic
    • “Write our company values” — agents can produce values; only humans can mean them

    The test: if the value of the output depends on being recognizably yours, agent involvement should be limited to research and drafting, not production.

    How to decide

    Desk with laptop, checklist notebook, and billing card ready before creating an Anthropic API key
    How to decide.

    Three questions before launching a new Custom Agent:

    1. Will I do this task at least 20 times in the next year? (No → don’t build an agent.)
    2. Is the cost of a wrong output bounded? (No → don’t automate it.)
    3. Is the value in the output, not the doing? (No → don’t outsource the doing.)

    If any answer is no, the task stays manual. That’s not a failure of AI. That’s discipline.

    AI shortcuts the doing, which destroys the value.

    Sources

    • Tygart Media editorial line
    • Operator practice notes

    Continue the journey

    This article is part of the May 3 Cliff Decision journey-pack on Tygart Media. Here’s where to go next:

  • Notion Custom Agent ROI: Calculating Your Time Reclaimed

    Notion Custom Agent ROI: Calculating Your Time Reclaimed

    Anchor fact: Notion Custom Agents cost $10 per 1,000 credits starting May 4, 2026. Credits reset monthly with no rollover. Simple agent runs use a handful of credits; complex multi-step runs can use dozens to hundreds.

    How do you calculate ROI on a Notion Custom Agent?

    Multiply the human-equivalent time saved per agent run by the dollar value of that time, subtract the credit cost per run (at $10/1000 credits starting May 4, 2026), then multiply by run frequency. An agent that saves 30 minutes of work per run at $50/hour, costs 5 credits ($0.05) per run, and runs daily produces ~$700/month in net value.

    The 60-second version

    Most operators don’t do the math because the math feels small. It isn’t. A Custom Agent that runs daily and saves 30 minutes of $50-an-hour work produces about $750/month in time savings and costs maybe $1.50 in credits. The ratio is so favorable for the right agents that the real ROI question isn’t whether agents pay back — it’s which agents to retire because the math doesn’t clear. After May 4, the bottom of the agent fleet stops being free. That’s good. That’s how you stop running agents that weren’t earning their keep.

    The simple formula

    Floor versus ceiling cards for commoditized work and human-network premium
    The simple formula.

    For any Custom Agent:

    • Time saved per run (minutes) × frequency (runs per month) × hourly value ($/hour ÷ 60) = monthly value
    • Credits per run × frequency × $0.01 (since $10/1000 = $0.01/credit) = monthly cost
    • Monthly value − monthly cost = net ROI

    Three worked examples:

    Example 1 — The weekly digest agent.
    Saves 45 minutes/run, runs 4×/month, your hourly value is $75. Monthly value: 45 × 4 × ($75/60) = $225. Credits: ~20/run × 4 × $0.01 = $0.80. Net: $224.20/month. Keep it.

    Example 2 — The lead enrichment agent.
    Saves 5 minutes/run, runs 200×/month (every new lead), hourly value $50. Monthly value: 5 × 200 × ($50/60) = $833. Credits: ~3/run × 200 × $0.01 = $6. Net: $827/month. Keep it.

    Example 3 — The exploratory analysis agent.
    Saves 15 minutes/run, runs 2×/month, complex multi-step (~80 credits). Monthly value: 15 × 2 × ($50/60) = $25. Credits: 80 × 2 × $0.01 = $1.60. Net: $23.40/month. Keep it, but barely. If credit cost rises or run complexity grows, retire it.

    Where the math turns negative

    Seven cards naming common AI chatbot failure modes
    Where the math turns negative.

    Three patterns where the ROI math fails:

    1. The fancy agent that runs occasionally. Complex agents cost dozens to hundreds of credits per run. Low frequency means the per-month cost is small but so is the value. Net is small. Better as a manual prompt.
    2. The agent that needs human review on every output. If you review 100% of the output anyway, the time saved is partial. Reduce the apparent monthly value by 40-60%. Many agents stop clearing the bar with that haircut.
    3. The agent that runs but the output isn’t used. This is the silent killer. Credits consumed, no value extracted. The fix is monthly observation: which agent outputs do you actually open?

    The portfolio approach

    Treat your Custom Agents as a portfolio. Three categories:

    • Anchors (top 3-5 agents producing outsized ROI). Protect their credit budget first.
    • Earners (agents producing positive but modest ROI). Watch monthly. Retire if drift.
    • Experiments (agents under evaluation). Cap at 20% of credit budget.

    Anything outside those three categories is waste.

    The monthly review ritual

    Desk with laptop, checklist notebook, and billing card ready before creating an Anthropic API key
    The monthly review ritual.

    Once a month, look at:

    • Credits consumed per agent (Notion’s dashboard will show this)
    • Outputs produced per agent
    • Outputs you actually used per agent
    • Time saved estimate per agent

    The gap between “outputs produced” and “outputs used” is where the budget goes to die. Close that gap or retire the agent.

    Treat your Custom Agents as a portfolio. Anchors, earners, experiments. Anything outside those three is waste.

    Sources

    • Notion Help Center — Custom Agent pricing
    • Notion 3.3 release notes (February 24, 2026)

    Continue the journey

    This article is part of the May 3 Cliff Decision journey-pack on Tygart Media. Here’s where to go next: