One question, with and without receipts
Same request on both sides: what did our competitor announce this week, and how should it change our pricing pitch? The left side answers from memory. The right side is allowed to work. Print the receipt and inspect any line of it. Notice what the receipt exposed. The memory answer invented a price cut, because a price cut is the kind of thing competitors announce and prediction fills gaps with the plausible. The real announcement was a distribution deal. Nothing about the memory answer sounded wrong. It was just never connected to the world.The four built-ins
Claude ships with a toolbox that is already switched on. No setup, no configuration, available in the app the day you get an account:Web searchQueries the live web and reads results. Anything after the training data, anything moving fast: this is where fresh comes from.
Web fetchOpens a specific page you point at and reads the whole thing, not a summary of a snippet of it.
Code executionA private computer inside the chat. Claude writes real code and runs it: exact math, data analysis, file crunching. Arithmetic stops being prediction.
File creationBuilds actual files you can download or send to Drive: spreadsheets, decks, documents, PDFs. Output that opens in the tools your colleagues use.
Say the tool’s name
Claude decides on its own when to reach for a tool, and it usually decides well. But on work where correctness matters, do not leave it implicit. Ask for the act, not just the answer:- “Search for this week’s coverage before answering” instead of “what happened this week”
- “Compute this with code and show the code” instead of “what do these numbers work out to”
- “Read the page at this link” instead of pasting a link and hoping
- “Build this as a spreadsheet” instead of accepting a table in chat you will retype anyway
The auditing habit
One behavior separates fluent users from everyone else: before trusting an answer, they glance at what ran. Tool calls are visible in the conversation as they happen. If an answer contains a number, a quote, a date, or a claim about the current world, and no tool ran, treat that piece as a guess wearing a suit. It might be right. It earned nothing.Turn this into a reflex with one question per answer: where did that come from? If the answer traces to a search result, a fetched page, or code output, trust accordingly. If it traces to nothing, verify or ask Claude to verify, which usually just means asking it to use the tool it skipped.
The agentic loop
One tool call answers a question. Chain many of them behind a goal, with the model checking its own results, and you get agentic work.