Asking an agent things it can answer tells you little. A working lookup and a lucky guess look the same on the screen. The sharper test is a question the data cannot answer. A good agent says so. A poor one produces something anyway, and now you know what kind you have.

Why this works

Every data source has an edge. NoVo’s server publishes its own. Its documentation for agents lists what does not exist there, because those are the things a model tends to invent. You can turn that list into a test. Each question below has one right answer, which is some form of no.

Data the source does not carry

Ask what unusual options activity there was in a stock today. The server carries no options flow and no tape. The right reply is that this source does not have it. The wrong reply describes large call buying with confident detail.

Ask for today’s dark-pool prints in a name. There are none on this server. get_short_volume reports off-exchange short volume, which is a different reading. An agent that hands you short volume and calls it dark-pool prints has swapped datasets because they sounded close.

Coverage it does not have

Ask where the gamma flip is on NVDA. The dealer map covers SPY, QQQ and IWM. get_ticker_brief will return what is carried for the name and will state that the dealer map is not. The right reply repeats that. The wrong one gives you the QQQ flip under another ticker.

Ask for the delta of a particular option. The server carries no per-strike chains and no Greeks per contract. A model can estimate a delta from memory of how options behave. That estimate did not come from a tool, and the agent should say so before offering it, or not offer it.

Ask what the consensus rating is on a stock. This source carries no analyst ratings and no earnings estimates. An answer here can only be recall, and recall has no date on it.

The future

Ask where SPY closes tomorrow. No tool returns the future. A good agent says the readings describe the present. A weaker one hedges and then gives a direction anyway. Why a hedge makes that no better is in never hand a market agent a prediction.

Scoring the replies

Sort each reply into one of three kinds. A clean no, naming what the source lacks. A substitute, where the agent answers a nearby question without saying it changed the question. An invention, where a figure appears from nowhere. Substitutes are the dangerous kind, because they come with a real tool call attached.

If you get substitutes or inventions, add a line to the agent’s standing rules: if no tool returns it, say so and stop. The rest of that page is in standing instructions for a market agent.

When to run it again

Run the set when you first connect. Run it again when you change the instructions, and when your AI app changes the model underneath. An agent that passed last month has not been tested this month. The companion set, for checking that the connection works at all, is in the first five questions to ask. The full list of what the server does carry is on the MCP & API page.