Research Instruments Method Findings Say Hello
Brainwork Independent AI research lab

An AI research lab
that reads
the logs.

Brainwork is the independent research lab run by Misha Manko. It operates an instrumented network of 50 real websites, records every visit from every AI crawler, measures which sources AI answers actually cite, and turns the data into published findings and working instruments.

No surveys of what people think AI does. No guesses. The lab runs on measurement.

50Instrumented sites in the network
200+Days of continuous bot logging
9Industries covered
1Operator. No account managers
access_log · tail -f network: 50 nodes filter: ai-fetchers Live
Observation: every AI fetcher in the network read the RSS feed before it read the homepage. Not one requested llms.txt.
03:12:04GPTBot/1.2/feed.xmlnode-07 · legal200
03:12:09ClaudeBot/1.0/research/schema-markup-that-ai-models-actually-usenode-19 · medical200
03:12:31PerplexityBot/1.0/sitemap.xmlnode-02 · finance200
03:13:02OAI-SearchBot/1.0/pricingnode-33 · saas200
03:13:47CCBot/2.0/node-11 · ecommerce200
03:14:15Claude-SearchBot/1.0/feed.xmlnode-41 · travel200
any fetcher/llms.txt50 nodes · 200+ days0 hits
01
Research programs

Three questions,
measured continuously.

Each program has its own instrument, its own data, and its own publication trail. None of them rely on vendor dashboards or third-party estimates.

Program A

How AI crawlers read the web

50 real websites across 9 industries, every request logged at the edge. Which fetchers visit, what they open first, how often they return, and what they ignore. This is the backbone of everything else the lab publishes.

Since 2025 · Edge logs · 50 nodes
Program B

What the open web index holds

Common Crawl feeds most training corpora. The lab measures how much of a domain each crawl actually captured, when robots.txt started blocking it, how far the stored copy has drifted from the live page, and where the domain sits in the 118-million-node web graph.

Since 2026 · Common Crawl · Web graph
Program C

Which sources AI answers cite

Nightly runs of real buyer prompts through ChatGPT and Google AI Overviews, with every cited source stored raw. The question is not "are we mentioned" but "which pages earn citations, and what do they have in common."

Since 2026 · Nightly · Raw responses kept
02
Instruments

Tools the lab built
because none existed.

Each one started as a measurement problem inside a research program. When the data needed a tool, the lab built it and kept it running.

03
Method

Instrument. Log.
Publish.

Step 01

Instrument

Every claim starts with a sensor. Real sites, real edge logs, real API responses stored raw. If it cannot be measured, the lab does not have an opinion on it yet.

Step 02

Log

Continuously, not as a one-off study. AI fetchers change behaviour month to month. A snapshot from last quarter is already a historical document.

Step 03

Publish

Findings go out in the open with the method attached, including the ones that contradict the industry's favourite advice. The next person should not have to repeat the work.

House rules
  • Measurement before enthusiasm. AI is a tool, sometimes the right one, often not.
  • No guarantees about opaque systems. Anyone promising AI citation rankings is uninformed or lying.
  • Vendor dashboards are inputs to check, never sources of truth.
  • Client data never appears in published work. Findings are aggregated across the network.
  • Tools are built only when a research program needs one. Nothing speculative.
  • One operator, no account managers. You talk to the person who ran the query.
04
Findings

What the logs
said.

Published research from the network, written up on Misha Manko's site. Method and data described in every piece.

Bring the lab
a hard
question.

Contact

Research collaborations, data questions, and consulting on the harder cases. One business day to reply. Productized audits and engagements run through mishamanko.com.