Product Strategy · Opportunity Map
New Lanes · Insider Focus Excluded

Beyond Insider Focus Insider-focus (the donor / person product) already sells great — so we set it aside and asked the query log a different question: what else are people repeatedly mining, and which of it could we build into a product now that inferno + OpenRouter do extraction at scale?

Source  users.q8_logs · 680k rows Scope  77,670 topical queries Excluded  ifocus · person-lookup · donor Window  2020 → 2026-07

01 The question, and the method

Strip out every insider-focus query and two-thirds of what remains is one client's accounting sweeps — but the leftover third, read by document type and buyer, exposes lanes with something the insider product's rivals lack: several independent customer segments already pulling the same filings by hand.

Already oursThe searchsec4 · kitems3 · agree2 full-text, proximity, facets — live today
+
Newly oursThe extractioninferno + OpenRouter code hits into structured rows
The productA finished datasetwhat buyers pay analysts to hand-assemble today
77.7k
topical queries once insider focus is removed
4
candidate product lanes surface from the residue
15k+
person-lookups filtered out (pcik/person) — the existing product
9+
distinct law / search / academic buyers behind the new lanes

02 The candidate products — RANKED BY BUYER BREADTH

01

Governance & Proxy Intelligence

Broad demandBuild: Med

A structured proxy dataset — director slates & election outcomes, say-on-pay results, shareholder & activist proposals, compensation packages — extracted from DEF 14A and kitems3 sections. The single strongest new signal in the residue.

Demand evidence
DEF 14A · top non-restatement form (1,516)feed: Proxy Stmts (807)PX14A6G · exempt solicitationsquery: ratify / "NASDAQ Rule 8635(d)"
Who's already pulling it
Kutak RockNixon PeabodyTarter KrinskyBuchananSaul EwingSpencer Stuart (exec search)Royal Goldacademics
02

SEC Enforcement & Regulatory Intelligence

Broad demandBuild: Low–MedNo client conflict

The cleanest net-new corpus: a family of enforcement/regulatory feeds nobody is productizing — coded into an action feed (type, respondent, allegation, resolution) and a comment-letter topic tracker.

Demand evidence
AAER · enforcement releasesLR · litigation releasesComment Letters (99)noact4 · no-action (177)ALJ / APR
Who buys it
securities-litigation firmscompliance teamsacademics (AAER data is heavily cited)

A proven academic market that competes with none of our current clients.

03

Material-Contract & Clause Mining

Focused demandBuild: Med–High

Clause-level search & extraction over the agreements corpus — change-of-control, indemnification, MAC clauses, precedent language — for transactional and M&A practices.

Demand evidence
agree2 corpus (2,358 topical)EX-10 exhibitscompany / counterparty lookups
Who buys it
transactional law firmsM&A / deal teams

Higher build effort — clause extraction is the hard part, and the natural place to prove inferno on nuanced text.

04

Accounting-Event Datasets — sold beyond Ideagen

Known demandChannel conflict

Restatements, insider pledging, changes-in-estimate, cyber (8-K 1.05), auditor changes — the dominant term clusters. Not new to our accounting reseller, but the same academics buying enforcement data want these as standalone feeds. Weigh against competing with a paying client (see the "Compete" sheet).

Demand evidence
restatement / out-of-period (~20k)pledge / margin / collateral (~7k)change-in-estimate (~11k)cyber attack (4k)

03 Prioritization

LaneBuyer breadthBuild effortClient conflictVerdict
Governance & Proxy Medium None Pursue
Enforcement & Regulatory Low–Med None Pursue
Contract / Clause Mining Med–High None Watch
Accounting datasets (broadened) Medium Ideagen Guard

04 Recommendation

Two to move on

Governance & Proxy and Enforcement & Regulatory are the picks: broadest, most diverse buyer bases, cleanly separable corpora we already index, and zero conflict with existing clients. Enforcement is the lowest-effort first build and lands in a proven academic market; Governance/Proxy has the widest commercial pull (law + exec-search + activists).

Next concrete step: point inferno + OpenRouter at one — pull the real query shapes, run an extraction pass over the matched sections, and grade precision against a hand-checked sample. One pilot proves the "search → dataset" machine, then it repeats across lanes.

Read the counts right

Raw term volume is dominated by one client's accounting sweeps, so it understates the new lanes. Buyer breadth — how many independent segments pull a document type — is the truer demand signal here, which is why the ranking leads with it, not with volume.