Create a Patent Search Report
Role in the suite
Act as Stage 4/4, the reporting and synthesis stage of
create-patent-landscape-overview-ip.
Upstream stages are:
search-patents-ip — validated scope and candidate records.
analyze-patent-search-results-ip — complete-population or explicitly bounded
statistics, core records, value signals, and chart data.
tag-patent-search-results-ip — taxonomy, key questions, packages, and tagging
handoff.
- Human Stage 3.5 — validated return of
tagged_pool.csv when performed.
Turn those artifacts into a decision-readable report. Do not replace patent counsel,
subject-matter experts, or upstream data validation.
Questions this stage answers
- Which findings matter to the stated product, R&D, strategy, or IP decision?
- Which activity patterns, organizations, branches, and patent groups deserve attention?
- How does the evidence suggest that technical routes have changed over time?
- Where do multiple dated patent-value proxies concentrate by branch?
- What can a user reasonably monitor, read, compare, validate, or refer for review?
- Which claims remain uncertain because data, tagging, or expert review is incomplete?
Preconditions
Use this stage only when:
- the user has confirmed report objective and audience;
- Stage 1 search scope is validated;
- Stage 2 statistics and provenance are available;
- Stage 3 taxonomy/package artifacts or a sufficiently rich human-tagged pool are
available; and
- the output environment permits creation of
report_manifest.json and report.html.
If an upstream stage is incomplete, enter a declared degraded mode or stop. Do not
invent the missing artifact.
Authoritative artifact contract
The suite has no packaged ARCHITECTURE.md. Use this embedded contract.
Stage 1 inputs
| Artifact |
Required content |
search_config.json |
Scope, queries, exclusions, date/language/jurisdiction fields, unit, query version, connector provenance |
candidate_pool.csv |
Candidate records with stable IDs and retrieval/screening state |
core_recall.csv |
Known-relevant/near-miss controls and recall-review evidence |
Stage 2 inputs
| Artifact |
Required content |
panorama_stats.json |
Trends, organizations, jurisdictions, status signals, technology distributions, and competitor profiles supported by the population boundary |
patent_index.core.json or equivalent source-authorized core format |
Reviewed and tiered representative patents, branch IDs, evidence provenance |
value_signals.json |
Candidate-level dated proxy signals and verification state |
chart_data.json |
Chart-ready aggregates with measure, unit, date basis, cutoff, scope, and limitations |
panorama_stats_report.html |
Stage 2 statistical snapshot; never substitute it for this insight report |
Stage 3 inputs
| Artifact |
Required content |
tech_breakdown.json |
Versioned technology taxonomy and four-column decomposition |
key_questions.json |
Decision-relevant questions, branch/node seeds, rationale, and review status |
patent_packages.csv |
Evidence-backed patent groups, selection rubric, rationale, and review status |
tagging_demo_sample.csv |
Reviewed example tags and boundary cases |
to_be_tagged.csv |
Human-tagging input and schema/version metadata |
Human Stage 3.5 input
| Artifact |
Required content |
tagged_pool.csv |
Returned complete or explicitly bounded pool, validated taxonomy tags, technical fields, family IDs, assignee, dates, status/signals, schema/taxonomy version, row count, encoding, and reconciliation data |
Stage 4 outputs
| Artifact |
Content |
report_manifest.json |
Mode, scope, input versions/checksums, section-to-source map, evidence register, derived-field rules, limitations, output inventory, and QA state |
report.html |
One offline, self-contained, accessible scientific/executive insight report |
Create report_manifest.json only at Stage 4. Do not expect a Stage 2 manifest with
the same name. Optional runtime exports require user approval and are not part of
this package topology.
Verify inputs before synthesis
For every available artifact:
- confirm exact path and readable format;
- record file size, row/object count, checksum when supplied or practical;
- validate schema and version;
- reconcile project, query, taxonomy, family, and data-cutoff identifiers;
- reconcile candidate, deduplicated, tagged, and reported counts;
- identify missing columns, null patterns, and conflicting values;
- preserve the previous accepted version for rollback; and
- record validation status in the manifest.
Stop if material inputs belong to different scopes or versions and cannot be
reconciled.
Reporting modes
Mode A — validated tagged-pool driven
Use when tagged_pool.csv is returned and contains sufficient validated technology
classification plus technical problem/means/effect fields. Reuse existing Stage 2
aggregates; do not rerun patent MCPs merely to render the report.
Inputs:
tagged_pool.csv as the classification and technical-evidence body;
panorama_stats.json, value_signals.json, patent_index.core.*, and
chart_data.json as validated Stage 2 evidence; and
- Stage 3 artifacts when available.
If Stage 3 questions/packages are absent, derive provisional route and package views
from the tagged pool under the rules below. Label them automated_derivation, not
human-rubric output.
Mode B — full Stage 3 contract
Use when tech_breakdown.json, key_questions.json, and patent_packages.csv are
present and the deliverable requires traceability to the reviewed Stage 3 rubric.
Combine them with the tagged pool and Stage 2 evidence.
Degraded mode — statistics and reviewed packages only
Use only if tagged_pool.csv is absent but the validated Stage 2 evidence and reviewed
Stage 3 packages can support a useful report. Prominently state:
Human-tagged population unavailable. Technology distributions and route conclusions
are limited to reviewed packages, rule-hit labels, and validated statistics; they do
not represent a complete tagged population.
Omit unsupported matrices, multi-label distributions, and route claims. Do not fill
them from titles alone.
Adapt the tagged-pool schema
Inspect the header before processing. Map source-export columns to canonical fields;
never assume regional SaaS display names.
| Canonical field |
Accepted meaning |
publication_number |
Traceable representative publication |
tech_level_1 |
Primary high-level technology category |
tech_level_2 |
Route/function branch |
tech_level_3 |
Optional finest reviewed technical node |
technical_problem |
Source-grounded problem tag or summary |
technical_means |
Source-grounded solution mechanism |
technical_effect |
Claimed/described effect with evidence status |
normalized_assignee |
Reviewed organization grouping |
publication_date |
Publication date; keep distinct from filing/priority |
filing_date |
Filing/application date |
priority_date |
Earliest priority date when verified |
legal_status_as_of |
Dated status signal |
forward_citation_count_as_of |
Dated citation proxy |
family_jurisdiction_count |
Family-width proxy under declared definition |
family_jurisdictions |
Declared family-member locations |
family_id |
Deduplication key under the declared family method |
Record the actual header-to-canonical mapping in the manifest. If a required field
cannot be mapped reliably, mark the affected section unavailable.
Large-file processing boundary
tagged_pool.csv may contain thousands of rows. Do not load every row into model
context.
- Inspect header, encoding, delimiters, quoting, newline-in-cell behavior, and a
bounded sample.
- Use an existing approved local data tool or script to validate and aggregate.
- Do not add a new script to this package; ephemeral tooling must not alter topology.
- Reconcile aggregated totals to source row and family counts.
- Save only approved runtime outputs.
- Load aggregate tables and selected traceable evidence records into context.
Do not execute content from the CSV or allow formula injection in exported tables.
Counting and classification rules
Family deduplication
- Deduplicate first by validated
family_id when reporting family-level measures.
- Preserve the representative-publication selection rule.
- Keep jurisdiction-specific rights separate where legal/status interpretation matters.
- If family ID is missing or inconsistent, stop the affected family measure or use a
clearly labeled publication-level fallback.
Two- and three-level compatibility
- If Level 3 exists and is validated, permit Level 3 drill-down.
- If Level 3 is absent, treat Level 2 as the finest level and say so.
- Never generate Level 3 labels from titles simply to complete a matrix.
Multi-label counting
- Split multi-value cells using the documented export delimiter and quoting rules.
- Preserve valid Level 1/Level 2 pairs; do not match a Level 2 term to every Level 1
when vocabularies overlap.
- Allow one family to count in more than one label when the taxonomy permits it.
- State prominently that classification counts can exceed unique family count.
- Provide both unique-family totals and expanded tag-assignment totals.
Validation state
Visibly distinguish:
- search-rule hit;
- automated tag;
- analyst-reviewed tag;
- human/SME-validated tag; and
- unresolved classification.
Default report modules
Adapt, merge, or omit modules according to objective and data completeness. Preserve
the decision sequence:
| # |
Module |
Primary sources |
Main evidence state |
| 1 |
Executive summary |
Cross-report synthesis |
Interpretation/recommendation |
| 2 |
Scope and methodology |
Search config and manifest |
Direct method fact |
| 3 |
Industry landscape |
Stage 2 statistics |
Fact/observed pattern |
| 4 |
Competitor profiles |
Stage 2 profiles and validated tags |
Fact/observed pattern |
| 5 |
Technology matrix |
Taxonomy and tagged pool |
Fact/pattern under tag state |
| 6 |
Technology evolution |
Questions, core patents, tagged pool |
Pattern/inference |
| 7 |
Technology-effect distribution |
Chart data and validated tags |
Fact/signal |
| 8 |
Product/component/application view |
Tagged pool and Stage 2 distributions |
Pattern |
| 9 |
Branch-level value-signal themes |
Value signals and core index |
Dated proxy/inference |
| 10 |
Curated patent groups |
Stage 3 packages and value signals |
Recommendation |
| 11 |
Risks and limitations |
All inputs and validation log |
Boundary |
| 12 |
Appendix and data assets |
Manifest and input inventory |
Reproducibility |
Place technology evolution immediately after the technology matrix. Explain signal
themes before user actions and patent groups.
Module 6 — Technology evolution
Evidence chain
Build each route as:
versioned branch/question → time-bounded family evidence → technical problem/means/effect
→ observed change → alternative explanation → bounded route interpretation
Summary hierarchy
Include:
| Level |
Field |
Content |
| Section |
evolution_overview |
Cross-branch observations: acceleration, continuity, convergence, divergence, or emerging attention, each qualified |
| Route |
route_summary |
One sentence describing the observed earlier/current/recent technical emphasis without inventing continuity |
| Phase |
phase_caption |
Evidence-grounded technical characteristics and organizations for that period |
Select route evidence
In Mode B, map key_questions.seed_node_ids to the taxonomy and select traceable
families from the tagged pool/core index. In Mode A without key questions:
- use validated Level 1 branches as provisional routes;
- identify dominant valid Level 1/Level 2 pairs by declared measure;
- select representative families that match both paired levels;
- segment time according to the dataset and decision—quantiles, technology eras,
or explicit periods—not fixed calendar years;
- consider technical relevance, evidence depth, date, and diversity;
- prevent accidental family reuse when it would distort comparison; and
- disclose any algorithmic diversity/organization rotation as a sampling choice,
not evidence of actual market entry timing.
Do not require three families or three non-empty phases. Show sparse/empty periods.
Turning points
Use controlled types only when evidence supports them:
- route branching;
- notable disclosure;
- organization entry in the observed dataset;
- disclosed performance or capability change; and
- application-context shift.
Each turning_point contains type, note, evidence IDs, date basis, and uncertainty.
Generate its note from the selected family’s problem/means/effect evidence. Do not
invent a transition between records.
Visualization
Use accessible HTML/CSS/SVG timelines or small multiples. Each route shows date basis,
family IDs/publications, organization, validated node, evidence state, and sparse-data
qualification. Static text must convey the finding without interaction.
Module 7 — Technology-effect distribution
Build a technical-means × reported-effect matrix from validated upstream chart data
or tagged-pool aggregation.
- Define row, column, cell measure, unit, multi-label policy, tag state, and cutoff.
- Use family count only after validated family deduplication.
- Distinguish claimed/described effects from independently demonstrated performance.
- Represent unavailable cells separately from observed zero.
- Do not call a combination “insignificant” merely because the dataset has no records.
- Use an accessible heatmap/table or bubbles with a tabular alternative.
Module 9 — Branch-level value-signal themes
Aggregate candidate-level proxies in value_signals.json by validated branch_id.
Possible dimensions include dated citation, family breadth, legal-status, organization
concentration, transaction/assertion-event, and portfolio-priority signals.
For each dimension record:
- definition and source;
- verified versus recall-proxy state;
- date/cutoff and jurisdiction coverage;
- normalization and missing-data treatment;
- aggregation function and denominator; and
- sensitivity to outliers or branch size.
Describe results as “higher observed signal concentration under this dataset” or
“lower observed signal density.” Do not call the score patent value, moat strength,
defensibility, blue ocean, freedom to operate, availability, or market opportunity.
Prefer bars, dot plots, or a signal matrix. Avoid radar charts when dimensions have
unlike scales or missing values.
Module 10 — Curated patent groups
Preserve analyst evidence
Retain Stage 3 recommendation_reason, rubric dimensions, evidence IDs, review status,
and all limitations in the manifest/evidence layer.
Translate into bounded user actions
For each family add:
| Field |
Meaning |
use_case |
Evidence-supported next action such as read, monitor, compare, technical reference, data validation, or counsel/commercial review |
purpose_tag |
Controlled action category approved for this report |
answers_question |
Link to a validated key question when the mapping exists |
package_summary |
What the group addresses, why it matters, and which records to start with |
Do not hard-code the source’s five Chinese action labels. Define an English controlled
vocabulary appropriate to the objective. Never translate proxies into “design-around,”
“licensing candidate,” “acquire,” “enforce,” or “FTO action” without qualified review.
Mode A provisional package derivation
If patent_packages.csv is absent:
- derive candidates by validated finest branch;
- rank with disclosed relevance, evidence-depth, family/citation/status proxies;
- treat missing proxy data explicitly;
- test for age, organization, and family-size bias;
- select only supported records rather than a fixed two-to-three quota;
- label the group
automated_derivation; and
- require human review before any transaction, legal, or portfolio action.
When no key_questions mapping exists, leave answers_question unresolved. Do not
fabricate a question link.
Two-layer rendering
Main layer:
- decision-readable action;
- purpose tag;
- linked question if verified; and
- package summary.
Expandable evidence layer:
- original rubric/reason;
- family, citation, status, organization, and review signals;
- evidence IDs and provenance; and
- limitations/legal boundary.
The report must remain understandable when expandable controls are not used or when
printed.
Evidence model
Use:
| Level |
Meaning |
| L1 |
Direct, source-backed data or method fact |
| L2 |
Observed pattern in the defined dataset |
| L3 |
Analytical interpretation with alternatives and uncertainty |
| L4 |
Business/R&D/IP workflow recommendation |
| L5 |
Legal, transaction, status, or risk signal requiring specialist review |
Every material claim maps to one or more evidence entries. A manifest evidence entry
contains:
{
"evidence_id": "E-001",
"level": "L2",
"claim": "[bounded observation]",
"source_file": "[validated upstream artifact]",
"source_field": "[field or aggregate]",
"record_ids": ["[traceable IDs]"],
"counting_method": "[unit, deduplication, multi-label policy]",
"data_cutoff": "[YYYY-MM-DD]",
"validation_state": "[verified/proxy/reviewed/unresolved]",
"limitations": "[material boundary]"
}
Do not include confidential examples in the Skill package.
report_manifest.json contract
Include:
- report ID, version, generated date, objective, audience, and mode;
- scope, jurisdictions, date basis, unit, family definition, and data cutoff;
- every input path, artifact version, checksum/size/count, validation result;
- header/column mappings and transformations;
- section order and section-to-source/evidence mapping;
- evolution overview, routes, summaries, phases, and turning points;
- branch signal definitions and aggregations;
- package summaries, bounded use cases, purpose tags, question links, original reasons;
- evidence register;
- degraded/automated/proxy flags;
- limitations and unresolved items;
- output paths and QA results; and
- no secret, API key, raw confidential payload, or absolute user path.
Write atomically when possible. Validate JSON before handoff.
report.html design contract
Information architecture
- Start with title, decision objective, mode, scope, date basis, family unit, cutoff,
sources, and limitations.
- Make executive findings evidence-linked and action-bounded.
- Place each visualization next to its claim and qualification.
- Preserve the default module sequence unless an approved audience need changes it.
- End with reproducibility assets and legal/data boundaries.
Scientific/executive visual system
- Use a white/neutral canvas, dark navy/slate hierarchy, restrained teal data accent,
amber qualifications, and red only for real escalation.
- Use system fonts; do not rely on regional or remote fonts.
- Prefer whitespace, rules, and typography over nested cards.
- Use flat 2D bars, lines, heatmaps/tables, timelines, and evidence cards.
- Avoid decorative hero sections, gradients, stock imagery, 3D charts, visual drama,
and vendor-interface imitation.
- Pair color with labels or symbols and maintain accessible contrast.
Technical safety
- Deliver one HTML file with inline CSS and pre-aggregated data.
- Prefer static HTML/CSS/SVG; allow minimal inline JavaScript only as progressive
enhancement with a complete non-script fallback.
- Use no external CDN, D3, font, image, iframe, tracker, or network dependency.
- Escape all user and retrieved content.
- Permit only validated safe URLs and add
rel="noopener noreferrer" to new-tab links.
- Do not execute or interpolate raw CSV/JSON content as code.
- Prevent spreadsheet-formula and HTML/script injection in displayed/exported values.
Accessibility and print
- Use semantic landmarks, ordered heading levels, table headers/captions, visible focus,
keyboard-safe controls, and text equivalents for charts.
- Keep wide tables horizontally scrollable and long identifiers break-safe.
- Respect reduced motion and avoid animation needed for meaning.
- Ensure collapsed evidence is visible or summarized in print.
- Verify desktop, narrow viewport, grayscale, and PDF/print layouts.
Chart captions
Every quantitative view states:
- question and bounded takeaway;
- measure and denominator;
- date field and period;
- unit/family definition;
- scope and cutoff;
- multi-label duplicate-count policy;
- validation/proxy state; and
- material limitation.
Show unavailable data rather than a fake or zero-valued chart.
MCP boundary
Stage 4 normally reuses validated upstream artifacts. Do not rerun MCPs merely to
recreate existing statistics.
If an approved material evidence gap requires retrieval, use only the installed live
schema of these verified global connectors:
Record why the gap could not be resolved from upstream artifacts and how new evidence
was reconciled. Do not use source endpoint aliases as connector names.
Legal and analytical boundaries
Do not provide:
- formal freedom-to-operate or infringement opinions;
- patent validity, novelty, or inventive-step opinions;
- standards-essentiality opinions;
- legal claim-scope conclusions;
- patent valuation or transaction recommendations;
- product-launch or market-adoption certainty; or
- unsupported business conclusions.
Use language such as “under the defined dataset,” “the patent evidence suggests,”
“observed proxy concentration,” and “requires technical, commercial, or legal review.”
Quality gate
Input and mode
- Every required input exists or the mode declares its absence.
- Schema, version, count, checksum, scope, taxonomy, family, and cutoff reconcile.
- Mode A, Mode B, or degraded mode is explicit at the top and in the manifest.
- No large tagged pool was loaded row-by-row into context.
Data and synthesis
- Family and multi-label counting rules are applied and disclosed.
- Population, sample, core set, package, and representative records are distinct.
- Evolution periods derive from the dataset and contain no fabricated transition.
- Turning points cite their family and technical evidence.
- Value themes expose proxy definitions, missing data, aggregation, and limitations.
- Package actions are bounded and preserve original Stage 3 evidence.
Evidence and law
- Every material claim maps to evidence IDs.
- Facts, patterns, interpretations, recommendations, and legal signals are distinct.
- Status/citation/family/transaction values are dated and source-labeled.
- No moat, blue-ocean, design-around, licensing, FTO, validity, or value conclusion is
inferred from proxies.
Outputs
report_manifest.json is valid, complete, and contains no secrets/absolute paths.
report.html opens offline and contains all promised sections or explicit omissions.
- HTML has no missing assets, remote dependencies, unsafe content, broken links,
inaccessible controls, blank charts, overlap, clipping, or print loss.
- Section/evidence counts reported in the handoff match the files.
Handoff
After writing and validating both outputs, return control to
create-patent-landscape-overview-ip. Report:
report.html written ([section count] sections, [chart count] charts);
report_manifest.json written ([evidence count] evidence entries).
Mode: [A/B/degraded]. Unresolved: [count and summary].
Do not paste the full HTML into the conversation unless the user asks.
Stop conditions
Stop or degrade when:
- upstream artifacts are missing, incompatible, or from different scopes;
tagged_pool.csv cannot be parsed or reconciled;
- family or taxonomy identifiers are unreliable for a requested measure;
- a population claim is unsupported by complete retrieval/aggregation;
- a required route, value, or package conclusion would need fabricated evidence;
- confidential data cannot be processed safely;
- the report cannot be written or validated in the authorized workspace; or
- a requested conclusion requires expert legal or commercial review.
Return the completed sections, failed validation, affected claims, and exact next step.
Do not silently omit a failed section.
1---2name: create-patent-search-report-ip3description: Create the final evidence-backed patent-landscape insight report from validated search, statistics, taxonomy, patent-package, and human-tagging artifacts. Use at Stage 4/4 of the create-patent-landscape-overview-ip suite to aggregate a large tagged patent pool safely, synthesize technology evolution and branch-level value signals, translate patent-package evidence into bounded user actions, and write report_manifest.json plus one self-contained scientific HTML report.4---56# Create a Patent Search Report78## Role in the suite910Act as Stage 4/4, the reporting and synthesis stage of11`create-patent-landscape-overview-ip`.1213Upstream stages are:14151. `search-patents-ip` — validated scope and candidate records.162. `analyze-patent-search-results-ip` — complete-population or explicitly bounded17 statistics, core records, value signals, and chart data.183. `tag-patent-search-results-ip` — taxonomy, key questions, packages, and tagging19 handoff.204. Human Stage 3.5 — validated return of `tagged_pool.csv` when performed.2122Turn those artifacts into a decision-readable report. Do not replace patent counsel,23subject-matter experts, or upstream data validation.2425## Questions this stage answers2627- Which findings matter to the stated product, R&D, strategy, or IP decision?28- Which activity patterns, organizations, branches, and patent groups deserve attention?29- How does the evidence suggest that technical routes have changed over time?30- Where do multiple dated patent-value proxies concentrate by branch?31- What can a user reasonably monitor, read, compare, validate, or refer for review?32- Which claims remain uncertain because data, tagging, or expert review is incomplete?3334## Preconditions3536Use this stage only when:3738- the user has confirmed report objective and audience;39- Stage 1 search scope is validated;40- Stage 2 statistics and provenance are available;41- Stage 3 taxonomy/package artifacts or a sufficiently rich human-tagged pool are42 available; and43- the output environment permits creation of `report_manifest.json` and `report.html`.4445If an upstream stage is incomplete, enter a declared degraded mode or stop. Do not46invent the missing artifact.4748## Authoritative artifact contract4950The suite has no packaged `ARCHITECTURE.md`. Use this embedded contract.5152### Stage 1 inputs5354| Artifact | Required content |55|---|---|56| `search_config.json` | Scope, queries, exclusions, date/language/jurisdiction fields, unit, query version, connector provenance |57| `candidate_pool.csv` | Candidate records with stable IDs and retrieval/screening state |58| `core_recall.csv` | Known-relevant/near-miss controls and recall-review evidence |5960### Stage 2 inputs6162| Artifact | Required content |63|---|---|64| `panorama_stats.json` | Trends, organizations, jurisdictions, status signals, technology distributions, and competitor profiles supported by the population boundary |65| `patent_index.core.json` or equivalent source-authorized core format | Reviewed and tiered representative patents, branch IDs, evidence provenance |66| `value_signals.json` | Candidate-level dated proxy signals and verification state |67| `chart_data.json` | Chart-ready aggregates with measure, unit, date basis, cutoff, scope, and limitations |68| `panorama_stats_report.html` | Stage 2 statistical snapshot; never substitute it for this insight report |6970### Stage 3 inputs7172| Artifact | Required content |73|---|---|74| `tech_breakdown.json` | Versioned technology taxonomy and four-column decomposition |75| `key_questions.json` | Decision-relevant questions, branch/node seeds, rationale, and review status |76| `patent_packages.csv` | Evidence-backed patent groups, selection rubric, rationale, and review status |77| `tagging_demo_sample.csv` | Reviewed example tags and boundary cases |78| `to_be_tagged.csv` | Human-tagging input and schema/version metadata |7980### Human Stage 3.5 input8182| Artifact | Required content |83|---|---|84| `tagged_pool.csv` | Returned complete or explicitly bounded pool, validated taxonomy tags, technical fields, family IDs, assignee, dates, status/signals, schema/taxonomy version, row count, encoding, and reconciliation data |8586### Stage 4 outputs8788| Artifact | Content |89|---|---|90| `report_manifest.json` | Mode, scope, input versions/checksums, section-to-source map, evidence register, derived-field rules, limitations, output inventory, and QA state |91| `report.html` | One offline, self-contained, accessible scientific/executive insight report |9293Create `report_manifest.json` only at Stage 4. Do not expect a Stage 2 manifest with94the same name. Optional runtime exports require user approval and are not part of95this package topology.9697## Verify inputs before synthesis9899For every available artifact:1001011. confirm exact path and readable format;1022. record file size, row/object count, checksum when supplied or practical;1033. validate schema and version;1044. reconcile project, query, taxonomy, family, and data-cutoff identifiers;1055. reconcile candidate, deduplicated, tagged, and reported counts;1066. identify missing columns, null patterns, and conflicting values;1077. preserve the previous accepted version for rollback; and1088. record validation status in the manifest.109110Stop if material inputs belong to different scopes or versions and cannot be111reconciled.112113## Reporting modes114115### Mode A — validated tagged-pool driven116117Use when `tagged_pool.csv` is returned and contains sufficient validated technology118classification plus technical problem/means/effect fields. Reuse existing Stage 2119aggregates; do not rerun patent MCPs merely to render the report.120121Inputs:122123- `tagged_pool.csv` as the classification and technical-evidence body;124- `panorama_stats.json`, `value_signals.json`, `patent_index.core.*`, and125 `chart_data.json` as validated Stage 2 evidence; and126- Stage 3 artifacts when available.127128If Stage 3 questions/packages are absent, derive provisional route and package views129from the tagged pool under the rules below. Label them `automated_derivation`, not130human-rubric output.131132### Mode B — full Stage 3 contract133134Use when `tech_breakdown.json`, `key_questions.json`, and `patent_packages.csv` are135present and the deliverable requires traceability to the reviewed Stage 3 rubric.136Combine them with the tagged pool and Stage 2 evidence.137138### Degraded mode — statistics and reviewed packages only139140Use only if `tagged_pool.csv` is absent but the validated Stage 2 evidence and reviewed141Stage 3 packages can support a useful report. Prominently state:142143```text144Human-tagged population unavailable. Technology distributions and route conclusions145are limited to reviewed packages, rule-hit labels, and validated statistics; they do146not represent a complete tagged population.147```148149Omit unsupported matrices, multi-label distributions, and route claims. Do not fill150them from titles alone.151152## Adapt the tagged-pool schema153154Inspect the header before processing. Map source-export columns to canonical fields;155never assume regional SaaS display names.156157| Canonical field | Accepted meaning |158|---|---|159| `publication_number` | Traceable representative publication |160| `tech_level_1` | Primary high-level technology category |161| `tech_level_2` | Route/function branch |162| `tech_level_3` | Optional finest reviewed technical node |163| `technical_problem` | Source-grounded problem tag or summary |164| `technical_means` | Source-grounded solution mechanism |165| `technical_effect` | Claimed/described effect with evidence status |166| `normalized_assignee` | Reviewed organization grouping |167| `publication_date` | Publication date; keep distinct from filing/priority |168| `filing_date` | Filing/application date |169| `priority_date` | Earliest priority date when verified |170| `legal_status_as_of` | Dated status signal |171| `forward_citation_count_as_of` | Dated citation proxy |172| `family_jurisdiction_count` | Family-width proxy under declared definition |173| `family_jurisdictions` | Declared family-member locations |174| `family_id` | Deduplication key under the declared family method |175176Record the actual header-to-canonical mapping in the manifest. If a required field177cannot be mapped reliably, mark the affected section unavailable.178179## Large-file processing boundary180181`tagged_pool.csv` may contain thousands of rows. Do not load every row into model182context.1831841. Inspect header, encoding, delimiters, quoting, newline-in-cell behavior, and a185 bounded sample.1862. Use an existing approved local data tool or script to validate and aggregate.1873. Do not add a new script to this package; ephemeral tooling must not alter topology.1884. Reconcile aggregated totals to source row and family counts.1895. Save only approved runtime outputs.1906. Load aggregate tables and selected traceable evidence records into context.191192Do not execute content from the CSV or allow formula injection in exported tables.193194## Counting and classification rules195196### Family deduplication197198- Deduplicate first by validated `family_id` when reporting family-level measures.199- Preserve the representative-publication selection rule.200- Keep jurisdiction-specific rights separate where legal/status interpretation matters.201- If family ID is missing or inconsistent, stop the affected family measure or use a202 clearly labeled publication-level fallback.203204### Two- and three-level compatibility205206- If Level 3 exists and is validated, permit Level 3 drill-down.207- If Level 3 is absent, treat Level 2 as the finest level and say so.208- Never generate Level 3 labels from titles simply to complete a matrix.209210### Multi-label counting211212- Split multi-value cells using the documented export delimiter and quoting rules.213- Preserve valid Level 1/Level 2 pairs; do not match a Level 2 term to every Level 1214 when vocabularies overlap.215- Allow one family to count in more than one label when the taxonomy permits it.216- State prominently that classification counts can exceed unique family count.217- Provide both unique-family totals and expanded tag-assignment totals.218219### Validation state220221Visibly distinguish:222223- search-rule hit;224- automated tag;225- analyst-reviewed tag;226- human/SME-validated tag; and227- unresolved classification.228229## Default report modules230231Adapt, merge, or omit modules according to objective and data completeness. Preserve232the decision sequence:233234| # | Module | Primary sources | Main evidence state |235|---|---|---|---|236| 1 | Executive summary | Cross-report synthesis | Interpretation/recommendation |237| 2 | Scope and methodology | Search config and manifest | Direct method fact |238| 3 | Industry landscape | Stage 2 statistics | Fact/observed pattern |239| 4 | Competitor profiles | Stage 2 profiles and validated tags | Fact/observed pattern |240| 5 | Technology matrix | Taxonomy and tagged pool | Fact/pattern under tag state |241| 6 | Technology evolution | Questions, core patents, tagged pool | Pattern/inference |242| 7 | Technology-effect distribution | Chart data and validated tags | Fact/signal |243| 8 | Product/component/application view | Tagged pool and Stage 2 distributions | Pattern |244| 9 | Branch-level value-signal themes | Value signals and core index | Dated proxy/inference |245| 10 | Curated patent groups | Stage 3 packages and value signals | Recommendation |246| 11 | Risks and limitations | All inputs and validation log | Boundary |247| 12 | Appendix and data assets | Manifest and input inventory | Reproducibility |248249Place technology evolution immediately after the technology matrix. Explain signal250themes before user actions and patent groups.251252## Module 6 — Technology evolution253254### Evidence chain255256Build each route as:257258```text259versioned branch/question → time-bounded family evidence → technical problem/means/effect260→ observed change → alternative explanation → bounded route interpretation261```262263### Summary hierarchy264265Include:266267| Level | Field | Content |268|---|---|---|269| Section | `evolution_overview` | Cross-branch observations: acceleration, continuity, convergence, divergence, or emerging attention, each qualified |270| Route | `route_summary` | One sentence describing the observed earlier/current/recent technical emphasis without inventing continuity |271| Phase | `phase_caption` | Evidence-grounded technical characteristics and organizations for that period |272273### Select route evidence274275In Mode B, map `key_questions.seed_node_ids` to the taxonomy and select traceable276families from the tagged pool/core index. In Mode A without key questions:2772781. use validated Level 1 branches as provisional routes;2792. identify dominant valid Level 1/Level 2 pairs by declared measure;2803. select representative families that match both paired levels;2814. segment time according to the dataset and decision—quantiles, technology eras,282 or explicit periods—not fixed calendar years;2835. consider technical relevance, evidence depth, date, and diversity;2846. prevent accidental family reuse when it would distort comparison; and2857. disclose any algorithmic diversity/organization rotation as a sampling choice,286 not evidence of actual market entry timing.287288Do not require three families or three non-empty phases. Show sparse/empty periods.289290### Turning points291292Use controlled types only when evidence supports them:293294- route branching;295- notable disclosure;296- organization entry in the observed dataset;297- disclosed performance or capability change; and298- application-context shift.299300Each `turning_point` contains `type`, `note`, evidence IDs, date basis, and uncertainty.301Generate its note from the selected family’s problem/means/effect evidence. Do not302invent a transition between records.303304### Visualization305306Use accessible HTML/CSS/SVG timelines or small multiples. Each route shows date basis,307family IDs/publications, organization, validated node, evidence state, and sparse-data308qualification. Static text must convey the finding without interaction.309310## Module 7 — Technology-effect distribution311312Build a technical-means × reported-effect matrix from validated upstream chart data313or tagged-pool aggregation.314315- Define row, column, cell measure, unit, multi-label policy, tag state, and cutoff.316- Use family count only after validated family deduplication.317- Distinguish claimed/described effects from independently demonstrated performance.318- Represent unavailable cells separately from observed zero.319- Do not call a combination “insignificant” merely because the dataset has no records.320- Use an accessible heatmap/table or bubbles with a tabular alternative.321322## Module 9 — Branch-level value-signal themes323324Aggregate candidate-level proxies in `value_signals.json` by validated `branch_id`.325Possible dimensions include dated citation, family breadth, legal-status, organization326concentration, transaction/assertion-event, and portfolio-priority signals.327328For each dimension record:329330- definition and source;331- verified versus recall-proxy state;332- date/cutoff and jurisdiction coverage;333- normalization and missing-data treatment;334- aggregation function and denominator; and335- sensitivity to outliers or branch size.336337Describe results as “higher observed signal concentration under this dataset” or338“lower observed signal density.” Do not call the score patent value, moat strength,339defensibility, blue ocean, freedom to operate, availability, or market opportunity.340341Prefer bars, dot plots, or a signal matrix. Avoid radar charts when dimensions have342unlike scales or missing values.343344## Module 10 — Curated patent groups345346### Preserve analyst evidence347348Retain Stage 3 `recommendation_reason`, rubric dimensions, evidence IDs, review status,349and all limitations in the manifest/evidence layer.350351### Translate into bounded user actions352353For each family add:354355| Field | Meaning |356|---|---|357| `use_case` | Evidence-supported next action such as read, monitor, compare, technical reference, data validation, or counsel/commercial review |358| `purpose_tag` | Controlled action category approved for this report |359| `answers_question` | Link to a validated key question when the mapping exists |360| `package_summary` | What the group addresses, why it matters, and which records to start with |361362Do not hard-code the source’s five Chinese action labels. Define an English controlled363vocabulary appropriate to the objective. Never translate proxies into “design-around,”364“licensing candidate,” “acquire,” “enforce,” or “FTO action” without qualified review.365366### Mode A provisional package derivation367368If `patent_packages.csv` is absent:3693701. derive candidates by validated finest branch;3712. rank with disclosed relevance, evidence-depth, family/citation/status proxies;3723. treat missing proxy data explicitly;3734. test for age, organization, and family-size bias;3745. select only supported records rather than a fixed two-to-three quota;3756. label the group `automated_derivation`; and3767. require human review before any transaction, legal, or portfolio action.377378When no `key_questions` mapping exists, leave `answers_question` unresolved. Do not379fabricate a question link.380381### Two-layer rendering382383Main layer:384385- decision-readable action;386- purpose tag;387- linked question if verified; and388- package summary.389390Expandable evidence layer:391392- original rubric/reason;393- family, citation, status, organization, and review signals;394- evidence IDs and provenance; and395- limitations/legal boundary.396397The report must remain understandable when expandable controls are not used or when398printed.399400## Evidence model401402Use:403404| Level | Meaning |405|---|---|406| L1 | Direct, source-backed data or method fact |407| L2 | Observed pattern in the defined dataset |408| L3 | Analytical interpretation with alternatives and uncertainty |409| L4 | Business/R&D/IP workflow recommendation |410| L5 | Legal, transaction, status, or risk signal requiring specialist review |411412Every material claim maps to one or more evidence entries. A manifest evidence entry413contains:414415```json416{417 "evidence_id": "E-001",418 "level": "L2",419 "claim": "[bounded observation]",420 "source_file": "[validated upstream artifact]",421 "source_field": "[field or aggregate]",422 "record_ids": ["[traceable IDs]"],423 "counting_method": "[unit, deduplication, multi-label policy]",424 "data_cutoff": "[YYYY-MM-DD]",425 "validation_state": "[verified/proxy/reviewed/unresolved]",426 "limitations": "[material boundary]"427}428```429430Do not include confidential examples in the Skill package.431432## `report_manifest.json` contract433434Include:435436- report ID, version, generated date, objective, audience, and mode;437- scope, jurisdictions, date basis, unit, family definition, and data cutoff;438- every input path, artifact version, checksum/size/count, validation result;439- header/column mappings and transformations;440- section order and section-to-source/evidence mapping;441- evolution overview, routes, summaries, phases, and turning points;442- branch signal definitions and aggregations;443- package summaries, bounded use cases, purpose tags, question links, original reasons;444- evidence register;445- degraded/automated/proxy flags;446- limitations and unresolved items;447- output paths and QA results; and448- no secret, API key, raw confidential payload, or absolute user path.449450Write atomically when possible. Validate JSON before handoff.451452## `report.html` design contract453454### Information architecture455456- Start with title, decision objective, mode, scope, date basis, family unit, cutoff,457 sources, and limitations.458- Make executive findings evidence-linked and action-bounded.459- Place each visualization next to its claim and qualification.460- Preserve the default module sequence unless an approved audience need changes it.461- End with reproducibility assets and legal/data boundaries.462463### Scientific/executive visual system464465- Use a white/neutral canvas, dark navy/slate hierarchy, restrained teal data accent,466 amber qualifications, and red only for real escalation.467- Use system fonts; do not rely on regional or remote fonts.468- Prefer whitespace, rules, and typography over nested cards.469- Use flat 2D bars, lines, heatmaps/tables, timelines, and evidence cards.470- Avoid decorative hero sections, gradients, stock imagery, 3D charts, visual drama,471 and vendor-interface imitation.472- Pair color with labels or symbols and maintain accessible contrast.473474### Technical safety475476- Deliver one HTML file with inline CSS and pre-aggregated data.477- Prefer static HTML/CSS/SVG; allow minimal inline JavaScript only as progressive478 enhancement with a complete non-script fallback.479- Use no external CDN, D3, font, image, iframe, tracker, or network dependency.480- Escape all user and retrieved content.481- Permit only validated safe URLs and add `rel="noopener noreferrer"` to new-tab links.482- Do not execute or interpolate raw CSV/JSON content as code.483- Prevent spreadsheet-formula and HTML/script injection in displayed/exported values.484485### Accessibility and print486487- Use semantic landmarks, ordered heading levels, table headers/captions, visible focus,488 keyboard-safe controls, and text equivalents for charts.489- Keep wide tables horizontally scrollable and long identifiers break-safe.490- Respect reduced motion and avoid animation needed for meaning.491- Ensure collapsed evidence is visible or summarized in print.492- Verify desktop, narrow viewport, grayscale, and PDF/print layouts.493494### Chart captions495496Every quantitative view states:497498- question and bounded takeaway;499- measure and denominator;500- date field and period;501- unit/family definition;502- scope and cutoff;503- multi-label duplicate-count policy;504- validation/proxy state; and505- material limitation.506507Show unavailable data rather than a fake or zero-valued chart.508509## MCP boundary510511Stage 4 normally reuses validated upstream artifacts. Do not rerun MCPs merely to512recreate existing statistics.513514If an approved material evidence gap requires retrieval, use only the installed live515schema of these verified global connectors:516517- `advanced_patent_search` — https://open.patsnap.com/marketplace/mcp-servers/patent-search518- `patent_briefing` — https://open.patsnap.com/marketplace/mcp-servers/patent-briefing519- `deep_patent_mining` — https://open.patsnap.com/marketplace/mcp-servers/patent-mining520- `global_core_patent_database` — https://open.patsnap.com/marketplace/mcp-servers/core-patents521522Record why the gap could not be resolved from upstream artifacts and how new evidence523was reconciled. Do not use source endpoint aliases as connector names.524525## Legal and analytical boundaries526527Do not provide:528529- formal freedom-to-operate or infringement opinions;530- patent validity, novelty, or inventive-step opinions;531- standards-essentiality opinions;532- legal claim-scope conclusions;533- patent valuation or transaction recommendations;534- product-launch or market-adoption certainty; or535- unsupported business conclusions.536537Use language such as “under the defined dataset,” “the patent evidence suggests,”538“observed proxy concentration,” and “requires technical, commercial, or legal review.”539540## Quality gate541542### Input and mode543544- Every required input exists or the mode declares its absence.545- Schema, version, count, checksum, scope, taxonomy, family, and cutoff reconcile.546- Mode A, Mode B, or degraded mode is explicit at the top and in the manifest.547- No large tagged pool was loaded row-by-row into context.548549### Data and synthesis550551- Family and multi-label counting rules are applied and disclosed.552- Population, sample, core set, package, and representative records are distinct.553- Evolution periods derive from the dataset and contain no fabricated transition.554- Turning points cite their family and technical evidence.555- Value themes expose proxy definitions, missing data, aggregation, and limitations.556- Package actions are bounded and preserve original Stage 3 evidence.557558### Evidence and law559560- Every material claim maps to evidence IDs.561- Facts, patterns, interpretations, recommendations, and legal signals are distinct.562- Status/citation/family/transaction values are dated and source-labeled.563- No moat, blue-ocean, design-around, licensing, FTO, validity, or value conclusion is564 inferred from proxies.565566### Outputs567568- `report_manifest.json` is valid, complete, and contains no secrets/absolute paths.569- `report.html` opens offline and contains all promised sections or explicit omissions.570- HTML has no missing assets, remote dependencies, unsafe content, broken links,571 inaccessible controls, blank charts, overlap, clipping, or print loss.572- Section/evidence counts reported in the handoff match the files.573574## Handoff575576After writing and validating both outputs, return control to577`create-patent-landscape-overview-ip`. Report:578579```text580report.html written ([section count] sections, [chart count] charts);581report_manifest.json written ([evidence count] evidence entries).582Mode: [A/B/degraded]. Unresolved: [count and summary].583```584585Do not paste the full HTML into the conversation unless the user asks.586587## Stop conditions588589Stop or degrade when:590591- upstream artifacts are missing, incompatible, or from different scopes;592- `tagged_pool.csv` cannot be parsed or reconciled;593- family or taxonomy identifiers are unreliable for a requested measure;594- a population claim is unsupported by complete retrieval/aggregation;595- a required route, value, or package conclusion would need fabricated evidence;596- confidential data cannot be processed safely;597- the report cannot be written or validated in the authorized workspace; or598- a requested conclusion requires expert legal or commercial review.599600Return the completed sections, failed validation, affected claims, and exact next step.601Do not silently omit a failed section.602