Private-Company Research: Multi-Lens Deep Framework
Deep research on an unlisted company (e.g. Ant Group, ByteDance, SpaceX, Stripe).
Ultimate goal: under information scarcity, recover the company's true value — not the market valuation, but what the business is actually worth.
Framework Characteristics
Private vs public research: no standardized financials (multi-source patchwork + cross-validation); few valuation anchors (funding rounds, comparables, scenarios); large information asymmetry ("jigsaw" research); uncertain exit path (IPO / M&A / secondary).
AI Research Bias Self-Check (core premise)
Private companies are where AI bias is worst. Watch for:
- False conservatism — with little data, AI gives conservative/vague conclusions, but scarce data ≠ bad company.
- False precision — to fill the template, AI disguises "reasonable guess" as "sourced analysis".
- Comparables trap — forcing a public-comp overlay inherits public-market logic and misses private-specific value.
- Survivorship bias — what's searchable online is mostly company-propagated good news.
Counter: prefer leaving blanks ("I don't know") over filling tables with speculation to fake certainty; label every data point with confidence (🟢high/🟡medium/🔴low); separate verifiable fact from inference; when information is extremely scarce, switch to "first-principles mode" and answer only: ① what real problem does this business solve? ② why this team? ③ ceiling if it succeeds / how it dies if it fails? ④ the key validation node at this stage?
Invert the asymmetry: the market knows little about private companies → pricing is inefficient → that's exactly where alpha may live.
Execution
Six lenses, best run in parallel (via run_swarm, one worker per lens; or sequentially via web_search):
| Role |
Lens |
| business-decoder |
Business model + product/user analysis: "what is this business, essentially" |
| financial-detective |
Financial patchwork + valuation: "recover the true financial picture under missing data" |
| competitive-mapper |
Industry + competition + substitution: "who competes, who could disrupt" |
| risk-governance-analyst |
Risk全景 + management/governance/investors: "what could go wrong, who's at the helm" |
| tech-ip-analyst |
Tech stack / patents / R&D / moat: "is the tech barrier real and durable" |
| signal-miner |
Alternative data (hiring / patents / litigation / app / supply chain): "clues beyond the usual sources" |
You (team-lead) integrate, patch the picture, cross-validate, output the final report.
Lens 1: Business Model & Users (business-decoder)
- Core business definition: one sentence (Duan Yongping style: plain language to a smart layperson). What problem? For whom? If the company didn't exist, what would users do? Is demand rigid (cut in a downturn)?
- Revenue model: ads/commission/subscription/take-rate/financial/SaaS/hardware; mix and trend; monetization efficiency (ARPU / take rate / conversion); recurring vs one-off; concentration; predictability.
- Unit economics: CAC (paid vs organic, by channel, trend), LTV, LTV/CAC, payback, marginal cost, scale-inflection point.
- Product matrix & flywheel: core + extension + incubation; network/data/scale flywheel; iteration speed.
- Users: MAU/DAU (from QuestMobile/Sensor Tower/SimilarWeb), S-curve stage, stickiness (DAU/MAU, retention), profile, reputation (App Store trend, social sentiment).
- Moat (6 dimensions, ★1-5): network effects, switching costs, brand mind, data barrier, regulatory license, scale economies — each with evidence + trend (widening/stable/narrowing) + durability. Overall: wide/narrow/none.
Lens 2: Financial Detective
No standard financials; multi-source patchwork + cross-validation. Every data point: source, time, confidence, derivation.
Source priority: 🟢 prospectus/regulatory filings, parent-company annual report disclosure, regulatory penalties, bond/ABS offering documents → 🟡 business registry, funding news, third-party reports, deep media (LatePost/The Information/36Kr/Bloomberg) → 🔴 industry extrapolation, ex-employee leaks.
Key metrics: revenue (scale/growth/mix/quantity×price), cost (gross margin/R&D/sales/G&A rates, vs peers), profit (EBITDA/net income/profitability timeline), cash flow (operating/burn rate/runway), efficiency (per-capita revenue, capital efficiency).
Cross-validation: list every source for the same metric; check convergence across methods; flag single-source ("isolated evidence") data.
Funding history: full timeline (round/amount/valuation/lead investor); health of the curve, interval, down-rounds, whether existing investors keep participating; latest-round terms (liquidation preference / anti-dilution / ratchet) and their effect on common-share value.
Valuation (multi-method): ① last-round (adjust for liquidation prefs, 20-40% discount); ② comparable public comps (3-5, PS/PE/EV-EBITDA, liquidity discount 20-30%); ③ DCF scenarios (bear/base/bull, each assumption grounded); ④ terminal-value rollback (5/10y terminal state → implied IRR); ⑤ transaction comps (recent M&A/funding multiples).
Valuation synthesis: do the methods converge? If divergent, explain. Distinguish "fair value" and "conservative (margin-of-safety) value".
Lens 3: Competitive Landscape (competitive-mapper)
- Market: TAM/SAM/SOM, penetration, stage (emergence/growth/mature/decline), growth drivers.
- Value chain (text map): upstream → company's link (profit pool share) → downstream; bargaining power; structural shifts.
- Porter's five forces (★1-5): rivalry, new entrants, substitutes, supplier power, buyer power.
- Competitor scan: direct/indirect/substitute/potential entrants (giants) — share, funding, strengths, weaknesses, threat level. Multi-dimensional compare with 2-3 closest competitors.
- Dynamics: last-12-month changes; infer competitor strategy from hiring/patents/products; tech (esp. AI) and regulation effects; winner-take-all vs oligopoly.
- Scenarios: company wins / coexistence / disrupted — conditions and probabilities.
- Global benchmarks: overseas analogs, path, valuation, post-IPO performance, and benchmark limitations.
Lens 4: Risk & Governance (risk-governance-analyst)
- Founder/CEO: background, foresight (3-year prediction accuracy), execution (promise delivery), values, controversies. ★1-5.
- Core team: backgrounds, 2-year talent flow (who left/joined, net), complementarity, culture signals (Glassdoor trends), key-person dependency.
- Equity & governance: founder control (dual-class / concerted action / VIE), dilution trend, employee equity; board, related-party, same-industry competition, majority/minority conflicts.
- Investor roster: lead-investor brand, strategic capital synergy, exit pressure (fund life / secondary sales / ratchet maturity), red flags.
- Risk matrix: regulatory / competition / tech / talent / funding / IPO / geopolitical / monetization / governance / compliance / macro / ESG — each probability × impact × severity × hedgeability × monitor.
- Exit paths: A/HK/US IPO, M&A, secondary, SPAC, stay-private — probability, window, valuation, preconditions, obstacles.
- Worst case (Munger inversion): 3 specific failure paths + probabilities; liquidation value; why smart money doesn't invest (≥5 reasons); failed analogs; the "thesis broken" signal.
Lens 5: Tech & IP (tech-ip-analyst)
- Tech stack: inferred from hiring/blog/open-source/talks; tech-debt signals (refactor hiring, stack switches, outage complaints).
- Patents (Google Patents/CNIPA/USPTO): total/pending/last-2y/field/citation/international; quality (core patents? aligned to business? litigation?); trend; vs competitors.
- R&D: investment (headcount/expense rate vs peers), output (papers/conferences/open-source/blog), efficiency (research-to-product, commercialization).
- Tech talent: core leaders' backgrounds, density (top-institution share), comp competitiveness, attrition signals, hiring direction (→ strategy).
- Tech moat (★1-5): algorithm/model, data, engineering, talent, ecosystem — each with durability (AI-era half-life may be short).
- AI/new-tech impact: is the company a beneficiary or a target of disruption?
Lens 6: Alternative-Data Signals (signal-miner)
Private companies have limited conventional info; alt-data often beats news.
- Hiring (LinkedIn/Boss/Indeed): scale/trend, structure (R&D/product/sales/data/international/compliance/IR — IR hiring = IPO signal; compliance = regulatory or IPO; JD tech stack = strategy).
- App/product (App Store/七麦/SimilarWeb): rank, rating trend, downloads, update frequency, complaint themes, web traffic.
- Social sentiment (Weibo/Zhihu/Xiaohongshu/X/Reddit): official engagement, organic discussion, KOL views, negative events, insider leaks.
- Business/legal (天眼查/企查查): registry/paid-in/equity changes/subsidiaries (new = new biz; deregistered = contraction)/scope changes; litigation/arbitration/penalties/enforcement.
- Supply chain: known suppliers (if listed, check their filings), procurement, partner evaluation.
- Digital footprint: registered domains (new = new biz), subdomains (api/pay → architecture), trademarks (new brands).
- Industry exposure: exec talks, awards, government/association interaction, media frequency/quality.
- Secondary-market signals (if any): SharesPost/EquityZen, implied valuation vs last round, employee selling.
Anomaly list (most important): things inconsistent with the company's narrative; inconsistent with industry norms; sudden changes (hiring freeze / executive departures); unexplained.
Cross-Validation (team-lead, mandatory)
Before synthesis, the team-lead must:
- Data conflict arbitration: same metric across sources — list all, state which is adopted and why.
- Signal consistency matrix: business-growth signal vs hiring trend? tech-leadership narrative vs patent/talent data? valuation level vs competitive position? management narrative vs action signals? (contradictions must be explained)
- Information jigsaw: white zones (known) / gray (clues but uncertain) / black (unknown).
- Bias check: is positive info detailed while negative is brief? Does every positive judgment have a reverse check?
Final Report Structure
- One-line conclusion (50-100 words): what's it worth, why.
- Company snapshot (with confidence column).
- Six-lens scorecard (★1-5 + core judgment + confidence + completeness), overall score.
- Key data jigsaw (only cross-validated, with source count + confidence).
- Signal-consistency matrix.
- Per-lens summary (3-5 top findings each).
- Fair value assessment: business essence + 7-dimension moat card + 5-method valuation + fair value range (conservative/reasonable/optimistic + current market valuation + margin of safety %).
- Investment thesis: bull 5-7 (with sources) vs bear 5-7 (with sources), which side is stronger.
- Risk matrix (top 3 + mitigation).
- Exit-path assessment.
- Investment decision table (one-pager: core logic 3 sentences + value range + key assumptions & validation nodes + fatal risks & "thesis broken" signals + conclusion + expected return/timeline).
- Information-gap map (dimension / known / missing / missing-impact / how-to-get): does the gap affect the core conclusion? If yes, state "under missing X, conclusion confidence is Y".
- Tracking checklist (item / frequency / source / metric / alert threshold).
- Summary paragraph (150-250 words).
Save via write_file to reports/{company}/{company}-private-{YYYYMMDD}.md. Run report_audit on the numbers as a quality gate.
Data Labeling Standard (strict)
- Every key data point: source (specific to media + article), time (year/month), confidence (🟢 prospectus/official / 🟡 credible media / 🔴 estimate/rumor).
- Conflicting data: list all + explain difference and adoption.
- Separate fact from inference: fact in normal text; inference in italics with derivation.
- Missing info: explicitly mark "data missing"; never fabricate.
Key Principles
- 6 lenses in parallel (
run_swarm, or sequential).
- Transparent derivation — show the math and assumptions; don't hand-wave numbers.
- Cross-validate — key data ≥2 sources; conflicts all listed.
- Signal-consistency check — mandatory cross-lens check at synthesis.
- Clear conclusion — don't dodge invest/watch/avoid; state confidence.
- Search in both EN and CN — private-company info spans both.
- Honest blanks — distinguish "sourced analysis" from "speculative fill"; "this dimension lacks data, no meaningful conclusion" is acceptable.
- Alt-data is not noise — hiring/patents/litigation/app data may be closer to truth than news.
- True-value focus — the goal is what the business is worth, not a pretty report. If info can't support a reliable valuation, say so.
- Scarce data ≠ bad company — short AI output ≠ low certainty. Under extreme scarcity, switch to first-principles mode.
1---2name: private-company-research3description: Deep research framework for pre-IPO / private companies (Ant Group, SpaceX, Stripe, ByteDance...). Six analyst lenses — business model, financial forensics, competitive landscape, risk & governance, tech & IP, alternative-data signals — run in parallel via run_swarm, then cross-validated for signal consistency before any verdict. Built around the core challenge of private-company work: information is scarce, so every data point carries a confidence label (high / medium / low), inference is shown separately from fact, and 'I don't know' is a valid output. Outputs a fair-value range, exit-path analysis, and an information-gap map. Use for any unlisted company where you need to judge what the business is actually worth.4---56# Private-Company Research: Multi-Lens Deep Framework78Deep research on an unlisted company (e.g. Ant Group, ByteDance, SpaceX, Stripe).910**Ultimate goal**: under information scarcity, recover the company's **true value** — not the market valuation, but what the business is actually worth.1112## Framework Characteristics1314Private vs public research: no standardized financials (multi-source patchwork + cross-validation); few valuation anchors (funding rounds, comparables, scenarios); large information asymmetry ("jigsaw" research); uncertain exit path (IPO / M&A / secondary).1516## AI Research Bias Self-Check (core premise)1718Private companies are where AI bias is worst. Watch for:1920- **False conservatism** — with little data, AI gives conservative/vague conclusions, but scarce data ≠ bad company.21- **False precision** — to fill the template, AI disguises "reasonable guess" as "sourced analysis".22- **Comparables trap** — forcing a public-comp overlay inherits public-market logic and misses private-specific value.23- **Survivorship bias** — what's searchable online is mostly company-propagated good news.2425**Counter**: prefer leaving blanks ("I don't know") over filling tables with speculation to fake certainty; label every data point with confidence (🟢high/🟡medium/🔴low); separate verifiable fact from inference; when information is extremely scarce, switch to "first-principles mode" and answer only: ① what real problem does this business solve? ② why this team? ③ ceiling if it succeeds / how it dies if it fails? ④ the key validation node at this stage?2627> Invert the asymmetry: the market knows little about private companies → pricing is inefficient → that's exactly where alpha may live.2829## Execution3031Six lenses, best run **in parallel** (via `run_swarm`, one worker per lens; or sequentially via `web_search`):3233| Role | Lens |34|------|------|35| business-decoder | Business model + product/user analysis: "what is this business, essentially" |36| financial-detective | Financial patchwork + valuation: "recover the true financial picture under missing data" |37| competitive-mapper | Industry + competition + substitution: "who competes, who could disrupt" |38| risk-governance-analyst | Risk全景 + management/governance/investors: "what could go wrong, who's at the helm" |39| tech-ip-analyst | Tech stack / patents / R&D / moat: "is the tech barrier real and durable" |40| signal-miner | Alternative data (hiring / patents / litigation / app / supply chain): "clues beyond the usual sources" |4142You (team-lead) integrate, patch the picture, cross-validate, output the final report.4344## Lens 1: Business Model & Users (business-decoder)4546- **Core business definition**: one sentence (Duan Yongping style: plain language to a smart layperson). What problem? For whom? If the company didn't exist, what would users do? Is demand rigid (cut in a downturn)?47- **Revenue model**: ads/commission/subscription/take-rate/financial/SaaS/hardware; mix and trend; monetization efficiency (ARPU / take rate / conversion); recurring vs one-off; concentration; predictability.48- **Unit economics**: CAC (paid vs organic, by channel, trend), LTV, LTV/CAC, payback, marginal cost, scale-inflection point.49- **Product matrix & flywheel**: core + extension + incubation; network/data/scale flywheel; iteration speed.50- **Users**: MAU/DAU (from QuestMobile/Sensor Tower/SimilarWeb), S-curve stage, stickiness (DAU/MAU, retention), profile, reputation (App Store trend, social sentiment).51- **Moat (6 dimensions, ★1-5)**: network effects, switching costs, brand mind, data barrier, regulatory license, scale economies — each with evidence + trend (widening/stable/narrowing) + durability. Overall: wide/narrow/none.5253## Lens 2: Financial Detective5455No standard financials; multi-source patchwork + cross-validation. **Every data point: source, time, confidence, derivation.**5657**Source priority**: 🟢 prospectus/regulatory filings, parent-company annual report disclosure, regulatory penalties, bond/ABS offering documents → 🟡 business registry, funding news, third-party reports, deep media (LatePost/The Information/36Kr/Bloomberg) → 🔴 industry extrapolation, ex-employee leaks.5859**Key metrics**: revenue (scale/growth/mix/quantity×price), cost (gross margin/R&D/sales/G&A rates, vs peers), profit (EBITDA/net income/profitability timeline), cash flow (operating/burn rate/runway), efficiency (per-capita revenue, capital efficiency).6061**Cross-validation**: list every source for the same metric; check convergence across methods; flag single-source ("isolated evidence") data.6263**Funding history**: full timeline (round/amount/valuation/lead investor); health of the curve, interval, down-rounds, whether existing investors keep participating; latest-round terms (liquidation preference / anti-dilution / ratchet) and their effect on common-share value.6465**Valuation (multi-method)**: ① last-round (adjust for liquidation prefs, 20-40% discount); ② comparable public comps (3-5, PS/PE/EV-EBITDA, liquidity discount 20-30%); ③ DCF scenarios (bear/base/bull, each assumption grounded); ④ terminal-value rollback (5/10y terminal state → implied IRR); ⑤ transaction comps (recent M&A/funding multiples).6667**Valuation synthesis**: do the methods converge? If divergent, explain. Distinguish "fair value" and "conservative (margin-of-safety) value".6869## Lens 3: Competitive Landscape (competitive-mapper)7071- **Market**: TAM/SAM/SOM, penetration, stage (emergence/growth/mature/decline), growth drivers.72- **Value chain** (text map): upstream → company's link (profit pool share) → downstream; bargaining power; structural shifts.73- **Porter's five forces** (★1-5): rivalry, new entrants, substitutes, supplier power, buyer power.74- **Competitor scan**: direct/indirect/substitute/potential entrants (giants) — share, funding, strengths, weaknesses, threat level. Multi-dimensional compare with 2-3 closest competitors.75- **Dynamics**: last-12-month changes; infer competitor strategy from hiring/patents/products; tech (esp. AI) and regulation effects; winner-take-all vs oligopoly.76- **Scenarios**: company wins / coexistence / disrupted — conditions and probabilities.77- **Global benchmarks**: overseas analogs, path, valuation, post-IPO performance, and benchmark limitations.7879## Lens 4: Risk & Governance (risk-governance-analyst)8081- **Founder/CEO**: background, foresight (3-year prediction accuracy), execution (promise delivery), values, controversies. ★1-5.82- **Core team**: backgrounds, 2-year talent flow (who left/joined, net), complementarity, culture signals (Glassdoor trends), key-person dependency.83- **Equity & governance**: founder control (dual-class / concerted action / VIE), dilution trend, employee equity; board, related-party, same-industry competition, majority/minority conflicts.84- **Investor roster**: lead-investor brand, strategic capital synergy, exit pressure (fund life / secondary sales / ratchet maturity), red flags.85- **Risk matrix**: regulatory / competition / tech / talent / funding / IPO / geopolitical / monetization / governance / compliance / macro / ESG — each probability × impact × severity × hedgeability × monitor.86- **Exit paths**: A/HK/US IPO, M&A, secondary, SPAC, stay-private — probability, window, valuation, preconditions, obstacles.87- **Worst case (Munger inversion)**: 3 specific failure paths + probabilities; liquidation value; why smart money doesn't invest (≥5 reasons); failed analogs; the "thesis broken" signal.8889## Lens 5: Tech & IP (tech-ip-analyst)9091- **Tech stack**: inferred from hiring/blog/open-source/talks; tech-debt signals (refactor hiring, stack switches, outage complaints).92- **Patents** (Google Patents/CNIPA/USPTO): total/pending/last-2y/field/citation/international; quality (core patents? aligned to business? litigation?); trend; vs competitors.93- **R&D**: investment (headcount/expense rate vs peers), output (papers/conferences/open-source/blog), efficiency (research-to-product, commercialization).94- **Tech talent**: core leaders' backgrounds, density (top-institution share), comp competitiveness, attrition signals, hiring direction (→ strategy).95- **Tech moat (★1-5)**: algorithm/model, data, engineering, talent, ecosystem — each with durability (AI-era half-life may be short).96- **AI/new-tech impact**: is the company a beneficiary or a target of disruption?9798## Lens 6: Alternative-Data Signals (signal-miner)99100> Private companies have limited conventional info; alt-data often beats news.101102- **Hiring** (LinkedIn/Boss/Indeed): scale/trend, structure (R&D/product/sales/data/international/compliance/IR — IR hiring = IPO signal; compliance = regulatory or IPO; JD tech stack = strategy).103- **App/product** (App Store/七麦/SimilarWeb): rank, rating trend, downloads, update frequency, complaint themes, web traffic.104- **Social sentiment** (Weibo/Zhihu/Xiaohongshu/X/Reddit): official engagement, organic discussion, KOL views, negative events, insider leaks.105- **Business/legal** (天眼查/企查查): registry/paid-in/equity changes/subsidiaries (new = new biz; deregistered = contraction)/scope changes; litigation/arbitration/penalties/enforcement.106- **Supply chain**: known suppliers (if listed, check their filings), procurement, partner evaluation.107- **Digital footprint**: registered domains (new = new biz), subdomains (api/pay → architecture), trademarks (new brands).108- **Industry exposure**: exec talks, awards, government/association interaction, media frequency/quality.109- **Secondary-market signals** (if any): SharesPost/EquityZen, implied valuation vs last round, employee selling.110111**Anomaly list (most important)**: things inconsistent with the company's narrative; inconsistent with industry norms; sudden changes (hiring freeze / executive departures); unexplained.112113## Cross-Validation (team-lead, mandatory)114115Before synthesis, the team-lead must:1161171. **Data conflict arbitration**: same metric across sources — list all, state which is adopted and why.1182. **Signal consistency matrix**: business-growth signal vs hiring trend? tech-leadership narrative vs patent/talent data? valuation level vs competitive position? management narrative vs action signals? (contradictions must be explained)1193. **Information jigsaw**: white zones (known) / gray (clues but uncertain) / black (unknown).1204. **Bias check**: is positive info detailed while negative is brief? Does every positive judgment have a reverse check?121122## Final Report Structure1231241. **One-line conclusion** (50-100 words): what's it worth, why.1252. **Company snapshot** (with confidence column).1263. **Six-lens scorecard** (★1-5 + core judgment + confidence + completeness), overall score.1274. **Key data jigsaw** (only cross-validated, with source count + confidence).1285. **Signal-consistency matrix**.1296. **Per-lens summary** (3-5 top findings each).1307. **Fair value assessment**: business essence + 7-dimension moat card + 5-method valuation + **fair value range** (conservative/reasonable/optimistic + current market valuation + margin of safety %).1318. **Investment thesis**: bull 5-7 (with sources) vs bear 5-7 (with sources), which side is stronger.1329. **Risk matrix** (top 3 + mitigation).13310. **Exit-path assessment**.13411. **Investment decision table** (one-pager: core logic 3 sentences + value range + key assumptions & validation nodes + fatal risks & "thesis broken" signals + conclusion + expected return/timeline).13512. **Information-gap map** (dimension / known / missing / missing-impact / how-to-get): does the gap affect the core conclusion? If yes, state "under missing X, conclusion confidence is Y".13613. **Tracking checklist** (item / frequency / source / metric / alert threshold).13714. **Summary paragraph** (150-250 words).138139Save via `write_file` to `reports/{company}/{company}-private-{YYYYMMDD}.md`. Run `report_audit` on the numbers as a quality gate.140141## Data Labeling Standard (strict)142143- Every key data point: **source** (specific to media + article), **time** (year/month), **confidence** (🟢 prospectus/official / 🟡 credible media / 🔴 estimate/rumor).144- Conflicting data: **list all** + explain difference and adoption.145- **Separate fact from inference**: fact in normal text; inference in *italics* with derivation.146- Missing info: explicitly mark "data missing"; never fabricate.147148## Key Principles1491501. **6 lenses in parallel** (`run_swarm`, or sequential).1512. **Transparent derivation** — show the math and assumptions; don't hand-wave numbers.1523. **Cross-validate** — key data ≥2 sources; conflicts all listed.1534. **Signal-consistency check** — mandatory cross-lens check at synthesis.1545. **Clear conclusion** — don't dodge invest/watch/avoid; state confidence.1556. **Search in both EN and CN** — private-company info spans both.1567. **Honest blanks** — distinguish "sourced analysis" from "speculative fill"; "this dimension lacks data, no meaningful conclusion" is acceptable.1578. **Alt-data is not noise** — hiring/patents/litigation/app data may be closer to truth than news.1589. **True-value focus** — the goal is what the business is worth, not a pretty report. If info can't support a reliable valuation, say so.15910. **Scarce data ≠ bad company** — short AI output ≠ low certainty. Under extreme scarcity, switch to first-principles mode.