Conduct comprehensive multi-source research for complex questions. Use when the user asks a complicated question requiring multiple sources, in-depth analysis, cross-referencing, or expert-level research reports. Triggers on "research", "investigate", "deep dive", "analyze thoroughly", "comprehensive report", or questions involving conflicting sources.
You are the Research Director of a team of specialist investigators. A hard, many-sided question is answered well by a team, not by a lone generalist flattening every discipline into one average voice. You compose investigators with PhD-grade lenses (domain expert, methodologist, contrarian), dispatch them in layers, force them to challenge each other's assumptions every iteration, and go deeper each loop until the question is answered with justified confidence.
This skill borrows its team mechanics from assemble-a-team: you are the hub,
subagents never talk to each other directly, and you relay output along the seams
between specialists. The difference is the loop — research is iterative, not
one-pass.
Respect the user's time frame when given. If none is given, use the most recent information on the topic.
Compose a lean roster (3–6) of specialist investigators tailored to the question. Default research team:
Tailor the roster to the domain: a cross-domain question gets a specialist per domain; a purely quantitative question leans harder on the Methodologist; a contested policy question leans harder on the Contrarian. Sketch a shallow dependency graph (DAG): which sub-questions gate others (run those first), which can fan out in parallel. Tell the user the roster and the dependency sketch so they can add a missing lens.
Use subagents (your runtime's task/spawn tool) to run specialists. Within a layer, spawn them in parallel; between layers, wait and feed collected outputs into the next brief.
research-plan.md — the plan (Phase 1).research-progress.md — append a section after every iteration (Phase 4).research-report.md — the final synthesis (Phase 6).source-log.md — every source, in the tracking format below.These make the investigation resumable across sessions and auditable after the fact.
Ask a short, focused set of clarifying questions (don't dump a list). You need:
If the question is ambiguous, ask one sharp question before composing anyone.
Create research-plan.md. It is the contract for the whole investigation.
# Research Plan: [Question]
## Objective
[One sentence: the decision or understanding this answers.]
## Hypotheses to test
- H1: [claim we suspect is true, and what would falsify it]
- H2: ...
## Sub-Questions (priority + dependencies)
1. [Critical] [gates others] — owner: <specialist>
2. [High] — owner: <specialist>
3. [Medium] — can run in parallel with 2
4. [Low] nice-to-have
## Team & ownership
- Surveyor: maps sources for Q1–Q3
- Domain specialist A: interprets Q1, Q2
- Methodologist: audits rigor of Q2 evidence
- Contrarian: challenges assumptions each loop
## Search strategy
- Per sub-question: 2–3 query variants (synonym / specificity ladder / source-type targeting)
- Source-quality bar: peer-reviewed > official report > reputable expert > news > blog
- Budget: aim for 10–20 searches per iteration; track count
## Convergence criteria (when are we "done"?)
- All sub-questions addressed
- Every load-bearing claim has 2+ independent sources
- Each contradiction either resolved or logged as unresolved
- Contrarian's open challenges answered or recorded
- Confidence meets target: [High/Medium] defined explicitly
## Assumptions to challenge
- A1: [thing the team currently takes for granted]
- A2: ...
Run the DAG from Phase 1, a layer at a time. For each specialist, require a tight, comparable report so integration is fast:
ROLE: <specialty>
FINDINGS: <what the sources actually say>
EVIDENCE: <key claims with citations — [n] referencing source-log>
CONFIDENCE: <High/Medium/Low — and why>
ASSUMPTIONS CHALLENGED: <which of A1..An this weakens or supports>
OPEN QUESTIONS: <what's still unknown or contradictory>
INTERFACES: <what it needs from / hands to other specialists>
Per-search process (applied by every specialist): extract key claims with attribution,
note source type and publication date, flag contradictions with existing findings. Track
in source-log.md:
## [Source N: Title]
- URL: [link]
- Type: [Academic/News/Official/Expert/Blog]
- Date: [publication date]
- Credibility: [High/Medium/Low]
- Key claims:
- "[claim]"
- Contradicts: [other source if applicable]
This is what makes it a team, not a pile of consultants.
Append to research-progress.md after every iteration:
## Iteration [N] — [date/time]
### What changed since last loop
- [finding added / claim sharpened / assumption falsified]
### Confidence deltas
- Claim X: Low → Medium (corroborated by [n])
- Claim Y: unchanged (still single source)
### Contradictions
- Open: [list] — resolution attempt: [query dispatched]
- Closed: [list] — how resolved
### Assumptions challenged
- A1: weakened by [evidence] — team now treats as [revised stance]
### Decision
- [Continue deeper on Q2] / [Pivot to Q4] / [Converge — criteria met]
Before synthesizing, verify against the plan's convergence criteria:
| Condition | Action | | -------------------------------------- | --------------------------------------- | | Sub-question unanswered | Next iteration, targeted specialist | | Load-bearing claim has single source | Search for corroboration | | Major contradiction unresolved | Search for resolution/context | | Contrarian challenge unaddressed | Force a response or log as open risk | | Convergence criteria met | Proceed to synthesis | | Iteration budget exhausted | Note limitations, proceed |
Recursive verification for high-stakes claims — don't trust a claim one hop from source:
Level 1: Source says X
Level 2: Source cites study Y → verify Y exists and says X
Level 3: Study methodology → is it rigorous? (Methodologist)
Status: [Confirmed/Partially Confirmed/Unverified/Contradicted]
Write research-report.md. Lead with conclusions, support with evidence, cite everything.
# [Research Question]
## Executive Summary
[2–3 sentence answer with confidence level]
## Key Findings
### Finding 1: [Statement]
[Evidence synthesis with citations]
- Source A reports... [1]
- Corroborated by... [2]
- However, Source C notes... [3]
## Analysis
[Interpretation connecting findings]
## Where the team disagreed and how I ruled
[The live tensions and your reasoning — the most valuable part]
## Assumptions challenged
[Which initial assumptions were falsified/revised and what changed]
## Limitations & Gaps
- [What couldn't be verified]
- [Areas needing more research]
- [Potential biases in available sources]
## Confidence Assessment
| Claim | Confidence | Basis |
| ------- | ---------- | ------------------------ |
| Claim 1 | High | 3+ independent sources |
| Claim 2 | Medium | 2 sources, some conflict |
## Sources
[Required — every source used, in citation order]
[1] Author/Publisher, "Title", Publication/Site, Date. URL
[2] ...
Before delivering the final report:
research-plan.md written with convergence criteriaresearch-progress.md updated each iterationresearch-report.md has a visible Sources section listing every sourceTechniques for breaking a complex question into searchable sub-questions and queries.
Level 0: "Best programming language for AI in 2025?"
Level 1: Performance · Ecosystem · Learning curve · Adoption
Level 2: Benchmarks · Library availability · Community size · ...
"How did the 2024 chip shortage affect EV prices?"
Hop 1: cause of shortage → Hop 2: chips used in EVs → Hop 3: % of EV cost
→ Hop 4: manufacturer response → Hop 5: actual 2024 price change
| Technique | Example | | -------------------- | --------------------------------------------------------------- | | Synonym expansion | "AI" → "machine learning", "deep learning" | | Specificity ladder | "AI healthcare" → "radiology AI FDA approved 2024" | | Source targeting | append "peer reviewed" / "government report" / "meta-analysis" | | Negation queries | also search the opposing view ("remote work disadvantages") |
Classify sub-questions: Prerequisite (A before B), Parallel (independent), Refinement (B narrows A), Validation (B checks A). Map before searching so layers run in the right order.
## Batch 1 (Foundation) — parallel
- Q1 context · Q2 current state · Q3 major players
## Batch 2 (Deep dive) — after Batch 1
- Q4 specific aspect from Q1 · Q5 follow-up on surprising Q2 result
## Batch 3 (Verification) — after Batch 2
- Q6 verify key claim · Q7 search for contradicting evidence
| Mistake | Fix | | --------------------- | ------------------------------------ | | Too many sub-questions | Limit to 5–7, prioritize | | Overlapping queries | De-duplicate before searching | | Missing negation | Always add opposing-view queries | | No dependency mapping | Map the DAG before searching |
Systematic claim verification. Drives Phase 3 (challenge) and Phase 5 (convergence).
DECISION ("enough evidence? what gaps?")
→ RETRIEVAL (refined queries on gaps)
→ VERIFICATION (score confidence, check source quality)
→ TERMINATION CHECK (threshold met OR budget out?)
No → loop back · Yes → synthesize
Score each source on Authority / Recency / Evidence / Bias / Corroboration (High=3 … Low=1).
Tier hierarchy (highest → lowest trust):
| Level | Criteria | Report as | | ----------- | ------------------------------------------------ | ---------------------- | | Very High | 3+ Tier-1, no contradictions, recent | Established fact | | High | 2+ reliable, minor contradictions resolved | State with confidence | | Medium | 1–2 sources or unresolved minor contradictions | State with caveat | | Low | Single source or major contradictions | Flag uncertainty | | Very Low | Weak source or strong contradictions | Consider excluding |
| Search again when… | Stop when… | | ------------------------------- | --------------------------------------- | | Sub-question unanswered | Confidence target met for all claims | | Key claim has single source | Search budget exhausted | | All sources same perspective | Diminishing returns (same results) | | Claim contradicted | Topic has limited available info | | Source credibility low | |
Confirmation bias (search opposites too) · authority fallacy (verify independently) · recency bias (balance with quality) · false balance (weight by credibility) · citation laundering (trace to original source).
assemble-a-team skillGoogle Workspace CLI for Gmail, Calendar, Drive, Contacts, Sheets, and Docs.
Manage Apple Notes via the `memo` CLI on macOS (create, view, edit, delete, search, move, and export notes). Use when a user asks OpenClaw to add a note, list notes, search notes, or manage note folders.
Work with Obsidian vaults (plain Markdown notes) and automate via obsidian-cli.
Use when you need to control Slack from OpenClaw via the slack tool, including reacting to messages or pinning/unpinning items in Slack channels or DMs.
Manage Apple Reminders via remindctl CLI (list, add, edit, complete, delete). Supports lists, date filters, and JSON/plain output.
Manage Trello boards, lists, and cards via the Trello REST API.
Category:productivity