跳到主要内容
知仓学习社ZHICANG

research-ideation

Generate structured research questions, hypotheses, and empirical strategies from a topic or dataset within the sewage/environmental economics space…

不碰外部(只输出文字)无严重或高危命中brycewang-stanford/Auto-Empirical-Research-Skills

它会碰到什么

扫了多少1 个文本文件,4 KB
它会碰到什么不碰外部(只输出文字)
命中总数0 处
命中统计严重 0 · 高 0 · 中 0 · 低 0

这一栏是扫描器报的事实,不是结论。命中多不等于有毒(安全工具、规则库、示例脚本本来就会包含危险写法),命中少也不等于干净。它和你手上的凭据、文件、网络有什么关系,需要你自己看。

技能内容

Research Ideation

Generate structured research questions, testable hypotheses, and empirical strategies from a topic, phenomenon, or dataset.

Input: $ARGUMENTS — a topic, phenomenon, dataset description, or extensions to brainstorm extensions of the current paper.


Project-Specific Context

Available Data

  • EDM data (2021-2024+): Event-level sewage spill records for all overflow sites in England
  • Land Registry: Universe of property transactions (prices, dates, postcodes)
  • Zoopla: Rental listings with prices and characteristics
  • Met Office: Daily rainfall at LSOA level
  • River networks: PostGIS topology of English waterways
  • Annual returns: Water company infrastructure data (sewer capacity, population served)
  • Geographic boundaries: LSOAs, MSOAs, constituencies, water company areas

Existing Identification Strategies

Hedonic, repeat sales, long difference, DiD with media coverage, upstream/downstream, dry spills, hydraulic capacity IV

Related Literature Areas

Environmental disamenity capitalisation, water quality and property values, UK water industry regulation, infrastructure and house prices, information and price discovery


Workflow

Step 1: Understand the Input

Read $ARGUMENTS and any referenced files.

  • If extensions: read current manuscript sections and analysis scripts to understand what's been done
  • If a topic: ground it in the available data and methods

Step 2: Generate 3-5 Research Questions

Order from descriptive to causal:

  • Descriptive: What are the patterns? (e.g. "How do spill frequencies vary across water companies?")
  • Correlational: What factors are associated? (e.g. "Is sewer age correlated with spill frequency after controlling for rainfall?")
  • Causal: What is the effect? (e.g. "What is the causal effect of media coverage of sewage spills on house prices?")
  • Mechanism: Why does the effect exist? (e.g. "Do dry spills affect prices through stigma or physical damage?")
  • Policy: What are the implications? (e.g. "Would mandated real-time spill disclosure affect property markets?")

Step 3: Develop Each Question

For each:

  • Hypothesis: Testable prediction with expected sign/magnitude
  • Identification strategy: How to establish causality
  • Data requirements: What data is needed — is it available in the project?
  • Key assumptions: What must hold
  • Potential pitfalls: Threats to identification
  • Related literature: 2-3 papers using similar approaches

Step 4: Rank by Feasibility and Contribution

| RQ | Feasibility | Contribution | Priority |

|----|-------------|-------------|----------|

| 1 | High/Med/Low | High/Med/Low | ... |

Step 5: Save Output

# Research Ideation: [Topic]
**Date:** YYYY-MM-DD

## Overview
[1-2 paragraphs situating the topic]

## Research Questions

### RQ1: [Question] (Feasibility: High/Medium/Low)
**Type:** Descriptive / Correlational / Causal / Mechanism / Policy
**Hypothesis:** [Prediction]
**Identification:** [Method + key assumption]
**Data:** [What's needed and what's available]
**Pitfalls:** [Top 2 threats]
**Related work:** [2-3 papers]

## Ranking
[Table]

## Suggested Next Steps
1. [Most promising direction]
2. [Data to obtain or analysis to run]
3. [Literature to review]

Save to output/log/research_ideation_[topic].md.


Principles

  • Be creative but grounded. Every suggestion must be empirically feasible with available or obtainable data.
  • Think like a referee. For each causal question, immediately identify the identification challenge.
  • Leverage existing infrastructure. The project already has spatial matching, spill aggregation, and river network tools — build on them.
  • Consider data availability. Flag when data is already available vs needs to be obtained.
  • Extensions are valuable. For extensions, focus on questions that use the same data infrastructure but ask different questions.

想直接用这个技能?

本站把开放许可(MIT / Apache 等)的技能按仓库打包整理到网盘,点一下转存到你自己的网盘,不用一个个从 GitHub 拉。许可未声明的技能只给原始仓库链接,不打包。

同名技能的其他版本

有 7 个不同仓库或目录里都有叫 research-ideation 的技能。它们内容并不相同,别混用: