xcrawl-search
Use this skill for XCrawl search tasks, including keyword search request design, location and language controls, result analysis, and follow-up crawl or scrape planning.
Browse reusable Agent Skills, each with a clear purpose and practical guidance.
Use this skill for XCrawl search tasks, including keyword search request design, location and language controls, result analysis, and follow-up crawl or scrape planning.
Goal-driven worker for exploring an unfamiliar target.
Run a fast 2-agent pre-submission check for an economics paper — focuses on contribution, identification, and causal overclaiming. Completes in ~1 minute.
Use when a task references a specific project, person, or note the user has saved (e.g. "about my project X", "what did I say about Y", "who is Z to me"). The user's overall voice/tone/personality is already loaded into the system prompt automatically and does NOT require this skill — only reach for read-wiki when the task names a specific entity that lives in the wiki. Read-only counterpart to the `save-wiki` skill (which writes to the wiki).
Verify scholarly references and claim-citation support for Light stage 10. Use when auditing a manuscript, claim map, bibliography, DOI/arXiv/PMID/ISBN/URL, BibTeX/CSL, citekeys, chimeric or fabricated citations, retraction/correction alerts, or preparing a canonical citation registry for typesetting. Builds provenance-preserving inventories, confirms metadata with independent authoritative sources, distinguishes CONFIRMED/CONFIRMED-MISSING/UNAVAILABLE/UNRESOLVED, records Crossref update direction, and emits the citation gate plus delivery artifacts.
Light 科研主线第 2 步·数据工程:**找得到且用得起的数据**(来源/许可/版本/大小/split)+ **提 idea 前先判数据可行性** (数据够不够支撑研究/统计功效)+ **防数据泄漏**(顶会拒稿高频雷)。何时用:用户要找/选/下载公开数据集,或给了数据问 "能不能做研究/够不够/质量行不行" / 要清洗·处理缺失异常·特征工程·划分数据集·数据增强 / 自建数据集(采集·标注规范· 隐私合规·发布) / 怀疑训练测试串了数据(泄漏) / 提 idea 前评数据基础。 触发词:数据够不够 / 数据可行性 / 数据质量 / 数据泄漏 / 防穿越 / train test 重叠 / 怎么划分 / 交叉验证 / 标注规范 / 一致性 IAA / 自建数据集 / 样本量够吗 / 统计功效 / 找数据集 / 数据许可 / dataset search / data leakage / feasibility / data split / annotation。核心纪律: **数据泄漏 = critical 一票否决**(标准化早于划分/时序穿越/实体重叠/目标编码穿越);**数据不足以支撑 idea = 拦在 idea 前(回边 2⊣3,补数据/改 idea)**;功效是经验阈值非 power analysis;泄漏检测是启发式有边界,不吹"查全了"。
Light 科研主线第 4 步·审 idea:以**顶会审稿人标准严审** idea,**撞车/无创新 fatal flaw 一票否决**(critical 门), 逼出真能发表的 idea。何时用:用户问"这 idea 行不行/够不够新/能不能发""帮我严审/挑刺/找致命问题" / idea 定稿前把关 / 收到 idea-generation 的候选要审 / 怀疑撞车(被人做过)。触发词:审 idea / 评审 / 严审 / 挑刺 / 这 idea 行不行 / 够不够新 / 创新性 / 撞车 / 被做过了吗 / 致命问题 / 能投顶会吗 / 拒稿风险 / critique my idea / review this idea / is this novel / fatal flaw / 一票否决。核心纪律:**撞车/无创新的 critical 一票否决在本技能**(不被其他高维度平均救回); **硬性反谄媚**(不被作者反复反驳顺从放行弱 idea);撞车判定**target/background 可追溯分解**非"感觉像";judge 不靠裸 自评(用可计算否决闸门 + 密度先验 + pairwise)。**消费上游 idea-generation 的撞车 findings + facet 槽位**。
Light 科研主线第 3 步·提 idea:从模糊方向/数据/文献**结构化发散**(激发算子系统生成,不是泛泛头脑风暴) → 产**值得做且做得成的分层候选 idea**(moonshot 冲刺/solid 稳妥/safe 保底),每个必答**为什么值得做·创新点· 比现有强在哪·解决什么具体问题·能投什么层次**,且**提出时就自带撞车前置自查**(最像的前作+delta,吃上游 literature-search 领域地图)。何时用:用户问"这个方向/数据能做什么" / 要创新点·研究思路·选题·突破口 / 帮我想 idea / brainstorm research ideas / 这 idea 行不行(先生成再送 idea-critique 严审)。触发词:提 idea / 想 idea / 创新点 / 研究思路 / 选题 / 突破口 / 差异化 / 这个方向能做什么 / 有什么可做的 / brainstorm / research idea / ideation / 新点子 / 立项。核心纪律:**不下 novel/无创新的最终判决**(那是 idea-critique 的 critical 门,生成端只产撞车 warn 信号);用**证据根→机制/假设 delta→信息增益→最小判别实验**谱系约束候选; 过 **innovation_engine 反拼接门**,把原创来源分型(新问题/新机制/新测量/新数据/新理论/跨域迁移/工程增量)与 claim 强度绑定; 数量/七角度只作 advisory,机制多样性才是硬门。
Light 科研主线第 1 步·文献调研:在线多源检索(OpenAlex/arXiv/Crossref/Europe PMC/DOAJ,全程免 key)→ 产出**不是论文列表,是可直接喂 idea 的"领域地图"**:近三年前沿 + 经典奠基 + 跨领域方法移植**三层分别 检索分别排序**,合成研究脉络 + 方法谱系 + 未解问题地图,揪出**最像你设想的那一篇**→喂 idea-critique 撞车预警,诚实标检索覆盖度。何时用:调研某方向 / 找综述 / 了解某领域有哪些工作 / 提 idea 前摸清前作 / 收集中英文论文·专利·标准·数据集·开源·竞赛方案·行业报告。触发词:文献调研 / 文献综述 / 调研方向 / 领域地图 / 研究现状 / 有哪些工作 / 相关工作 / related work / literature review / survey / 前沿 / 综述 / 找论文 / 撞车 / 新颖性摸底 / 跨领域 / 检索式。核心纪律:不臆造 DOI/被引(查不到写 unknown);三层分别排序 不一锅 relevance;撞车只产**信号**不下 novel 判决(归 idea-critique);覆盖度按真实 HTTP 码诚实标。
Light 科研诚信与伦理全生命周期常驻红线门:在研究分诊、审批、采集、变更、分析、投稿、发布与出版后检查 学术不端/数据造假/统计自洽/结论夸大/幻觉与撤稿 引用/自我抄袭/隐私/版权/署名与 AI 披露/软著专利权属/论文工厂洗稿等风险,把"别造假别夸大"从口头建议 落成**可机检、可阻断、可被总控 run_checkpoint 聚合的机读门**(产 light.findings.v1,Critical fail → exit 1)。 何时用:投稿/camera-ready 前、数据或代码发布前、软著提交前、涉人或涉动物实验设计时——这些是硬闸门; 以及任意写作/分析/引用任务的后台诚信扫描。触发词:伦理 / 诚信 / 学术不端 / 造假 / p-hacking / 夸大 / SOTA / 撤稿 / 查重 / 抄袭 / 隐私 / IRB / 知情同意 / 署名 / AI 披露 / 论文工厂 / 软著权属。核心纪律:**信号≠定罪**, AI 不能自评 → 一律"机读门 + 人工复核";查不到写"待核查/UNRESOLVED",绝不编造;全程在线核实、零本地知识库、零付费 key。