科研技能库/幻灯片全流程编排器
科研效率
未发现用户侧风险

幻灯片全流程编排器

用于从零开始制作完整演示文稿,包括规划、设计幻灯片、编辑和导出。推荐导出为 PDF 和每页 PNG;PPTX/Figma 导出为实验性功能,不稳定。

文件预览

2 个文件
references
SKILL.md
5.4 KB · 可预览
---
name: slides-grab
description: End-to-end presentation workflow for Codex. Use when making a full presentation from scratch — planning, designing slides, editing, and exporting. PDF and per-slide PNG are preferred; PPTX/Figma export is experimental / unstable.
metadata:
  short-description: Full pipeline from topic to PDF/PNG + experimental / unstable PPTX/Figma export
---

# slides-grab Skill (Codex) - Full Workflow Orchestrator

Guides you through the complete presentation pipeline from topic to exported file.

---

## Workflow

### Stage 1 — Plan

Use the installed **slides-grab-plan** skill.

1. Take user's topic, audience, and tone.
2. **Style selection (mandatory before outline):** Run `slides-grab list-styles`, analyze the topic/tone, and shortlist 2–3 bundled styles that fit. Present the shortlist with reasons. Optionally offer `slides-grab preview-styles` for visual preview. If none of the 35 bundled styles fit, propose a fully custom visual direction. **Get explicit style approval before writing the outline.**
3. Create `slide-outline.md` with the chosen style ID in the meta section (`style: <id>`).
4. Present outline to user.
5. Revise until user explicitly approves.

**Do not proceed to Stage 2 without approval of both style and outline.**

### Stage 2 — Design

Use the installed **slides-grab-design** skill.

1. Read approved `slide-outline.md` and apply the style specified in its meta section (`style: <id>`). Do not re-open style selection — the style was already approved in Stage 1.
3. Generate `slide-*.html` files in the slides workspace (default: `slides/`).
4. Run validation: `slides-grab validate --slides-dir <path>`
5. If validation fails, automatically fix the slide HTML/CSS until validation passes.
6. For bespoke slide imagery, use `slides-grab image --prompt "<prompt>" --slides-dir <path>` so the default god-tibo-imagen provider (reuses local Codex ChatGPT login — no API key required) saves a local asset under `<slides-dir>/assets/`.
7. For complex diagrams (architecture, workflows, relationship maps, multi-node concepts), prefer `tldraw` over hand-built HTML/CSS diagrams. Render the asset with `slides-grab tldraw`, store it under `<slides-dir>/assets/`, and place it in the slide with a normal `<img>`.
8. Keep local videos under `<slides-dir>/assets/`, prefer `poster="./assets/<file>"` thumbnails, and use `slides-grab fetch-video --url <youtube-url> --slides-dir <path>` (or `yt-dlp` directly) when the source starts on a supported web page.
9. The default provider, god-tibo-imagen, reuses the local Codex ChatGPT login (`~/.codex/auth.json`) — run `codex login` once; no API key required. ⚠️ god-tibo-imagen uses an unsupported private Codex backend that may break without notice. Optional alternatives: `--provider codex` (Codex/OpenAI gpt-image-2 via `OPENAI_API_KEY`; maps `--aspect-ratio` to the nearest supported OpenAI image size; `--image-size 2K|4K` is Nano Banana-only) or `--provider nano-banana` (Google `gemini-3-pro-image-preview` via `GOOGLE_API_KEY` or `GEMINI_API_KEY`; supports `--image-size 2K|4K`). If credentials are unavailable, fall back to web search/download into `<slides-dir>/assets/`.
10. Launch the interactive editor for review: `slides-grab edit --slides-dir <path>`
11. Revise slides based on user feedback via the editor, then re-run validation after each edit round.
12. When the user confirms editing is complete, suggest next steps: build the viewer (`slides-grab build-viewer --slides-dir <path>`) for a final preview, or proceed directly to Stage 3 for PDF/PPTX export.

**Do not proceed to Stage 3 without approval.**

### Stage 3 — Export

Use the installed **slides-grab-export** skill.

1. Confirm user wants conversion.
2. Pick the primary target:
   - Card-news / Instagram-style decks → `slides-grab png --slides-dir <path> --slide-mode card-news --resolution 2160p` (see `slides-grab-card-news`).
   - Widescreen decks → `slides-grab pdf --slides-dir <path> --output <name>.pdf`.
3. Per-slide PNG (any mode): `slides-grab png --slides-dir <path> --output-dir <path>/out-png --resolution 2160p`.
4. PPTX (optional, **experimental / unstable**): `slides-grab convert --slides-dir <path> --output <name>.pptx`.
5. Figma-importable PPTX (optional, **experimental / unstable**): `slides-grab figma --slides-dir <path> --output <name>-figma.pptx`.
6. Report results.

---

## Rules

1. **Always follow the stage order**: Plan → Design → Export.
2. **Get explicit user approval** before advancing to the next stage.
3. **Read each stage's SKILL.md** for detailed rules — this skill only orchestrates.
4. **Use `decks/<deck-name>/`** as the slides workspace for multi-deck projects.
5. **Call out export risk clearly**: PPTX and Figma export are experimental / unstable and must be described as best-effort output.
6. Use the stage skills as the source of truth for plan, design, and export rules.
7. When a slide needs a complex diagram, default to a `tldraw`-generated asset unless the user explicitly asks for a different approach.
8. When a slide needs bespoke imagery, prefer the default god-tibo-imagen provider via `slides-grab image` (reuses local Codex ChatGPT login — no API key required) and keep the saved asset local under `<slides-dir>/assets/`.

## Reference
- `references/presentation-workflow-reference.md` — archived end-to-end workflow guidance from the legacy skill set

SKILL.md

元数据
nameslides-grab
descriptionEnd-to-end presentation workflow for Codex. Use when making a full presentation from scratch — planning, designing slides, editing, and exporting. PDF and per-slide PNG are preferred; PPTX/Figma export is experimental / unstable.
metadata{ "short-description": "Full pipeline from topic to PDF/PNG + experimental / unstable PPTX/Figma export" }

slides-grab 技能(Codex)- 全流程编排器

引导您完成从主题到导出文件的完整演示文稿流程。


工作流程

阶段 1 — 规划

使用已安装的 slides-grab-plan 技能。

  1. 接受用户的主题、受众和语气。
  2. 样式选择(在大纲之前必须完成): 运行 slides-grab list-styles,分析主题/语气,并筛选出 2–3 个适合的内置样式。展示筛选列表并说明理由。可选地,使用 slides-grab preview-styles 进行视觉预览。如果 35 个内置样式都不适合,提出完全自定义的视觉方向。在编写大纲之前获得明确的样式批准。
  3. 在元数据部分(style: <id>)包含所选样式 ID 创建 slide-outline.md。
  4. 向用户展示大纲。
  5. 修改直到用户明确批准。

在样式和大纲均未获得批准之前,不要进入阶段 2。

阶段 2 — 设计

使用已安装的 slides-grab-design 技能。

  1. 阅读已批准的 slide-outline.md 并应用其元数据部分指定的样式(style: <id>)。不要重新打开样式选择——样式已在阶段 1 批准。
  2. 在幻灯片工作区(默认:slides/)中生成 slide-*.html 文件。
  3. 运行验证:slides-grab validate --slides-dir <path>
  4. 如果验证失败,自动修复幻灯片的 HTML/CSS 直到验证通过。
  5. 对于定制的幻灯片图像,使用 slides-grab image --prompt "<prompt>" --slides-dir <path>,这样默认的 god-tibo-imagen 提供者(重用本地 Codex ChatGPT 登录——无需 API 密钥)会将本地资产保存在 <slides-dir>/assets/ 下。
  6. 对于复杂图表(架构、工作流程、关系图、多节点概念),优先使用 tldraw 而非手动构建 HTML/CSS 图表。使用 slides-grab tldraw 渲染图表资产,存储在 <slides-dir>/assets/ 下,然后通过普通的 <img> 标签放置到幻灯片中。
  7. 将本地视频保存在 <slides-dir>/assets/ 下,优先使用 poster="./assets/<file>" 缩略图,当源网页是受支持的网页时,使用 slides-grab fetch-video --url <youtube-url> --slides-dir <path>(或直接使用 yt-dlp)下载视频。
  8. 默认提供者 god-tibo-imagen 重用本地 Codex ChatGPT 登录(~/.codex/auth.json)——运行一次 codex login;无需 API 密钥。 ⚠️ god-tibo-imagen 使用了不受支持的私有 Codex 后端,可能会在没有通知的情况下中断。可选替代方案:--provider codex(Codex/OpenAI gpt-image-2 通过 OPENAI_API_KEY;将 --aspect-ratio 映射到最近支持的 OpenAI 图片尺寸;--image-size 2K|4K 仅限 Nano Banana)或 --provider nano-banana(Google gemini-3-pro-image-preview 通过 GOOGLE_API_KEY 或 GEMINI_API_KEY;支持 --image-size 2K|4K)。如果凭据不可用,回退到网络搜索/下载,放入 <slides-dir>/assets/。
  9. 启动交互式编辑器进行审核:slides-grab edit --slides-dir <path>
  10. 根据用户反馈通过编辑器修订幻灯片,之后每轮编辑后重新运行验证。
  11. 当用户确认编辑完成时,建议下一步:构建查看器 (slides-grab build-viewer --slides-dir <path>) 进行最终预览,或直接进入阶段 3 进行 PDF/PPTX 导出。

在未获得批准之前,不要进入阶段 3。

阶段 3 — 导出

使用已安装的 slides-grab-export 技能。

  1. 确认用户想要转换。
  2. 选择主要目标:
    • 卡片新闻/Instagram 风格幻灯片 → slides-grab png --slides-dir <path> --slide-mode card-news --resolution 2160p (参见 slides-grab-card-news)。
    • 宽屏幻灯片 → slides-grab pdf --slides-dir <path> --output <name>.pdf。
  3. 每页 PNG(任何模式):slides-grab png --slides-dir <path> --output-dir <path>/out-png --resolution 2160p。
  4. PPTX(可选,实验性/不稳定):slides-grab convert --slides-dir <path> --output <name>.pptx。
  5. 可导入 Figma 的 PPTX(可选,实验性/不稳定):slides-grab figma --slides-dir <path> --output <name>-figma.pptx。
  6. 报告结果。

规则

  1. 始终遵循阶段顺序:规划 → 设计 → 导出。
  2. 在进入下一阶段前获得用户的明确批准。
  3. 阅读每个阶段的 SKILL.md 以获取详细规则——此技能仅负责编排。
  4. 对于多幻灯片项目,使用 decks/<deck-name>/ 作为幻灯片工作区。
  5. 明确说明导出风险:PPTX 和 Figma 导出为实验性/不稳定,必须描述为尽力而为的输出。
  6. 使用各阶段技能作为规划、设计和导出规则的权威来源。
  7. 当幻灯片需要复杂图表时,默认使用 tldraw 生成的资产,除非用户明确要求其他方法。
  8. 当幻灯片需要定制图像时,优先通过 slides-grab image 使用默认的 god-tibo-imagen 提供者(重用本地 Codex ChatGPT 登录——无需 API 密钥),并将生成的资产保存在 <slides-dir>/assets/ 下。

参考

  • references/presentation-workflow-reference.md — 存档的端到端工作流指南,来自旧版技能集