{"repo":"AI272/speaker","free":true,"listed":false,"github":"https://github.com/AI272/speaker","clone":"git clone https://github.com/AI272/speaker.git","description":"Speaker is a Codex skill project for academic presentations: read real.pptx, combine text extraction, PPTX structure parsing, page-by-page rendering, OCR, and visual review to generate page-by-page speaker notes, and write a clean version of the lecture into the PowerPoint comment area.","language":"Python","stars":421,"topics":["ai","claude-code","claude-code-skill","codex","codex-skill","skill"],"license":null,"category":"media-processing","readme_excerpt":"speaker 中文说明 speaker is a Codex skill project for academic presentations. It reads a real .pptx , combines text extraction, PPTX structure inspection, slide rendering, OCR, and vision review, then generates grounded speaker notes and injects the clean script into PowerPoint's speaker notes pane. Current skill package: speaker-v8.skill Internal skill name: ppt-speech-writer What's New in v0.8 This release focuses on tighter, evidence-grounded output and realistic speech pacing. - Pause-aware pacing model. Speech length is now budgeted with a deterministic model (English 110 wpm, Chinese 165 characters/min) that reserves time for slide transitions and [PAUSE] marks. This fixes the previous problem where a \"15-minute\" script ran to 1,800 words and overran to 30 minutes; a 15-minute talk now targets roughly 1,300–1,400 words. Both Chinese and English notes are tuned per slide. - Per-slide word budget. SKILL.md now computes and records a per-slide budget so each slide stays within the overall time target, and the timing table reports words/characters, budget, and pauses with a TOTAL row. - Glossary toggle. A new on/off switch (default on) can skip the \"Key Parameters And Methods\" glossary table end to end. - Compact extraction ( read slides.py --mode compact ). Drops redundant raw OOXML dumps and non-visual geometry while keeping picture bounding boxes, producing much smaller intermediate JSON. - Region-scoped OCR ( visual inventory.py --ocr-scope image-regions ). OCR runs only on","default_branch":null,"files":null,"tree":[],"storefront":"/r/AI272","claimed":false,"request_supported":{"post":"https://gitbuyer.com/r/AI272/speaker/request-supported","requests":0},"note":"indexed from public GitHub; nothing is for sale on this page. Clone it from GitHub. Paid listings live at /search."}