A code-driven presentation generation framework. 像构建软件工程一样生成演示文稿。
-
Updated
Jun 8, 2026 - Python
A code-driven presentation generation framework. 像构建软件工程一样生成演示文稿。
A Codex plugin and local MCP runtime for context-bounded image generation and inspection.
AI가 청중에 맞는 HTML 발표자료와 발표 노트를 만들고 Chromium 기반 시각 검증까지 수행하도록 돕는 스킬·플러그인
AI Computer Use for Claude Code — The open-source alternative to OpenAI Codex's playwright-interactive. Dual-engine: Win32 API + Playwright. Control WeChat, DingTalk, Feishu, QQ, Slack, Teams, and any web/Electron app. Automated QA, viewport testing, visual feedback loops.
AI skill for mirroring public webpages into local pixel-perfect replicas with captured assets, visual QA, and rebranding inventories. 像素级复刻公开网站,内置抓取资源、视觉QA能力和品牌替换清单。
Native AppKit/UIKit-first Pixiv client for macOS, with SwiftUI glue and visual QA.
Claude Code plugin that audits your project before you deploy. 40+ checks across security, visual QA, code quality, testing, error handling, build config, and performance. One command, structured report, actionable fixes. Auto-detects your stack.
Rebuild reference images as editable Microsoft Visio diagrams using Codex, MCP, COM/ShapeSheet, resumable rendering, and visual QA.
Six-skill spec-driven engineering pod for Claude Code — scaffold specs, research, discover direction, generate design variants, implement, and visually verify, with hard handoffs that prevent silent assumptions
Reasoning-based, vectorless RAG over a large document using a hierarchical tree (PageIndex) and a Vision-Language Model (Llama 4 Scout), no embeddings, no vector store, no text chunking.
A source-aware Draw.io skill for Codex with collaborative intake, semantic icons, and visual QA.
A visual-first web design Skill: explore, approve, decompose, build, and refine.
Codex skills for professional Word and PDF publishing with a reproducible TAIZHOU open-source case.
Production-grade Codex v2 pet workflow with reactions, resumable generation, repair, QA, and cache-safe packaging.
Agent skill for inspecting, editing, and visually QAing Figma files. Works with any Agent Skills host. Requires Figma Console MCP.
Visual QA, browser automation, accessibility auditing, screenshot diffing, and MCP tooling for coding agents.
Visual QA CLI for AI-built frontends that renders pages, finds design and accessibility issues, and writes agent-ready fix prompts.
A FastAPI-based backend utilizing Google Cloud Vision API to provide intelligent, real-time visual question-answering via REST endpoints and WebSockets.
VisualJudge: continuous browser walkthrough capture plus Gemini-backed visual QA for web app UI/UX gaps.
Agent skill for LaTeX-based bilingual PDF translation with page-level visual QA.
Add a description, image, and links to the visual-qa topic page so that developers can more easily learn about it.
To associate your repository with the visual-qa topic, visit your repo's landing page and select "manage topics."