Source profileQuality 78/100Review permissions

github/awesome-copilot/skills/convert-word-to-md/SKILL.md

convert-word-to-md

Converts Word (.docx) documents into Markdown so their contents can be accurately analyzed, summarized, searched, or extracted from. Use this skill whenever the user shares, references, or asks about a .docx file — even if they don't say "convert" or "markdown" explicitly. This includes requests to "read", "summarize", "review", "extract data from", "compare", or "analyze" a Word document, resume, report, contract, or proposal. Always run the bundled conversion script to produce Markdown first;

Source repository stars
37,126
Declared platforms
0
Static risk flags
2
Last source update
2026-07-28
Source checked
2026-07-28

Decision brief

What it does—and where it fits

Converts Word (. docx) documents into Markdown so their contents can be accurately analyzed, summarized, searched, or extracted from.

Best for

  • convert-pdf-to-md for any .pdf files
  • convert-excel-to-md for any .xlsx files

Not for

  • Tasks that require unconfirmed production actions or broad system permissions.
  • Environments where the pinned source and install steps cannot be inspected.

Compatibility matrix

Platform support, with evidence labels

PlatformStatusEvidenceWhat to check
CodexNot declaredNo explicit evidencePortability before use
Claude CodeNot declaredNo explicit evidencePortability before use
CursorNot declaredNo explicit evidencePortability before use
Gemini CLINot declaredNo explicit evidencePortability before use
Open the compatibility checker

Installation

Inspect first. Install second.

The source command is displayed only when detected. A safe inspection prompt is always available so your agent can explain every action before execution.

Source-detected install commandSource
npx skills add https://github.com/github/awesome-copilot --skill "skills/convert-word-to-md"
Safe inspection promptEditorial

Inspect the Agent Skill "convert-word-to-md" from https://github.com/github/awesome-copilot/blob/9933dcad5be5caeb288cebcd370eeeb2fc2f1685/skills/convert-word-to-md/SKILL.md at commit 9933dcad5be5caeb288cebcd370eeeb2fc2f1685. List every install step, command, network request, credential, file read/write, external action, and rollback step. Explain whether it fits my task. Do not install or execute anything until I approve.

Workflow

What the source asks the agent to do

  1. 01

    Setup (once per environment)

    Before the first conversion in a given environment, follow references/setup.md step by step to ensure Python, pip, and the markitdown package are installed. Do this proactively rather than guessing whether the environment is ready — the script itself will also fail with a clear…

    Before the first conversion in a given environment, follow references/setup.md step by step to ensure Python, pip, and the markitdown package are installed. Do this proactively rather than guessing whether the environme…
  2. 02

    Usage

    The conversion script lives at scripts/convertwordtomd.py.

    The conversion script lives at scripts/convertwordtomd.py.Output structure: MarkItDown embeds images as a truncated data:image/png;base64... URI placeholder (not real image data), so the script extracts real images directly from the .docx and writes a self-contained folder per…If the document has no embedded images, no img/ folder is created.
  3. 03

    When to use this skill

    Trigger this skill any time there is a .docx file that needs to be understood or processed — for example, a user attaches a Word document and asks questions about it, wants a summary, wants specific data pulled out, or wants multiple Word documents in a folder processed together…

    convert-pdf-to-md for any .pdf filesconvert-excel-to-md for any .xlsx filesTrigger this skill any time there is a .docx file that needs to be understood or processed — for example, a user attaches a Word document and asks questions about it, wants a summary, wants specific data pulled out, or…
  4. 04

    Windows

    python scripts\convertwordtomd.py "C:\path\to\document.docx" bash

    python scripts\convertwordtomd.py "C:\path\to\document.docx" bash
  5. 05

    macOS / Linux

    python scripts/convertwordtomd.py "/path/to/document.docx" powershell python scripts\convertwordtomd.py "C:\path\to\document.docx" -o "C:\path\to\outputfolder" powershell python scripts\convertwordtomd.py "C:\path\to\folder" powershell python scripts\convertwordtomd.py "C:\path\…

    python scripts/convertwordtomd.py "/path/to/document.docx" powershell python scripts\convertwordtomd.py "C:\path\to\document.docx" -o "C:\path\to\outputfolder" powershell python scripts\convertwordtomd.py "C:\path\to\fo…Each .docx found gets its own \ output folder next to it by default. Pass -o "C:\path\to\outputparent" to collect all the generated \ folders under a separate parent directory instead (subfolder structure is preserved w…After conversion, read the resulting .md file(s) to perform the actual analysis the user asked for — the script's job is only to produce accurate Markdown (and images), not to interpret the content.

Permission review

Static risk signals and limitations

Reads files

low · line 11

The documentation asks the agent to read local files, directories, or repositories.

skill rather than trying to open or parse the file directly.

Runs scripts

medium · line 61

The documentation asks the agent to run terminal commands or scripts.

python scripts\convert_word_to_md.py "C:\path\to\document.docx"

Runs scripts

medium · line 66

The documentation asks the agent to run terminal commands or scripts.

python scripts/convert_word_to_md.py "/path/to/document.docx"

Reads files

low · line 94

The documentation asks the agent to read local files, directories, or repositories.

After conversion, read the resulting `.md` file(s) to perform the actual

Evidence record

Why each signal appears

EvidenceSourceComputedTestedEditorial
SignalValueEvidence typeMeaning
Quality score78/100ComputedDocumentation, specificity, maintenance, and trust rules
Repository stars37,126SourceRepository attention, not individual Skill quality
Compatibility0 platformsSourceDeclared in the catalog source record
Usage guideautomated source guideEditorialGenerated or reviewed according to the visible evidence level

Pinned source

Provenance and original SKILL.md

Repository
github/awesome-copilot
Skill path
skills/convert-word-to-md/SKILL.md
Commit
9933dcad5be5caeb288cebcd370eeeb2fc2f1685
License
MIT
Collected
2026-07-28
Default branch
main
View the original SKILL.md

Convert Word to Markdown

When to use this skill

Trigger this skill any time there is a .docx file that needs to be understood or processed — for example, a user attaches a Word document and asks questions about it, wants a summary, wants specific data pulled out, or wants multiple Word documents in a folder processed together. Word's native .docx format is a zipped XML bundle that is not reliably readable as plain text, so always convert it to Markdown first using the script in this skill rather than trying to open or parse the file directly.

This skill only supports .docx. If asked to convert a legacy .doc file, tell the user it isn't supported and ask them to re-save it as .docx (Word: File > Save As > Word Document (.docx)) first.

Mixed file types: When the user references a folder or set of documents containing multiple supported file types (.pdf, .docx, .xlsx), this skill handles only .docx files. The agent MUST also invoke the sibling skills in parallel:

  • convert-pdf-to-md for any .pdf files
  • convert-excel-to-md for any .xlsx files

Never process a folder and silently skip a supported file type. All three skills must be invoked together when mixed types are present.

Setup (once per environment)

Before the first conversion in a given environment, follow references/setup.md step by step to ensure Python, pip, and the markitdown package are installed. Do this proactively rather than guessing whether the environment is ready — the script itself will also fail with a clear pointer back to that file if markitdown turns out to be missing, so it's safe to just try the conversion first if you're reasonably confident setup was already done.

Usage

The conversion script lives at scripts/convert_word_to_md.py.

Output structure: MarkItDown embeds images as a truncated data:image/png;base64... URI placeholder (not real image data), so the script extracts real images directly from the .docx and writes a self-contained folder per document instead of a single loose .md file:

<name>/
    img/
        img001.<ext>
        img002.<ext>
        ...
    <name>.md          (image references are relative: img/imgNNN.ext)

If the document has no embedded images, no img/ folder is created.

Single file:

# Windows
python scripts\convert_word_to_md.py "C:\path\to\document.docx"
# macOS / Linux
python scripts/convert_word_to_md.py "/path/to/document.docx"

This creates a document\ folder next to the source file (containing document.md and, if present, document\img\). To control the destination folder explicitly:

python scripts\convert_word_to_md.py "C:\path\to\document.docx" -o "C:\path\to\output_folder"

A folder of Word documents (batch mode):

python scripts\convert_word_to_md.py "C:\path\to\folder"

Add --recursive to also include subfolders:

python scripts\convert_word_to_md.py "C:\path\to\folder" --recursive

Each .docx found gets its own <name>\ output folder next to it by default. Pass -o "C:\path\to\output_parent" to collect all the generated <name>\ folders under a separate parent directory instead (subfolder structure is preserved when combined with --recursive).

After conversion, read the resulting .md file(s) to perform the actual analysis the user asked for — the script's job is only to produce accurate Markdown (and images), not to interpret the content.

Deciding where output goes

Default — always output next to the source file. The <name>/ folder is created in the same directory as the source .docx. This is the required default for every case. Do NOT override it unless the user explicitly asks for a different location.

Only use -o when the user explicitly provides an output path (e.g., "save the output to C:\output", "put the results in D:\work"). Do NOT pass -o based on the agent's current working directory, the session state folder, or any implied location.

If the source file path cannot be fully resolved — for example, the user provides only a filename with no directory, or the path is ambiguous — use ask_user to confirm the full absolute path before running the conversion. Never guess or assume the directory.

Troubleshooting

SymptomLikely causeFix
ModuleNotFoundError: No module named 'markitdown' / exit code 2MarkItDown not installedFollow references/setup.md
ERROR: Unsupported file type '.doc' / exit code 3Legacy .doc, not .docxAsk the user to re-save as .docx
ERROR: Input path not found / exit code 3Wrong path, or file movedConfirm the correct path with the user
FAILED <file> -> ... in batch outputThat specific file is corrupt, password-protected, or otherwise unreadableReport which file(s) failed; other files in the batch still succeed
NOTE: skipped N non-.docx file(s)Folder contains non-Word filesExpected — those files are intentionally ignored
WARNING: found N image placeholder(s) ... but extracted M image file(s)Mismatch between MarkItDown's placeholder count and images found in word/media/ (unusual/malformed docx)Placeholders are left unreplaced rather than risk wrong images; inspect the source file's media manually if images are needed

Alternatives

Compare before choosing

Computed 10042,015

coreyhaines31/marketingskills

ab-testing

When the user wants to plan, design, or implement an A/B test or experiment, or build a growth experimentation program. Also use when the user mentions "A/B test," "split test," "experiment," "test this change," "variant copy," "multivariate test," "hypothesis," "should I test this," "which version is better," "test two versions," "statistical significance," "how long should I run this test," "growth experiments," "experiment velocity," "experiment backlog," "ICE score," "experimentation program

Computed 997

event4u-app/agent-config

design-review

Use when the user says "review the design", "check the UI", or wants a comprehensive UI/UX review. Uses a 7-phase methodology covering interaction, responsiveness, accessibility, and more.

Computed 9831,966

K-Dense-AI/scientific-agent-skills

dask

Distributed computing for larger-than-RAM pandas/NumPy workflows. Use when you need to scale existing pandas/NumPy code beyond memory or across clusters. Best for parallel file processing, distributed ML, integration with existing pandas code. For out-of-core analytics on single machine use vaex; for in-memory speed use polars.

Computed 9831,966

K-Dense-AI/scientific-agent-skills

neurokit2

Use NeuroKit2 to build or audit reproducible research workflows for physiological time-series preprocessing, event/interval analysis, multimodal alignment, variability, and complexity. Trigger when code imports neurokit2 or needs its current APIs, schemas, and method-aware validation—not for diagnosis or device validation.