Document Metadata Extractor
Upload any Office document (DOCX, XLSX, or PPTX) to extract all hidden metadata. The Document Metadata Extractor reads the internal ZIP structure and parses the XML metadata files (docProps/core.xml and docProps/app.xml) to display author information, creation dates, revision history, document statistics, company details, and custom properties. Fields are organized by namespace (File, Core, App, Custom) and categorized by type (Identity, Content, Dates, Statistics, Software). With tab-based filtering, export as text, and a clear field-by-field breakdown - all processing runs locally in your browser with zero uploads. Free, no signup required.
Upload any Office document (DOCX, XLSX, or PPTX) to extract all embedded metadata. The tool reads the internal ZIP structure and parses the XML metadata files to find author information, creation dates, revision history, document statistics, and more. All processing runs entirely in your browser - no uploads, no signup required.
Drop an Office document here or click to browse
Supports DOCX, XLSX, and PPTX files - no file upload
Core & App Metadata Extraction
Reads both docProps/core.xml (title, creator, dates, revision, keywords) and docProps/app.xml (application, word count, pages, slides, company) from any Office Open XML document. Also extracts custom properties when present.
Three Format Support
Supports DOCX (Word), XLSX (Excel), and PPTX (PowerPoint) files. Each format exposes different metadata fields - Word documents show page counts and word statistics, Excel shows sheets, and PowerPoint shows slides and presentation format.
Organized Field Display
Metadata fields are grouped by namespace (File, Core, App, Custom) and categorized by type (Identity, Content, Dates, Statistics, Software). Tab-based filtering lets you focus on specific namespaces. Export all metadata as plain text.
100% Browser-Based Processing
All processing happens entirely in your browser using JSZip for ZIP extraction and fast-xml-parser for XML parsing. Your documents never leave your device - no uploads, no servers, no third-party services.
Privacy & Data Leak Prevention
Before sharing documents externally, extract metadata to check for hidden author names, company information, revision history, and other personally identifiable information that could expose your organization.
Document Forensics & Attribution
Identify the author and creator of a document using embedded metadata. The creator, last modified by, and company fields can help determine the origin of an anonymous or suspicious document.
Intellectual Property Verification
Verify copyright ownership and creation dates of digital documents. The creation and modification timestamps provide a verifiable record of when a document was first created and how it has evolved.
Document Template Analysis
When working with document templates, extract metadata to identify the template source, company association, and application version. Useful for standardizing templates across an organization.
Collaboration History Review
Review the revision history and last modified by fields to understand who has worked on a document. The revision number and total edit time provide insight into the document's development lifecycle.
Application Compatibility Checking
Check which application and version created a document to ensure compatibility. Documents created in newer versions of Office may have features not available in older versions or alternative software.
What Is Document Metadata?
Document metadata is embedded information stored inside Office files that describes the document's properties - who created it, when it was created, what software was used, and statistics about its content. This metadata is stored in XML files within the Office Open XML (ZIP) package and is separate from the visible document content.
How Metadata Extraction Works
Office Open XML files (DOCX, XLSX, PPTX) are ZIP archives containing XML files that define the document structure. The Document Metadata Extractor opens these ZIP files using JSZip, reads the docProps/core.xml (Dublin Core metadata) and docProps/app.xml (application-specific properties) files, and parses them with fast-xml-parser to extract all metadata fields.
Types of Metadata Found
Core metadata includes: title, subject, creator, keywords, description, last modified by, revision number, category, content status, created/modified timestamps, and language. Application metadata includes: application name and version, total edit time, page/word/character counts (Word), slides/notes/hidden slides (PowerPoint), company, and manager.
Browser-Based & Private
All processing runs entirely in your browser - your documents are never uploaded to any server. The tool uses JSZip to read the ZIP structure and parses XML metadata files locally. This makes it suitable for analyzing confidential documents, legal files, and sensitive business data without any privacy concerns.
Related Tools
Steganography Detector
Upload an image to detect potential hidden data using LSB (Least Significant Bit) steganography analysis. Analyzes pixel-level modifications, shows a heatmap of altered pixels, detects statistical anomalies, and extracts hidden text messages if found. Compatible with standard LSB encoding and the STEG magic header format. All processing happens locally in your browser - free online steganography detector, no signup required.
Image File Size Analyzer
Upload any image to see a detailed binary-level breakdown of what contributes to its file size - pixel data, compression overhead, metadata (EXIF/IPTC/XMP), color profiles (ICC), and structural headers. Get prioritized optimization tips with estimated savings for web performance, mobile app optimization, and storage reduction. Supports JPEG, PNG, GIF, WebP, BMP, TIFF, AVIF, and HEIC - all processing runs locally in your browser, no server upload required. Free online image file size analyzer.
File Magic Byte Detector
Upload any file to instantly detect its true file type by reading the magic bytes (file signature / header). The tool bypasses incorrect file extensions and reveals the real format. Shows hex dump with matching signature bytes highlighted, ASCII interpretation, MIME type, file category, and extension match analysis. Supports over 120 file signatures across 16 categories: images, audio, video, documents, archives, executables, fonts, certificates, disk images, and more. All processing runs locally in your browser - free online File Magic Byte Detector, no signup required.
Image Clone Detector
Detect copy-move forgeries in images using pixel-block matching. Upload any image to find cloned/copied regions, view heatmap overlays showing the location and intensity of detected clones, and get confidence scores with region details. Three sensitivity presets for different detection needs. 100% private browser-based processing - free online Image Clone Detector.