Getting Started
What is IDM Crawler?
A specialized auditing engine designed to detect unauthorized marketing tags and validate data layer integrity across enterprise websites.
Governance Focus
Built for technical auditors to identify "Shadow IT" and rogue pixels operating outside your tag management system.
Pillar Scoring
Automatic 0-100 grading across five technical governance domains with actionable recommendations.
Fast Diagnostics
Complete single-page audits in under 30 seconds with full network waterfall maps and vendor detection.
Quick Start Guide
Download & Install
Download the Windows executable from our website. Run the installer and follow the setup wizard. Launch IDM Crawler from your Start Menu or Desktop shortcut.
Launch & Start Crawling
No browser installation needed! The Chromium engine is already bundled with the application. Simply launch IDM Crawler, enter a URL, and click START. The browser is ready to use immediately.
Run Your First Audit
Enter a URL (e.g., https://example.com), select Single Page Audit, and click START. Review results across the 5 tabs to understand your tag governance health.
System Requirements
| Resource | Minimum | Recommended |
|---|---|---|
| Operating System | Windows 10 (64-bit) | Windows 11 |
| RAM | 4 GB | 8 GB |
| Disk Space | 500 MB | 1 GB (for logs and exports) |
| Internet | Broadband connection | Broadband connection |
Installation
Windows Installer (Recommended)
-
Download
IDMCrawler-Setup.exefrom our website - Run the installer and follow the prompts
- Launch IDM Crawler from your Start Menu or Desktop shortcut
- Start crawling immediately - browser is bundled with the application
Portable Version
-
Download
IDM_Crawler_Portable.zip -
Extract to any folder (e.g.,
C:\IDMCrawler) -
Run
IDM_Crawler_Portable.exe -
The
browsersfolder must remain in the same directory as the executable - Start crawling immediately - no additional setup required
Core Capabilities
5-Pillar Governance Framework
The engine evaluates every page against these key technical standards:
Performance Scoring Algorithm
The performance score (0-100) is calculated using weighted metrics based on industry benchmarks:
Data Layer Validation
IDM Crawler validates data layer implementations across major platforms:
GTM dataLayer
- Checks for array format and proper structure
- Validates page view events (gtm.js, gtm.dom, gtm.load)
- Detects ecommerce objects and transaction data
- Identifies event duplication and race conditions
Tealium utag.data
- Validates tealium_event presence (view, link, etc.)
- Checks for page_name, dom.title, and site section data
- Detects ecommerce variables (_corder, _ctotal, _cprod)
Adobe digitalData
- Basic structure validation
- Custom validation rules available per implementation
Consent Detection
- CMP Presence - Identifies OneTrust, Cookiebot, Tealium Consent Manager, and other CMPs
- Pre-consent Firings - Flags tags that fire before user consent is given
- Consent Category Validation - Coming in v0.10+
Supported Vendors & Detection
Tag Manager Detection
IDM Crawler can detect and validate the following tag management systems:
Analytics & Tracking Vendors
The engine identifies these analytics platforms and their implementation IDs:
Advertising & Marketing Vendors
IDM Crawler detects these advertising and marketing pixels:
Media & Video Platforms
IDM Crawler detects embedded media players and video analytics:
Detection Methodology
IDM Crawler uses multiple techniques to identify vendors with high accuracy:
Unknown Vendor Handling
When IDM Crawler cannot identify a vendor, it still provides valuable information:
- Domain tracking - The requesting domain is always captured
- Loader identification - Whether the script came from a tag manager or was hardcoded
- Blocking status - Whether the script blocks page rendering
- Execution order - When the script loaded relative to other resources
UI Guide
Overview Tab
The dashboard showing your tag governance health at a glance:
- 5 Pillar Scores - Visual representation with letter grades (A-F)
- Active Stack - All detected vendors organized by category (Tag Managers, Analytics, Marketing, etc.)
- Key Metrics - Pages crawled, total scripts found, unmanaged tags count
- Recent Issues - Top priority findings requiring attention
Use this tab to get a quick health check of your tag implementation.
Network Tab
Complete network request waterfall with vendor detection:
- URL - The full request URL
- Domain - Third-party domain (helps identify external vendors)
- Vendor - Detected vendor name (Google Analytics, Meta Pixel, etc.)
- Loader - How the script was loaded (Tag Manager, Hardcoded, etc.)
- Managed - Whether it's governed by an authorized tag manager
- Blocking - Whether the request blocks page rendering
- Status - HTTP status code (200, 204, 404, etc.)
Scripts Tab
All JavaScript resources with execution order analysis:
- Execution Order - Sequential order in which scripts loaded
- Name/Vendor - Identified vendor or "Unknown"
- Loader Type - How the script was initiated (HTML/Parser, Tag Manager, Hardcoded)
- Issue - Any governance issues detected (Hardcoded, Pre-consent, etc.)
- Severity - Issue severity (0-10) with color coding
- URL - Full script URL for debugging
DataLayer Tab
All data layer events captured during the crawl:
- Event Name - The event type (gtm.js, gtm.dom, page_view, etc.)
- Vendor - Which data layer (GTM dataLayer, Tealium utag.data, etc.)
- Data - The complete event payload with all variables
- Page URL - Which page the event fired on (for multi-page crawls)
- Timestamp - When the event occurred
Issues Tab
Aggregated findings organized by pillar for easy prioritization:
- Severity - Color-coded (Red=Critical, Orange=High, Yellow=Medium, Blue=Low)
- Description - Clear explanation of the issue
- Details - Specific technical details and affected resources
- Pages Count - How many pages are affected by this issue
- Example Pages - Sample URLs where the issue was found
Use this tab to prioritize your remediation efforts based on severity and impact.
Permissions & Roles
Role-Based Permission Matrix
IDM Crawler respects your organization's role-based access control. The following permissions apply to all crawler operations:
| Permission / Action | Owner | Admin | Editor | Viewer | Client |
|---|---|---|---|---|---|
| View Organizations & Properties | |||||
| Run Crawl Audits | |||||
| Export Reports (PDF/CSV/JSON) | * | * | |||
| Sync to Cloud Workspace | |||||
| Add/Edit Properties | |||||
| Manage Organization Settings | |||||
| Invite/Remove Members |
Role Descriptions
Full control over the organization. Can manage all settings, members, properties, and audit data.
- Full read/write access to all crawler features
- Can sync audits to cloud workspace
- Can add/edit/delete properties
- Can manage organization members and roles
Full operational control. Can manage properties, audits, and members, but cannot delete the organization.
- Full read/write access to all crawler features
- Can sync audits to cloud workspace
- Can add/edit/delete properties
- Can manage organization members
Can create and modify content. Ideal for technical auditors who need to run crawls and save results.
- Can run crawls and sync audits to cloud
- Can add/edit properties
- Cannot manage organization members or settings
Read-only access. Can view audit results and export locally, but cannot sync to cloud.
- Can run crawls and view results
- Can export reports locally (PDF/CSV/JSON)
- Cannot sync to cloud workspace
- Cannot add or modify properties
Limited access for external stakeholders. Can view audit results for specific assigned properties.
- Can run crawls on assigned properties
- Can export reports locally
- Cannot sync to cloud workspace
- Cannot modify properties or settings
Sync to Cloud Limitations
The "Sync to Cloud" feature is only available to users with Owner, Admin, or Editor roles. This ensures that audit data is only written to your workspace by authorized personnel.
- Overwriting existing audit records
- Creating misleading or incorrect audit entries
- Accessing sensitive audit data outside their permission scope
The "Sync to Cloud" button will be hidden from your toolbar. You can still run crawls and export results locally, but cloud synchronization is disabled.
If you need cloud sync capabilities, contact your organization administrator to upgrade your role to Editor, Admin, or Owner.
How to Check or Upgrade Your Role
Check Your Current Role
- Log in to IDM Crawler
- Select your organization from the dropdown in the left sidebar
- Your role is indicated next to the organization name (e.g., "Acme Corp (Professional) - Owner")
-
You can also check by looking for the "Sync to
Cloud" button after a crawl:
- If visible → You have write permissions (Owner/Admin/Editor)
- If hidden → You have read-only permissions (Viewer/Client)
Request a Role Upgrade
- Contact your organization's Owner or Admin
- They can change your role in the IDM Web App under Organization Settings → Members
- After the role is updated, restart IDM Crawler to see the changes
Need Help?
If you're unsure about your permissions or need assistance, contact support@idmartech.com
Privacy & Data Collection
What Data We Collect
IDM Crawler is designed with privacy-first principles. We collect only what's necessary for functionality and troubleshooting.
| Data Type | Collected By Default | Purpose | Stored Where |
|---|---|---|---|
| Error Signatures | ✓ Yes | Crash reporting and bug fixes | Local (your computer) + optional anonymous upload |
| Domain Names | ✓ Yes | Troubleshooting crawl issues | Local only (logs) |
| Full URLs | ✗ No (opt-in) | Deep troubleshooting | Local only (debug logs, requires consent) |
| Page Content / HTML | ✗ Never | N/A | Never collected |
| Passwords / Tokens | ✗ Never | N/A | Never collected |
| Personal Information | ✗ Never | N/A | Never collected |
| Anonymous Usage Stats | ✗ No (opt-in) | Product improvement | Our analytics (anonymous, aggregated) |
- Enable telemetry in privacy settings
- Export a support bundle and share it with us
- Log in and save crawl results to your workspace
Privacy Modes
IDM Crawler offers three privacy modes that control what information is written to logs. You can change this anytime via Help → Privacy Settings.
What's logged:
- Error types only
- No URLs, domains, or personal data
Best for: Maximum privacy, but limits troubleshooting capability.
What's logged:
- Domain names only (e.g., "example.com")
- No full URLs or query parameters
- No page content or personal data
Best for: Balancing privacy with effective troubleshooting.
What's logged:
- Full URLs (paths only, query params stripped)
- Detailed operation logs
- Data stays on YOUR computer
Best for: Deep troubleshooting with explicit consent required.
Telemetry & Crash Reports
Anonymous Usage Statistics (Opt-In)
When enabled, we collect completely anonymous usage data to help improve IDM Crawler:
- Feature usage counts (e.g., "Single Page Audit used 10 times")
- Error frequency and types
- Performance metrics (crawl duration, pages per crawl)
- App version and operating system
- NEVER URLs or domain names
- NEVER page content or personal data
- NEVER browsing history outside IDM Crawler
Crash Reports (Opt-Out by Default, Recommended)
When the application crashes, we may receive:
- Error type and stack trace (no personal data)
- App version and operating system
- Privacy mode setting
- NEVER URLs or crawl data
Crash reports help us identify and fix bugs faster. You can disable this in privacy settings.
Log Files (Stored Locally)
IDM Crawler maintains several log files on your
computer for troubleshooting. These are stored in %APPDATA%\IDMCrawler\logs\.
| File | Content | Privacy Mode Impact | Retention |
|---|---|---|---|
| idm_crawler.log | Main application log | Respects privacy mode | Rotated at 10MB (keeps 5 backups) |
| errors.log | Error-only log | Always safe (no PII) | Rotated at 5MB (keeps 3 backups) |
| debug.log | Detailed debug log | Only created in Debug Mode | Rotated at 20MB (keeps 2 backups) |
| audit.json | User action audit trail | No PII, action names only | Append-only |
| crashes.log | Crash reports | No PII | Append-only |
Your Privacy Controls
You have complete control over your data at all times:
Privacy Settings
Change privacy mode, telemetry, and crash reporting anytime via Help → Privacy Settings.
View Logs
Access all log files directly via Help → Open Log Folder.
Delete Logs
Manually delete log files from the logs folder. Old logs are automatically rotated.
Support Bundle
Generate a support bundle via Help → Export Support Bundle. You control what's included.
Cloud Sync
Save crawls to your workspace only when you log in and explicitly save.
Opt-Out
Disable telemetry and crash reporting at any time without affecting functionality.
Legal Compliance
GDPR Compliant (EU)
- No personal data processed by default
- Explicit consent required for telemetry
- Right to access, delete, and export your data
- Data stays on your computer unless you share
CCPA Compliant (California)
- No "sale" of personal information
- Opt-out available for telemetry
- Clear disclosure of data collection
PIPEDA Compliant (Canada)
- Meaningful consent for data collection
- Data limited to what's necessary
- Transparent privacy practices
Questions About Privacy?
Contact our Data Protection Officer at privacy@idmartech.com
How-To Guides
Single Page Audit
- Launch IDM Crawler
- Enter the URL you want to audit (e.g., https://example.com)
- Select "Single Page Audit" from the dropdown
- Click START
- Wait for the audit to complete (typically 10-30 seconds)
- Review results across the 5 tabs
Full Site Crawl
- Launch IDM Crawler
- Enter your website's starting URL (usually the homepage)
- Select "Full Site Crawl" from the dropdown
- Click START
- The crawler will discover and analyze pages based on your plan limit (Starter: 50, Professional: 500, Agency: Unlimited)
- Monitor progress in the progress bar and console output
- Review aggregated results across all pages in each tab
- Starter (Free): Up to 50 pages per crawl
- Professional ($59/mo): Up to 500 pages per crawl
- Agency ($149/mo): Unlimited pages per crawl
Plan Limits & Upgrading
The IDM Crawler is completely free for local audits. However, certain features have limits based on your workspace plan:
| Feature | Starter (Free) | Professional ($59/mo) | Agency ($149/mo) |
|---|---|---|---|
| Pages per Crawl | 50 | 500 | Unlimited |
| Local Audits | ✓ Unlimited on all plans | ||
| Export Reports | ✓ CSV, PDF (login required) | ||
| Cloud Sync | 1 Workspaces | Up to 10 Workspaces | Unlimited |
How to Check Your Current Plan
- Log in to IDM Crawler
- Your plan is displayed next to your organization name (e.g., "Acme Corp - Professional")
- You can also check in the WebApp under Workspace Settings → Billing
How to Upgrade
- Go to IDM WebApp
- Navigate to Workspace Settings → Billing
- Click "Upgrade to Professional" or contact sales for Agency
- Choose monthly or annual billing (annual saves 17%)
- Enter payment details (processed securely by LemonSqueezy)
Export Results
- Complete a crawl (single page or full site)
- Log in to your IDM Crawler account (free registration required for export). Export is available on all plans including Starter.
- Navigate to any tab (Overview, Network, Scripts, etc.)
- Click the Export button in the tab toolbar
-
Choose your preferred format:
- CSV - For spreadsheet analysis
- PDF - For sharing with stakeholders and reports
- Save the file to your computer
Browser Configuration
IDM Crawler includes a complete Chromium browser engine. No additional setup is required!
Browser Management
- Location:
%APPDATA%\IDMCrawler\browsers\ - Size: ~300MB
- First Run: The browser is automatically extracted from the application bundle
Clearing Browser Cache
- Go to File → Browser → Clear Browser Cache
- Confirm the action
- The browser will be re-extracted on your next crawl
Interpret Results
Here's how to read and act on your audit results:
Severity Levels
Fix immediately - blocks compliance or severely impacts performance
Fix soon - significant governance risk
Plan to fix - moderate impact
Monitor - low priority
Informational only - no action needed
Pillar Score Interpretation
Common Actions by Issue Type
- Hardcoded Tags → Migrate to tag management system
- Pre-consent Firings → Review consent integration and tag firing rules
- Missing Data Layer → Implement GTM, Tealium, or Adobe data layer
- High Blocking Scripts → Add async/defer attributes or move to footer
- Multiple Analytics → Consolidate to single platform
Troubleshooting
Browser & Firewall
Windows Firewall Alert - "Windows Defender Firewall has blocked some features"
When you run your first crawl, Windows Defender Firewall may display a dialog asking if you want to allow IDM Crawler's browser to access the network.
How to resolve:
- When the firewall dialog appears, click "Allow access"
- You only need to do this once - Windows will remember your choice
"Browser not found" or "Browser failed to launch"
Solution: Try these steps in order:
- Clear browser cache via File → Browser → Clear Browser Cache
- Restart IDM Crawler
- If the issue persists, reinstall IDM Crawler
- Ensure you have enough disk space (~300MB free)
- Temporarily disable antivirus/firewall (then re-enable after testing)
"Browser already running" error
Solution: Close all IDM Crawler instances, wait 5 seconds, then relaunch. If persists, restart your computer.
Portable version: "browsers" folder not found
Solution: Ensure the browsers folder is in the same directory as the executable.
The portable version requires this folder to be present.
Need to crawl more than 50 pages?
Solution: The Starter plan has a 50-page limit per crawl. You can:
- Upgrade to Professional (500 pages) or Agency (unlimited) for higher limits
- Break your crawl into smaller segments by crawling specific sections of your site
- Use Single Page Audit mode for individual important pages
To upgrade, visit IDM WebApp → Workspace Settings → Billing.
Crawl Issues
No pages crawled (0 pages)
Possible causes and solutions:
- Invalid URL format → Ensure URL includes https:// or http://
- Site inaccessible → Check if the website is online
- Browser not installed → Install browser via File → Browser Setup
- Site blocking crawler → Some sites block automated access
- Authentication required → Login support coming in future version
Crawl stops before reaching expected pages
Solution: The Starter plan has a 50-page limit per crawl. Professional and Agency plans offer higher limits (500 and unlimited respectively). Check your workspace settings to confirm your current plan limits.
Some pages are not being discovered
Possible causes:
- Pages may require JavaScript rendering (our crawler supports this)
- Pages may be behind login walls
- Pages may have no internal links from discovered pages
- Site may use complex SPA routing
Solution: Try crawling specific URLs individually using Single Page Audit mode.
Vendor detection showing "Unknown" for known vendors
Solution: The vendor signature database is continuously updated. Report unknown vendors to support@idmartech.com for inclusion in future updates.
Export Issues
Export button is disabled/grayed out
Solution: You need to log in to your free IDM Crawler account to enable exports. Create an account at our website and log in from the application.
Export fails or produces empty file
Solution:
- Ensure you've completed at least one crawl
- Check that you have write permissions to the save location
- Try exporting to a different folder (e.g., Desktop or Documents)
- Restart the application and try again
PDF export looks different than expected
Solution: PDF formatting is optimized for printing and sharing. For exact data replication, use JSON or CSV export formats.
Release Notes
v0.9.0 (Beta)
May 20, 2026Current VersionWhat's New
- 5-Pillar Governance Framework - Comprehensive scoring across Tracking, Consent, Data Layer, Performance, and Governance
- Unmanaged Tag Detection - Identifies 50+ vendors including Google, Meta, Adobe, Tealium
- Bundled Browser - Chromium engine included - no separate installation required!
- Performance Impact Scoring - Weighted algorithm measuring third-party domains, tag bloat, blocking scripts, redundant pixels, and analytics overlap
- Data Layer Validation - Supports GTM dataLayer, Tealium utag.data, and Adobe digitalData
- Consent Detection - Identifies CMP presence and pre-consent tag firings
- Single Page & Full Site Crawl - Starter: 50 pages, Professional: 500 pages, Agency: Unlimited
- Export Capabilities - JSON, CSV, and PDF exports for logged-in users
- Modern UI - 5 tabs for detailed analysis (Overview, Network, Scripts, DataLayer, Issues)
- Guest Mode - Crawl unregistered sites with 50-page limit
Known Limitations (MVP)
- Starter plan limited to 50 pages per crawl (upgrade to Professional or Agency for higher limits)
- Team collaboration requires Professional or Agency plan (Starter: 1 user only)
- Exports require free account login
- Windows 64-bit only (Mac/Linux coming)
- Guest mode results cannot be saved to cloud
Coming in v0.10.0
- Custom vendor signature addition
- Detailed consent category validation
- Enhanced ecommerce funnel validation
- Improved PDF report formatting
Planned for Future Releases
- Enhanced consent management validation (Google Consent Mode, OneTrust, Cookiebot)
- Deeper ecommerce and data layer validation tools
- Custom validation rule builder for enterprise implementations
- White-label reporting for Agency plans
- Client portal access for Agency plan clients