Skip to content

Latest commit

 

History

1 Commit

Folders and files

NameName
Last commit message
Last commit date
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 

Repository files navigation

 __  ___            _ ____     
/  |/  /___ _____  (_) __/_  __
/ /|_/ / __ `/ __ \/ / /_/ / / /
/ /  / / /_/ / /_/ / __/ /_/ / 
/_/  /_/\__,_/ .___/_/_/  \__, / 
            /_/          /____/   

SEO Sitemap Generator with HTTP/3 Stealth & TLS Impersonation

v2.0 | Developed By Hayder | Property of Qamrix Tech

Python 3.10+ License: MIT

Building intelligent solutions that drive progress and shape the future.


What is Mapify?

Mapify is a website crawler that generates SEO sitemaps while respecting robots.txt. It browses like a real visitor -- using real browser fingerprints and HTTP/3 -- so websites see it as normal traffic.

Feed it a URL, and it gives you a complete XML sitemap with SEO data: titles, descriptions, headers, images, and link structure.

Contact: qamrixtech@gmail.com | GitHub


Screenshots

Launcher
Auto-detects Python, installs deps
Main Menu
Navigate with arrow keys
Crawl Depth
How deep to follow links
Request Delay
Time between requests
Max Pages
Limit total pages crawled
Feature Toggle
Stealth, images, XenForo filter
Proxy Config
SOCKS5/HTTP proxy support
Config Summary
Review before starting
Ready to Crawl
Confirm and start
Crawl in Progress
Live progress tracking
SEO Analysis
Title, meta, headers, links
Crawl Complete
Summary with stats

Quick Start (Easiest Way)

1. Clone the repo:

git clone https://github.com/qamrixtech/mapify.git
cd mapify

2. Run the launcher -- it handles everything:

Platform Command
Linux ./run.sh
macOS ./run.sh
Windows python run.py
WSL ./run.sh

The launcher will:

  • Find Python 3.10+ on your system
  • Create a virtual environment automatically
  • Install all required packages (curl_cffi, PySocks, rich, etc.)
  • Launch Mapify

You don't need to install anything manually.


Features

Stay Hidden Crawl Smart
  • Looks like a real browser to websites
  • Uses HTTP/3 (the latest protocol)
  • Mimics Chrome, Firefox, Safari, Edge TLS fingerprints
  • 24 browser identities to rotate through
  • Real browser headers on every request
  • Respects robots.txt rules
  • Skips junk URLs (login pages, search, pagination)
  • Deduplicates pages automatically
  • Filters XenForo/vBulletin forum junk
  • Follows canonical tags to avoid repeats
SEO Data Flexible Output
  • Title length analysis
  • Meta description coverage
  • H1/H2 header structure
  • Image ALT text check
  • Internal/external link count
  • Content size metrics
  • XML sitemap for search engines
  • Image sitemap with captions
  • JSON file with full crawl data
  • Choose output directory
  • Proxy rotation support

How to Use

Interactive Mode (Recommended)

./run.sh

The launcher opens a menu where you use arrow keys to select:

  1. Enter target URL -- the website you want to crawl
  2. Set crawl depth -- how many links deep to follow (default: 3)
  3. Set request delay -- seconds between requests (default: 1.0)
  4. Set max pages -- limit total pages (default: 1000)
  5. Toggle features -- stealth mode, images, XenForo filter
  6. Proxy config -- optional HTTP/SOCKS5 proxy
  7. Confirm and start

Command Line

# Basic crawl
python mapify.py https://example.com

# Full options
python mapify.py https://example.com \
  --depth 5 \
  --delay 0.5 \
  --max-pages 5000 \
  --images \
  --proxy socks5://127.0.0.1:1080 \
  --verbose

# With proxy list
python mapify.py https://example.com --proxy-list proxy.txt

# Quick defaults (depth 3, 1000 pages, stealth on)
python mapify.py https://example.com --stealth

CLI Options

Option Default What it does
--depth 3 How deep to follow links (1-10)
--delay 1.0 Seconds to wait between requests
--max-pages 1000 Maximum pages to crawl
--images off Include images in the sitemap
--proxy none Use a single proxy (HTTP/HTTPS/SOCKS5)
--proxy-list none File with proxy list (rotates per request)
--stealth on Use HTTP/3 + browser TLS impersonation
--strip-queries on Remove tracking params from URLs
--skip-xenforo on Filter out forum junk URLs
--verbose off Show each page as it's crawled
--output output Where to save results
--config none Load settings from YAML file
-i off Interactive TUI mode

Proxy Support

Create a proxy.txt file with one proxy per line:

http://127.0.0.1:8080
https://proxy.example.com:8443
socks5://127.0.0.1:1080
http://user:pass@proxy.example.com:3128

Mapify rotates through proxies automatically, using a different proxy for each request.


Configuration File

Create a config.yaml:

crawler:
  max_depth: 5
  delay: 0.5
  max_pages: 5000
  max_workers: 4
  timeout: 15
  strip_queries: true
  skip_xenforo_junk: true
  stealth_mode: true
  proxy: "socks5://127.0.0.1:1080"
python mapify.py https://example.com --config config.yaml

Output

After crawling, Mapify saves to the output/ directory:

output/
  sitemap.xml          # XML sitemap for search engines
  sitemap-images.xml   # Image sitemap (if --images)
  sitemap.json         # Full crawl data + SEO metadata

File Structure

mapify/
  mapify.py            # Main tool (entry point)
  run.py               # Launcher (Windows / cross-platform)
  run.sh               # Launcher (Linux / macOS / WSL)
  requirements.txt     # Python dependencies
  setup.py             # Package setup
  config.yaml.example  # Example config
  LICENSE              # MIT License
  README.md            # Documentation
  images/              # Screenshots
  src/
    cli.py             # CLI and interactive menu
    crawler.py         # Core crawling engine
    ui.py              # TUI rendering

Platform Support

Platform Status How to launch
Linux Supported ./run.sh
macOS Supported ./run.sh
Windows Supported python run.py
WSL Supported ./run.sh

Requirements

  • Python 3.10 or higher
  • That's it -- the launcher installs everything else automatically

License

Copyright (c) 2026 Qamrix Tech. All Rights Reserved.
Developed By Hayder

See LICENSE for details.


Qamrix Tech

Made with <3 by Hayder -- Qamrix Tech

About

SEO Sitemap Generator with HTTP/3 Stealth & TLS Impersonation. Crawls websites like a real browser, generates XML sitemaps with SEO analysis. Built with Python, curl_cffi, Rich TUI.

Topics

Resources

Stars

0 stars

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages