Category 07 · 14 resources

Web Search / Archives / Historical Research

General web search engines and historical archives. Mastering operators and the Wayback Machine recovers content that was deleted or never indexed by Google.

google Web

https://www.google.com/

Google search

How to use it

Master operators: site:, filetype:, intitle:, inurl:, quotes, minus, and date ranges via Tools. Combine operators to find documents others miss.

Open resource →

bing Web

https://www.bing.com/

Bing search

How to use it

Often indexes content Google misses; its visual search and maps are useful OSINT pivots.

Open resource →

brave Web

https://search.brave.com/

Brave independent search index

How to use it

An independent index (not Google-based) — use it to cross-check results and escape filter bubbles.

Open resource →

duckduckgo Web

https://duckduckgo.com/

Privacy search engine

How to use it

Anonymous searching with !bang shortcuts (e.g., !gh, !so) that jump directly into thousands of site searches.

Open resource →

startpage Web

https://www.startpage.com/

Anonymous Google proxy

How to use it

Google results without tracking; its Anonymous View lets you open pages through a proxy.

Open resource →

mojeek Web

https://www.mojeek.com/

Independent crawler search engine

How to use it

Fully independent index — finds pages absent from Google/Bing.

Open resource →

qwant Web

https://www.qwant.com/

European privacy search

How to use it

EU-based search with its own index, strong for French/European content.

Open resource →

baidu Web

https://www.baidu.com/

China's dominant search engine

How to use it

Essential for Chinese-language content; indexes WeChat ecosystem and Chinese sites Google can't reach.

Open resource →

sogou Web

https://www.sogou.com/

Sogou search

How to use it

Tencent-backed engine; its WeChat search (weixin.sogou.com) searches public WeChat articles.

Open resource →

so Web

https://www.so.com/

360 search

How to use it

Chinese search engine useful as a second source for Chinese content.

Open resource →

archive Web

https://archive.org/web/

Internet Archive

How to use it

Search archived web pages, books, video and software. The /web path is the Wayback Machine.

Open resource →

archive Web

https://web.archive.org/

Internet Archive

How to use it

Search archived web pages, books, video and software. The /web path is the Wayback Machine.

Open resource →

commoncrawl Web

https://commoncrawl.org/

Massive web-crawl archive

How to use it

Free petabyte-scale crawl data; query the index API to find historical captures of URLs programmatically.

Open resource →

webcheck Web

https://webcheck.xyz/

Website analyzer

How to use it

Enter a URL to get headers, DNS, SSL, redirects and tech-stack info in one report.

Open resource →

Frequently asked questions

Do I need to install anything for these 07 tools?

No — these are web-based. Open the link in your browser; some platforms require a free account or API key.

Is it legal to use these resources?

Public-record lookups and defensive/research use are generally lawful, but rules vary by country and tool. Only test systems you own or have written authorization to test, respect each site's terms of service, and never use personal data unlawfully.

Where should a beginner start in this category?

Start with the web-based tools at the top of the list — they need no setup. Read each card's “How to use it” panel, run one real query, and only then move to installable tools.