PageSourceSearch

Site profile

ad-datasets.com

Built with @findify/bundle 7.1.108, google-analytics, bootstrap-vue 2.23.1, @tolgee/web 7.2.1, @pdftron/pdfjs-express-viewer 8.7.5, @memori.ai/memori-webcomponent 8.3.3, @accusoft/pdf-viewer 3.11.158639, @okta/okta-signin-widget 7.49.1, @arco-design/web-react 2.66.16. 1 page and 8 stored files, last crawled 2026-10-07.

rank 12,868,785 9 libraries 0 identifiers 104 third-party hosts 1 crawl open the site

Technologies and libraries

Recognised by fingerprinting the stored files and bundles against known releases; a version is the release the bytes match.

LibraryVersionSeen
@findify/bundle 7.1.108 2026-10-07
google-analytics 2026-10-07
bootstrap-vue 2.23.1 2026-10-07
@tolgee/web 7.2.1 2026-10-07
@pdftron/pdfjs-express-viewer 8.7.5 2026-10-07
@memori.ai/memori-webcomponent 8.3.3 2026-10-07
@accusoft/pdf-viewer 3.11.158639 2026-10-07
@okta/okta-signin-widget 7.49.1 2026-10-07
@arco-design/web-react 2.66.16 2026-10-07

Tracking, tag and verification IDs

No analytics property, tag manager container, pixel or site-verification token was found in the stored source.

Third-party hosts

Hosts outside ad-datasets.com that its pages load scripts, frames or stylesheets from, and that its own scripts name in absolute URLs (connect). Each links to every site loading from the same host.

HostLoaded asFound in
7dlabs.com connect view source
ai.stanford.edu connect view source
aimagelab.ing.unimore.it connect view source
ankitshah009.github.io connect view source
apolloscape.auto connect view source
arxiv.org connect view source
avdata.ford.com connect view source
benchmark.ini.rub.de connect view source
cadcd.uwaterloo.ca connect view source
cg.cs.tsinghua.edu.cn connect view source
cmp.felk.cvut.cz connect view source
coda-dataset.github.io connect view source
crashd-cars.github.io connect view source
cs.stanford.edu connect view source
cvgl.stanford.edu connect view source
cvrr.ucsd.edu connect view source
cvssp.org connect view source
daniilidis-group.github.io connect view source
data.nvision2.eecs.yorku.ca connect view source
data.vision.ee.ethz.ch connect view source
docs.google.com connect view source
download.visinf.tu-darmstadt.de connect view source
epan-utbm.github.io connect view source
fishyscapes.com connect view source
github.com connect view source
hci-benchmark.iwr.uni-heidelberg.de connect view source
hci.iwr.uni-heidelberg.de connect view source
idd.insaan.iiit.ac.in connect view source
idda-dataset.github.io connect view source
ieeexplore.ieee.org connect view source
interaction-dataset.com connect view source
its.acfr.usyd.edu.au connect view source
joonyoung-cv.github.io connect view source
journals.sagepub.com connect view source
level-5.global connect view source
link.springer.com connect view source
mui.com connect view source
multimodal-distill.cs.uni-freiburg.de connect view source
multispectral.kaist.ac.kr connect view source
npm3d.fr connect view source
once-for-auto-driving.github.io connect view source
openaccess.thecvf.com connect view source
outreach.didichuxing.com connect view source
oxford-robotics-institute.github.io connect view source
pandaset.org connect view source
panoptic-bev.cs.uni-freiburg.de connect view source
paperswithcode.com connect view source
pdf.sciencedirectassets.com connect view source
people.ee.ethz.ch connect view source
prevention-dataset.uah.es connect view source
pro.hw.ac.uk connect view source
radar-scenes.com connect view source
reactjs.org connect view source
research.comma.ai connect view source
robotcar-dataset.robots.ox.ac.uk connect view source
robots.engin.umich.edu connect view source
rugd.vision connect view source
segmentmeifyoucan.com connect view source
sites.google.com connect view source
soda-2d.github.io connect view source
synthia-dataset.net connect view source
talk2car.github.io connect view source
thudair.baai.ac.cn connect view source
unmannedlab.github.io connect view source
unsupervised-llamas.com connect view source
usa.honda-ri.com connect view source
waymo.com connect view source
wilddash.cc connect view source
www.4seasons-dataset.com connect view source
www.6d-vision.com connect view source
www.a2d2.audi connect view source
www.argoverse.org connect view source
www.bdd100k.com connect view source
www.boreas.utias.utoronto.ca connect view source
www.cityscapes-dataset.com connect view source
www.cs.toronto.edu connect view source
www.cv-foundation.org connect view source
www.cvlibs.net connect view source
www.dbehavior.net connect view source
www.gavrila.net connect view source
www.google-analytics.com connect view source
www.highd-dataset.com connect view source
www.ind-dataset.com connect view source
www.ini.rub.de connect view source
www.kaggle.com connect view source
www.mapillary.com connect view source
www.mohamedaly.info connect view source
www.mpi-inf.mpg.de connect view source
www.mrpt.org connect view source
www.nec-labs.com connect view source
www.nightowls-dataset.org connect view source
www.nuscenes.org connect view source
www.poss.pku.edu.cn connect view source
www.robesafe.uah.es connect view source
www.robots.ox.ac.uk connect view source
www.round-dataset.com connect view source
www.spiedigitallibrary.org connect view source
www.uni-ulm.de connect view source
www.vis.xyz connect view source
www.vision.caltech.edu connect view source
www.w3.org connect view source
wwwlehre.dhbw-stuttgart.de connect view source
xingangpan.github.io connect view source
yuxng.github.io connect view source

Stored pages and scripts

The newest stored version of each file, newest first@if (p.FilesTruncated) { (the 500 newest of 8) }. Versions are the crawls at which the file was new or its content changed; each opens the file as it was then, rebuilt from the same stored chunks.

FileKindSizeCollectedVersions
https://www.google-analytics.com/analytics.js js library not fetched 2026-10-07 2026-10-07
https://www.google-analytics.com/analytics_debug.js js library not fetched 2026-10-07 2026-10-07
https://ad-datasets.com/static/js/3.3124b288.chunk.js js 4.5 KB 2026-10-07 2026-10-07
https://ad-datasets.com/static/js/runtime-main.12be113c.js js 2.3 KB 2026-10-07 2026-10-07
https://ad-datasets.com/asset-manifest.json js 1019 B 2026-10-07 2026-10-07
https://ad-datasets.com/static/js/2.4f652300.chunk.js js 902.6 KB 2026-10-07 2026-10-07
https://ad-datasets.com/static/js/main.e58e24fd.chunk.js js 219.5 KB 2026-10-07 2026-10-07
https://ad-datasets.com/ html 3.2 KB 2026-10-07 2026-10-07

Timeline

One entry per crawl, newest first, with what changed since the crawl before: libraries, identifiers, third-party hosts and files. Historical versions stay viewable because the stored chunks are shared between versions, never copied.

  1. first crawl

    1 page, 8 files, 1.1 MB; 9 libraries, 0 identifiers, 104 third-party hosts

    9 libraries added
    • @accusoft/pdf-viewer 3.11.158639
    • bootstrap-vue 2.23.1
    • @pdftron/pdfjs-express-viewer 8.7.5
    • @findify/bundle 7.1.108
    • @memori.ai/memori-webcomponent 8.3.3
    • @okta/okta-signin-widget 7.49.1
    • @tolgee/web 7.2.1
    • @arco-design/web-react 2.66.16
    • google-analytics
    104 third-party hosts added
    8 files added

Questions about ad-datasets.com

How does PageSourceSearch know what ad-datasets.com is built with?
From the site's own source code. The crawler stores the exact bytes of its pages and first-party scripts; the libraries are recognised by fingerprinting that code against known releases, the identifiers are read out of the tag snippets, and the third-party hosts are the script, iframe and stylesheet sources in the HTML and the absolute URLs inside the scripts. Nothing is inferred from headers or guessed.
How far back does the history of ad-datasets.com go?
To the first crawl the timeline lists. Every crawl records what the site looked like; when a file's content changes, its earlier version stays viewable because the stored chunks are never deleted, only mapped. A crawl that finds a file unchanged adds no copy.
Can I see an older version of a script from ad-datasets.com?
Yes. In the stored files list, every date under a file opens the version that was live at that crawl, with the same viewer as the current one. The raw bytes can be downloaded from there.
Which other websites use the same tracking IDs as ad-datasets.com?
Each identifier links to its reverse lookup, the list of every site in the index whose source carries the same id: the sites one analytics property or tag manager container is shared across.

Look up another site, browse sites by tracking ID, third-party host or technology. Site owners: see the crawler page for how stored pages are removed.