Last updated: 16 September 2026
Fabien Raquidel
Freelance SEO consultant
Paris, France
Email: contact@referenceur-web.pro
Publication director: Fabien Raquidel.
o2switch
SAS with share capital of €100,000
Chemin des Pardiaux, 63000 Clermont-Ferrand, France
SIRET 510 909 807 00032, Clermont-Ferrand trade register
www.o2switch.fr
All elements of this site, including text, layout, code, logo and illustrations, are the exclusive property of Fabien Raquidel unless stated otherwise. Any reproduction or representation, in whole or in part, without prior written permission is prohibited.
The data displayed by the tool comes from the public index of Common Crawl, a non-profit organisation, and from the servers of the sites analysed. It remains the property of its respective owners. The trade marks and crawler names mentioned, CCBot, GPTBot, ClaudeBot, PerplexityBot and Google-Extended, belong to their respective vendors and are referred to descriptively only.
This tool is provided free of charge, for information purposes. It performs three measurements at the moment you request them: reading the robots.txt file of the domain analysed, a request to that domain, and a query to the public Common Crawl index.
Results reflect the state of those sources at the time of measurement. They may vary from one run to the next, in particular because the Common Crawl index is at times momentarily incomplete, and because a server may answer differently depending on the moment. They are provided without warranty of accuracy or completeness, and constitute neither a contractual audit nor professional advice.
The presence of a site in Common Crawl indicates that it was crawled. It does not demonstrate that a language model was trained on its content, nor that it will be cited in an answer. The publisher cannot be held liable for decisions taken on the basis of the information displayed, nor for the consequences of any configuration change made as a result.
The tool is intended for people analysing a site they are responsible for. The requests it issues are limited in number and strictly identical to those of an ordinary indexing crawler.
The site requires no sign-up and no account. It collects no data about you, with one exception described below: the address you leave if you request a backlinks report.
Only two items are recorded on the server, with each analysis:
Your IP address is never stored. It is transformed by a one-way function with a secret salt, and only four characters of the result are kept. This makes it possible to tell two visitors apart within a single day, not to identify any one of them. The data controller is therefore unable to link these records to a person.
These items serve two purposes: measuring use of the service and protecting it against abuse, by limiting the number of analyses per visitor. The legal basis is the publisher's legitimate interest in ensuring the proper operation and security of the service.
Anti-abuse counters are erased after one hour. The analysis log is kept for 12 months, then deleted automatically. No data is passed on, sold or transferred to a third party, and no audience measurement service is installed.
Under the General Data Protection Regulation you have rights of access, rectification, erasure, restriction and objection. You may exercise them at contact@referenceur-web.pro. Given the above, the publisher is not in a position to identify you within the records kept, and may ask for further details or inform you that no data identifiable as yours exists. You may also lodge a complaint with the French data protection authority, the CNIL.
That report cannot be computed while you wait: it requires reading Common Crawl's entire link graph, close to ten gigabytes, which takes about twenty minutes. The computation is therefore batched and run once a night for every pending request. That is the only reason an address is asked for.
The address is used to notify you once, and nothing else. It is stored with the requested domain, used to send the message announcing the report is ready, then deleted straight away. It is neither passed on, nor sold, nor used to write to you again. There is no newsletter and no further mailing.
The report itself stays available for thirty days through a link containing a random identifier, then it is deleted automatically along with the request that produced it. The legal basis is the performance of the request you make.
The host keeps technical connection logs of its own, under the conditions set by the applicable regulations and by its own terms.
This site uses no cookies. It stores nothing in your browser, uses neither local storage, nor tracking pixels, nor advertising identifiers, and installs no audience measurement tool.
No resource is loaded from a third-party service: no remote font, no external library, no image hosted elsewhere. Your browser therefore communicates with this domain alone. That is also why no consent banner is needed.
The site links to external resources, in particular the documentation of Common Crawl, Cloudflare and OpenAI, and the datasets cited. The publisher exercises no control over their content and accepts no responsibility for it.
This notice is governed by French law. In the event of a dispute, and failing an amicable settlement, the French courts have jurisdiction.