What powers paperarchive.io?
Last seen Aug 2026 · 15 technologies across 11 categories
Analytics
CDN
Databases
Development
JavaScript frameworks
Miscellaneous
Performance
Programming languages
Security
UI frameworks
Webmail
AI Crawlers
Blocks 2 of 33 AI crawlers we track by name in its robots.txt. Checked Aug 2026.
Training crawlers
- Bytespider Blocked
- Applebot-Extended Allowed
- CCBot Allowed
- ClaudeBot Allowed
- FacebookBot Allowed
- Google-Extended Allowed
- GPTBot Allowed
- meta-externalagent Allowed
Search and AI assistants
- Amazonbot Blocked
- Claude-SearchBot Allowed
- OAI-SearchBot Allowed
- PerplexityBot Allowed
User-request fetchers
- ChatGPT-User Allowed
- Claude-User Allowed
- meta-externalfetcher Allowed
The other 18 AI crawlers we track are not mentioned and only inherit the catch-all rules that apply to every crawler.
- robots.txt found
- llms.txt found
- llms-full.txt found
- ai.txt not found
Content Signals:
search=yes ai-train=no declared in robots.txt
AI Crawlers
2 of 33 AI crawlers blocked
Standards: llms.txt, llms-full.txt, Content Signals
Audience
- 63% phone
- 37% desktop
India
Where the site's Chrome visitors come from, largest estimated audience first.
Open Page Rank
1.11 / 10 · ranked #10,936,585 of all domains
3 referring domains · OpenPageRank
Core Web Vitals
| LCP | 2.3 s |
| CLS | 0 |
| TTFB | 1100 ms |
Real user p75, Chrome UX Report, Aug 2026.
Page
- Language: English (United States)
- CDN: Cloudflare
- Weight: 0.6 MB · 36 requests
- Structured data: JSON-LD, OpenGraph, Twitter Cards
Homepage as crawled by the HTTP Archive, Aug 2026.