import.io
Link authority for import.io (registrable domain: import.io) from the Common Crawl web graph: PageRank and harmonic centrality history since 2017, current scores, and the site's top backlink sources.
Web-graph authority (0-100)
Domain scores measure import.io as a whole: every subdomain's links rolled up into one node of Common Crawl's domain-level web graph. Host scores measure the exact hostname import.io. WWW scores cover the www variant of that hostname, which the graph tracks as its own entry since January 2026 (before that, www and bare host were merged). A big gap between domain and host usually means the domain's authority lives on other subdomains. PageRank is link-endorsement authority: how much weight the sites linking here carry themselves. Harmonic centrality measures reach: how close this site sits to the rest of the web when following links, which favors broadly connected sites over ones boosted by a few heavy linkers. All are normalized 0 to 100 within each crawl snapshot; a dash means that series does not exist for this site.
Homepage
What import.io presents on its own homepage: the page title, description, and the social share markup it publishes. Pulled from https://import.io.
| Title | Home New |
|---|---|
| Description | Import.io helps businesses scale web scraping and data extraction for real-time pricing, competitor tracking, and high-quality web data. |
| OG Title | Import.io | Web Scraping and Data Extraction at Scale |
| OG Description | Managed web scraping and data extraction for pricing intelligence, competitor monitoring, and digital shelf analytics. |
| Twitter Title | Import.io | Web Scraping and Data Extraction at Scale |
| Twitter Card | summary_large_image |
Social accounts and contact points
Where this site maintains a presence and every reachable route its homepage exposes: social accounts, contact pages, addresses, and content feeds.
| X | importio |
|---|---|
| import-io (business) | |
| Youtube | importiovideos |
| Contact pages |
https://import.io/contact/contact-sales https://www.import.io/contact/contact-sales |
|---|---|
| Blog | https://import.io/resources/blog |
| Feeds | https://cdn.import.io/sitemap.xml |
Structured data (JSON-LD)
The schema.org markup this homepage publishes: the machine-readable identity search engines and AI systems read. Types found: SoftwareApplication, Organization, FAQPage.
JSON-LD block (1,616 bytes)
{
"@context": "https://schema.org",
"@type": "SoftwareApplication",
"name": "Import.io",
"description": "AI-powered web data extraction platform that transforms raw web data into compliant, AI-native intelligence streams for analytics, AI, and enterprise decision systems",
"url": "https://www.import.io/",
"applicationCategory": "BusinessApplication",
"operatingSystem": "Web",
"offers": {
"@type": "Offer",
"availability": "https://schema.org/InStock",
"url": "https://www.import.io/pricing"
},
"featureList": [
"AI-Native Automation with prompt-based data extraction",
"Self-healing pipelines with continuous monitoring",
"Compliance-first filters with PII masking",
"GDPR and CCPA compliant data extraction",
"Complete audit trails and dashboards",
"Secure data delivery pipelines",
"Enterprise-grade reliability"
],
"provider": {
"@type": "Organization",
"name": "Import.io",
"logo": {
"@type": "ImageObject",
"url": "https://cdn.prod.website-files.com/64fb455a1764c0e56de8a1cb/64fb455a1764c0e56de8a1fa_import-logo.svg"
},
"sameAs": []
},
"audience": [
{
"@type": "Audience",
"audienceType": "Retail & E-Commerce"
},
{
"@type": "Audience",
"audienceType": "Finance & Alternative Data"
},
{
"@type": "Audience",
"audienceType": "Healthcare & Life Sciences"
},
{
"@type": "Audience",
"audienceType": "Climate & ESG"
},
{
"@type": "Audience",
"audienceType": "Legal & Regulatory"
}
],
"inLanguage": "en-US"
}
JSON-LD block (889 bytes)
{
"@context": "https://schema.org",
"@type": "Organization",
"@id": "https://www.import.io/#organization",
"name": "Import.io",
"url": "https://www.import.io",
"logo": "https://www.import.io/images/importio-logo.png",
"sameAs": [
"https://www.linkedin.com/company/import-io/",
"https://twitter.com/importio",
"https://www.youtube.com/@importio",
"https://www.facebook.com/Importio"
],
"description": "Import.io is an AI-powered web data extraction platform that helps enterprises collect, structure, and integrate web data at scale.",
"foundingDate": "2012",
"founder": {
"@type": "Person",
"name": "David White"
},
"contactPoint": {
"@type": "ContactPoint",
"contactType": "Sales",
"email": "sales@import.io",
"url": "https://www.import.io/contact",
"areaServed": "Global",
"availableLanguage": ["English"]
}
}
JSON-LD block (1,991 bytes)
{
"@context": "https://schema.org",
"@type": "FAQPage",
"mainEntity": [
{
"@type": "Question",
"name": "How does Import.io use AI in data extraction?",
"acceptedAnswer": {
"@type": "Answer",
"text": "Import.io uses machine learning to detect patterns, adapt to changes in website layouts, and validate results automatically. This means pipelines don’t just run, they adjust themselves to keep your data accurate and consistent."
}
},
{
"@type": "Question",
"name": "Can Import.io provide training data for AI models?",
"acceptedAnswer": {
"@type": "Answer",
"text": "Yes. Import.io delivers model-ready datasets that are clean, structured, and documented. Our customers use this data for fine-tuning large language models, powering retrieval-augmented generation (RAG) systems, and building custom AI solutions."
}
},
{
"@type": "Question",
"name": "What does “self-healing pipelines” mean?",
"acceptedAnswer": {
"@type": "Answer",
"text": "When a website changes, most extractors break. Import.io pipelines don’t. They automatically detect drift, re-map fields, and continue delivering data. Combined with continuous monitoring, this ensures your AI systems and dashboards are always fed with reliable inputs."
}
},
{
"@type": "Question",
"name": "How does Import.io help me use AI from the very beginning?",
"acceptedAnswer": {
"@type": "Answer",
"text": "Import.io takes an AI-first approach to every project. Instead of starting with manual scoping or guesswork, you can engage directly with our platform using natural language to define your requirements. The AI helps refine the scope, shape the schema, and even build your first proof-of-concept automatically. This way, you see value faster, and the POC already reflects the exact business outcomes you care about."
}
}
]
}
Authority over time
Every Common Crawl web-graph release since 2017, roughly quarterly. The domain series goes back to 2017; Common Crawl only began publishing the host-level graph in 2020, so host lines start there. Click a legend entry to show or hide that line. Flat lines are normal for established sites; watch for sustained rises or drops, which track real gains or losses in who links here.
A note on www: through the end of 2025 the source web graph merged www and non-www hostnames into the bare name, so one host line covers both. Starting in January 2026 the graph splits them into separate entries, which is why www can carry its own scores going forward.
Backlink counts
Unique linking domains seen in Common Crawl's link graphs. This is a different dataset than the 0-100 authority scores above: these are raw counts, host-level for import.io and domain-level for import.io with all its subdomains combined. Common Crawl covers a large sample of the web, not all of it, so treat counts as comparable between sites rather than absolute totals.
Top backlinks in the domain graph
Common Crawl publishes two link graphs. The DOMAIN graph connects whole domains: every link from any page of one domain to any page of another. This table shows the highest-authority domains linking to import.io (including its subdomains), ranked by the linker's PageRank. The table shows the top 851; the download carries the full stored list (up to 10,000 linkers) as JSON. Click a linker for its own site profile.
Top backlinks in the host graph
The HOST graph connects exact hostnames, so subdomains count separately. This table shows the highest-authority hosts linking to import.io specifically, ranked by the linker's PageRank. The table shows the top 281; the download carries the full stored list (up to 10,000 linkers) as JSON. Click a linker for its own site profile.