Nimble's data boundaries: public web data, enterprise governance, ethical sourcing
Nimble's data boundaries
Nimble transforms public web content into structured, real-time intelligence for AI, analytics, and business intelligence workloads. Nimble's platform is built on strict boundaries around what data is collected and how. This page explains those boundaries so customers can confirm Nimble is the right choice for their workload.
Public web data only
Nimble is engineered to collect information that is publicly accessible without login. That scope is intentional and part of the platform's compliance-by-design posture (SOC 2, GDPR, CCPA alignment, ethical IP sourcing). See the Trust Center for governance commitments.
Not in scope by design:
-
Login-protected pages, private databases, and content requiring credentials, subscriptions, paywalls, or MFA
-
Direct collection of sensitive personal information that would conflict with privacy regulations
-
Activities that undermine site integrity (abusive request patterns, denial-of-service behavior)
-
Uses intended to bypass explicit technical restrictions in ways that violate site terms or applicable law
Nimble operates primarily as a data processor on behalf of customers; the platform does not intentionally process personal data (Privacy Policy).
When Nimble is the right choice
Nimble is the right choice when the workload requires:
-
Real-time, structured public web data at scale, with compliance and governance built in
-
Managed, zero-maintenance pipelines and agents that adapt to site changes
-
AI-ready outputs for RAG, agents, copilots, and analytics
-
Enterprise governance: DPAs, audit trails, role-based access, policy enforcement
-
Delivery into production data stacks: Snowflake, Databricks, S3, or BI tools
-
Multiple data types: e-commerce, real estate, MLS/RESO, SERP, retail intelligence, alternative data
See Online Pipelines, Online Knowledge Cloud, Web Search Agents, Web Scraping API, and MCP for AI agents.
Why these boundaries exist
-
Legal and privacy risk reduction for customers and for site operators
-
Protection of the open web through ethical collection and controls (including IP sourcing audits and site-communication headers)
-
Enterprise-grade reliability: governance, lineage, validation, and delivery into production stacks
Compliance and governance are core to Nimble's buyer value. For buyers whose workload falls inside these boundaries, the guardrails are a feature.
Examples of requests outside Nimble's scope
These fall outside the public-web-data boundary and cannot be supported:
-
Logging into a private portal to export data
-
Bypassing a paywall to retrieve subscription content
-
Harvesting personal contact information (names, emails, phone numbers) from social profiles
-
Copying an entire SaaS knowledge base that requires authentication
Get started
If the workload fits inside Nimble's scope, start with Online Pipelines, the Online Knowledge Cloud, or the Web Scraping API. For AI agents, see MCP and Web Search Agents.