Know what's in it before your customers do.
As the publisher, you carry controller liability for personal data in published datasets, regardless of what your suppliers told you.
PII in shared datasets creates liability for the sharing party. Technical controls at the point of publication are the appropriate measure.
Healthcare, financial, and legal datasets carry additional vertical-specific obligations on top of GDPR.
Scanning triggered on S3 event, SFTP arrival, or API upload. No polling. Webhook callbacks on completion. Designed for high-throughput ingestion pipelines.
Each publisher tenant gets isolated scanning context, results, and audit records. No cross-contamination of findings or audit data.
Python, Node.js, Java, and .NET SDK clients. OpenAPI spec for custom integration. Embed the detection engine directly inside your pipeline if needed.
GLiNER v2 handles any sector-specific entity type without retraining. Health identifiers, financial references, legal codes: all from the same engine.
Raw datasets with PII are never published. Governed clean copies generated automatically and placed in the publish queue. Full audit trail per dataset.
Complete REST API with OpenAPI 3.1 spec. Generates typed clients in any language. Integrates with your existing API gateway and authentication layer.
VestraShield intercepts every AI-assisted dataset review, analysis, and enrichment step. PII in the datasets your team is working with doesn't reach any LLM endpoint without being governed first.
For engineering teams building data marketplaces. SDK documentation and pipeline integration questions welcome.