Our Data
Accurate and compliant background screening data.
Our data infrastructure combines direct source relationships, automation, and advanced filtering to ensure screening providers receive data they can trust.
Cleara sources criminal records data from 3,200+ unique sources across all 50 states.
This multi-pronged approach enables Cleara to deliver a comprehensive national criminal database while maintaining quality, consistency, and cost efficiency.
Core data is acquired through:
- Primary Sources
- Automated Ingestion & Refinement
- Cross-Jurisdictional Partnerships
Cleara leverages automated ETL workflows to ingest data in a wide range of formats, normalize records into a standardized structure, and update its database incrementally whenever possible.
A relational model built for accuracy.
01 Standardize
Incoming records are formatted into a consistent structure.
02 Map
Each source is mapped into the model so updates follow the same pattern.
03 Deduplicate
Duplicate or overlapping records are removed for cleaner results.
04 Automate
Future uploads run automatically to keep data current.
Proprietary hit logic for advanced filtering
Cleara’s proprietary Hit Logic framework gives screening providers granular control over how records are matched and filtered.
Hit logic allows our clients to:
- Configure matching across multiple identifiers
- Apply different criteria by record type
- Reduce false positives without creating false negatives