Information infrastructure management tools with variable and configurable filters and segmental data stores

Patent No. US10182073 (titled "Information infrastructure management tools with variable and configurable filters and segmental data stores") on Jan 15, 2015. The application was issued on Jan 15, 2019.

What is this patent about?

’073 is related to the field of information management and data security within distributed computing systems. It specifically addresses the challenges of managing unstructured and semi-structured data—such as emails, word processing documents, and web content—which often contain sensitive trade secrets or mission-critical information. The invention provides a framework for identifying, classifying, and isolating this content to prevent unauthorized access, accidental loss, or disclosure through open enterprise ecosystems.

The underlying idea behind ’073 is the granular deconstruction of data streams into atomic elements to separate high-value content from common remainder data. Rather than relying on traditional perimeter defenses or simple file-level encryption, the system uses dynamic categorical filters to identify specific words, images, or data objects. By stripping sensitive elements from a source document and replacing them with placeholders, the invention creates a 'formless' data structure where the most critical information is physically and logically isolated from the rest of the file.

The claims of ’073 focus on a method for creating an information infrastructure that processes data throughput using a plurality of configurable filters. These claims cover the identification of sensitive and select content, the grouping of this data into multiple sensitivity levels, and the use of distributed data stores specifically designated for different classifications. A key aspect of the independent claims is the ability to modify these filters by expanding, contracting, or imposing hierarchical and orthogonal classifications to organize further data throughput.

In practice, the invention works by passing data through a multi-tiered filtering process that includes content-based, contextual, and taxonomic analysis. When a filter detects sensitive content, it extracts those specific elements and disperses them to secure storage locations across a network. This process effectively 'sanitizes' the original document, leaving behind a remainder file that is safe for general distribution. To view the complete information, a user must possess the appropriate security clearance to trigger a reconstruction process that pulls the granular pieces back from their respective data stores.

This approach differs from prior solutions by moving away from static classification labels, which are vulnerable to tampering, and instead focusing on the semantic essence of the data itself. By utilizing adaptive filters that can be expanded or contracted based on operator selection or environmental triggers, the system can respond to real-time threats or changing enterprise policies. This creates a resilient ecosystem where the value of information is protected through fragmentation and dispersal, ensuring that even if a single storage node is compromised, the full context of the sensitive information remains hidden.

How does this patent fit in bigger picture?

Technical Landscape

In the late 2000s when ’073 was filed, enterprise information management was typically implemented using centralized indexing and firewalls at a time when systems commonly relied on perimeter-based security rather than granular data-level controls. During this era, the proliferation of unstructured data across distributed environments—such as portable media, email bodies, and diverse document formats—made the consistent enforcement of corporate security policies non-trivial, as hardware and software constraints often limited real-time semantic analysis and automated classification across open ecosystems.

Prosecution Position

The disclosed invention represents a meaningful technical advancement through the integration of dynamic, adaptive categorical filters—including content-based, contextual, and taxonomic classification filters—directly into a distributed data processing architecture. By transposing security-sensitive content into designated distributed stores while maintaining remainder data in the system, the architecture enables a transformation of data that achieves higher levels of organization and security. This structural shift allows for the automated implementation of complex enterprise policies, such as data retention and privacy compliance, by associating specific data processes like extraction or destruction with the output of the categorical filters, thereby overcoming the technical constraint of managing sensitive information within unstructured and semi-structured data streams.

Claims

This patent contains 20 claims, with claims 1, 11, and 20 serving as the independent claims. The independent claims focus on a method for establishing an information infrastructure that processes data throughput in a distributed computing system by utilizing a plurality of filters to identify and organize sensitive and select content. These claims specifically address the dynamic modification of filters through expansion, contraction, or classification changes to generate modified configurations for organizing data. The dependent claims further refine this process by introducing operator commands, event triggers, policy-level classifications, and the integration of an inference engine to generate relevant keywords for filter supplementation and data reorganization.

Key Claim Terms New

Definitions of key terms used in the patent claims.

Term (Source)Support for SpecificationInterpretation
Hierarchical or an orthogonal classification
(Claim 1, Claim 11, Claim 20)
Ownership classification may include hierarchical or orthogonal security classification. The system enables access and search techniques to be granular and multi-level, representing five informational attributes. This allows for the separation of strategic information from tactical information so that access is granular by role and user.A multi-dimensional categorization scheme applied to filters to organize data based on either ranked levels of authority/importance (hierarchical) or independent, non-overlapping categories (orthogonal).
Initially configured filters
(Claim 1, Claim 11)
A data input is processed through at least one activated categorical filter to obtain select content. The system provides data processing tools to manage and organize data processed by an enterprise using these basic filter modules. These filters screen data for enterprise policies such as privacy, financial data handling, and document retention.The primary set of screening modules (categorical, sensitivity, or taxonomic) that are first applied to a data stream before being modified by expansion, contraction, or classification changes.
Saved profiles
(Claim 20)
The system and process can be pre-set to automatically trigger a keyword search based on preset filtering processes. The user has the ability to set the system ON for a continuous, non-stop cycle of filtering keywords. Generating modified filters as saved profiles for future use enables the enterprise to establish a policy for that information and implement the policy recommendation.Stored configurations of modified filters, resulting from operator-selected alterations (such as content expansion or time limits), preserved for future data throughput organization.
Select content
(Claim 1, Claim 11, Claim 20)
Select content (SC) is represented by one or more predetermined words, characters, images, data elements or data objects. The system employs a dynamic, adaptive filter to enhance select content collection and classification systems to organize such SC. It is stored in select content data stores separately from security sensitive content.Specific data elements represented by predetermined words, characters, or objects that are important to an enterprise and are organized via categorical filters (content-based, contextual, or taxonomic).
Sensitive content
(Claim 1, Claim 11, Claim 20)
The invention relates to the protection of confidential information and identification of such information. Sensitive content (sec-con) includes secret or security sensitive data in the enterprise computer system. This content is extracted from a data input to obtain extracted security sensitive data for a corresponding security level and remainder data.Data elements (words, characters, images, or objects) identified as confidential or security-sensitive that are grouped into specific sensitivity levels for extraction and storage in secure data stores.
Sensitivity levels
(Claim 1, Claim 11, Claim 20)
Extracted security sensitive data is stored for a corresponding security level. The controlled release of corresponding extracted security sensitive data from the respective extract stores is permitted based on associated security clearances for corresponding security levels. Security levels of a document or data stream can be upgraded or downgraded based on the results of inference tests.Distinct categories or tiers assigned to sensitive content, each potentially associated with different security clearances, used to determine how data is extracted and stored.

Litigation Cases New

US Latest litigation cases involving this patent.

Case NumberFiling DateTitle
2:25-cv-00595Apr 18, 2025Digitaldoors, Inc. v. SouthPoint Bank
8:25-cv-00002Jan 1, 2025Digital Doors, Inc. v. Sandy Spring Bank

Patent Family

Patent Family

File Wrapper

The dossier documents provide a comprehensive record of the patent's prosecution history - including filings, correspondence, and decisions made by patent offices - and are crucial for understanding the patent's legal journey and any challenges it may have faced during examination.

  • Get instant alerts for new documents

US10182073

Application Number
US14597314A
Filing Date
Jan 15, 2015
Publication Date
Jan 15, 2019
External Links
Slate, USPTO , Google Patents