Patentable/Patents/US-9824160
US-9824160

Computer implemented method and device for accessing a data set

PublishedNovember 21, 2017
Assigneenot available in USPTO data we have
Inventorsnot available in USPTO data we have
Technical Abstract

A computer implemented method of accessing a data set comprising a plurality of records, wherein each record is associated with one or more items of data. The method comprises using the computer to receive a data query on the data set. Each record is assigned to an in-group or to an out-group with respect to the query. Words appearing in records of the in-group are determined and a user interface representative of said words is generated. Words appearing in records of the out-group are determined and a user interface representative of said words is generated.

Patent Claims
11 claims

Legal claims defining the scope of protection, as filed with the USPTO.

1

1. A non-transitory computer readable medium storing instructions for causing an electronic processor to access a data set that includes a plurality of records, wherein each record is associated with at least one item of data, the instructions causing the processor in real-time to perform the following steps: preprocess, where a conversion unit converts the plurality of records into a list of representations and concordance; receive a data query on the data set; assign each record that complies with the query to an in-group using an assignation unit; assign each record that does not comply with the query to an out-group using the assignation unit; determine, for each of said at least one item of data, a first indicator and a second indicator, wherein, for the entire data set, only NI sets of the first indicator and the second indicator are determined, where NI is a number of items of the data in the concordance; determine, for each of said at least one item of data a score S representative of a discriminative power of the at least one tern of data, wherein the score is calculated using a formula S=(I 1 1.5 −I 2 1.5 )/(I 1 +I 2 ); wherein I 1 is the first score and I 2 is the second score; determine a first plurality of words appearing in records of the in-group; determine a second plurality of words appearing in records of the out-group; and generate a displayable user interface, which includes a dual word cloud, representative of the first plurality of words and the second plurality of words such that each of the first plurality of words share a common trait so as to be distinguishable from each of the second plurality of words, wherein: the user interface includes a first view comprising data representative of the in-group and the out-group, and the user interface includes at least one further view including data representative of data representative of the records in different formats, the dual word cloud includes the first plurality of words relating to the in-group of records that comply with a particular query. the dual word cloud includes the second plurality of words relating to the out-group of records that do not comply with the query, and the dual word cloud includes both words having high discriminative powers for the records of the in-group and the out-group, and all views are updated upon user selection of at least one item of data in one of the views.

2

2. The non-transitory computer readable medium of claim 1 , wherein the at least one further view comprises data representative of geographical information relating to the records, temporal information relating to the records, or relational information relating to the records.

3

3. The non-transitory computer readable medium of claim 1 , wherein the first view and the at least one further view are coupled.

4

4. The non-transitory computer readable medium of claim 1 , wherein the data query on the data set includes selection of one or more data items in the first view or one of the at least one further views.

5

5. The non-transitory computer readable medium of claim 4 , wherein all views are updated instantaneously.

6

6. The non-transitory computer readable medium of claim 1 , the instructions causing the processor to: determine a first plurality of words having a high discriminative power favoring records of the in-group and generate a user interface representative of said words; and determine a second plurality of words having a high discriminative power favoring records of the out-group and generate a user interface representative of said words.

7

7. A non-transitory computer readable medium storing instructions for causing an electronic processor to access a data set that include a plurality of records, wherein each record is associated with at least one item of data, the instructions causing the processor in real-time to perform the following operations: preprocess, where a conversion unit converts the plurality of records into a list of representations and concordance; receive a data query on the data set; assign each record that complies with the query to a first group using an assignation unit; assign each record that does not comply with the query to a second group using the assignation unit; determine, for each of that at least one item of data, a first indicator and a second indicator, wherein, for the entire data set, only NI sets of the first indicator and the second indicator are determined, where NI is a number of iters of the data in the concordance; determine, for each of the at least one tern of data, a score S representative of a discriminative power of that said at least one item of data, wherein the score is calculated using a formula S+(I 1 1.5 −I 2 1.5 )/(I 1 +I 2 ); wherein I 1 is the first score and I 2 is the second score; determine a first plurality of items of data appearing in records of the first group; determine a second plurality of items of data appearing in records of the second group; and generate a displayable user interface representative of the first plurality of items and the second plurality of items such that each of the first plurality of items share a common trait so as to be distinguishable from each of the second plurality of items, wherein: the user interface includes a first view including data representative of the first group and the second group, and the user interface includes at least one further view comprising data representative of data representative of the records in different formats, the dual word cloud includes the first plurality of words relating to the first group of records that comply with a particular query, the dual word cloud includes the second plurality of words relating to the second group of records that do not comply with the query, and the dual word cloud includes both words having high discriminative powers for the records of the first group and the second group, and all views are updated upon user selection of at least one item of data in one of the views.

8

8. The non-transitory computer readable medium of claim 7 , wherein the items of data are one or more of words, groups of words, texts, image fragments, images, video fragments, audio fragments, numbers, chemical formula fragments, chemical formulae, mathematical formula fragments, mathematical formulae.

9

9. A non-transitory computer readable medium storing instructions for causing an electronic processor to generate a user interface including data representative of a reference item of data included in a record in a data set comprising a plurality of records, the instructions causing the processor in real-time to perform the following operations: preprocess, where a conversion unit converts the plurality of records into a list of representations and concordance; determine, for each of said at least one item of data, a first indicator and a second indicator, wherein, for the entire data set, only NI sets of the first indicator and the second indicator are determined, where NI is a number of items of the data in the concordance; determine, for each of the at least one item of data, a score S representative of a discriminative power of that said at least one item, of data, wherein the score is calculated using a formula S=(I 1 1.5 −I 2 1.5 )/(I 1 +I 2 ); wherein I 1 is the first score and I 2 is the second score: determine a first plurality of items of data appearing in records including the reference item; determine a second plurality of items of data appearing in records not including the reference item; and generate a displayable user interface representative of the first plurality of items and the second plurality of items such that each of the first plurality of items share a common trait so as to be distinguishable from each of the second plurality of items, wherein: the user interface includes a first view comprising data representative of an in-group and an out-group, and the user interface includes at least one further view comprising data representative of data representative of the records in different formats, the dual word cloud includes the first plurality of words relating to the in-group of records that comply with a particular query, the dual word cloud includes the second plurality of words relating to the out-group of records that do not comply with the query, and the dual word cloud includes both words having high discriminative powers for the records of the in-group and the out-group, and all views are updated upon user selection of at least one item of data in one of the views.

10

10. A data processing system for accessing a data set including a plurality of records, wherein each record is associated with at least one item of data, the system comprising: a conversion unit for preprocessing a conversion of the plurality of records into a list of representations and concordance; an input unit for receiving a data query on the data set; a an assignation unit for assigning each record to one of a first group and to a second group with respect to the query; and a processing unit for determining, for each of said at least one itern of data, a first indicator and a second indicator, wherein: for the entire data set, only NI sets of the first indicator and the second indicator are determined where the NI is a number of items of the data in the concordance; the processing unit determines, for each of said at least one item of data, a score S representative of a discriminative power of the at least one item of data; the score is calculated using a formula S=(I 1 1.5 −I 2 1.5 )/(I 1 +I 2 ); wherein I 1 is the first score and I 2 is the second score; the processing unit is arranged for determining a first plurality of items of data appearing in records of the first group and a second plurality of items of data appearing in records of the second group, and generating a displayable user interface representative of the first plurality of items and the second plurality of items such that each of the first plurality of items share a common trait so as to be distinguishable from each of the second plurality of items, the user interface includes a first view including data representative of the first group and the second group, and the user interface includes at least one further view comprising data representative of data representative of the records in different formats, the dual word cloud includes the first plurality of words relating to the first group of records that comply with a particular query, the dual word cloud includes the second plurality of words relating to the second group of records that do not comply with the query, and the dual word cloud includes both words having high discriminative powers for the records of the first group and the second group, and all views are updated upon user selection of at least one item of data in one of the views.

11

11. A non-transitory computer readable medium storing computer implementable instructions, which when implemented by a programmable computer, cause the computer in real-time to perform the following operations; preprocess, where a conversion unit converts the plurality of records into a list of representations and concordance; receive a data query on the data set; assign each record to one of a first group and to a second group; determine for each of said at least one item of data a first indicator and a second indicator, wherein, for the entire data set, only NI sets of the first indicator and the second indicator are determined, where NI is a number of items of the data in the concordance; determine, for each of the at least one item of data, a score S representative of a discriminative power of that the at least one item of data, wherein the score is calculated using a formula S=(I 1 1.5 −I 2 1.5 )/(I 1 +I 2 ); wherein I 1 is the first score and I 2 is the second score; determine a first plurality of items of data appearing in records of the first group; determine a second plurality of items of data appearing in records of the second group; and generate a displayable user interface representative of the first plurality of words and the second plurality of words such that each of the first plurality of words share a common trait so as to be distinguishable from each of the second plurality of words, wherein; the user interface includes a first view comprising data representative of the first group and the second group, and the user interface includes at least one further view comprising data representative of data representative of the records in different formats, the dual word cloud includes the first plurality of words relating to the first group of records that comply with a particular query, the dual word cloud includes the second plurality of words relating to the second group of records that do not comply with the query, and the dual word cloud includes both words having high discriminative powers for the records of the first group and the second group, and all views are updated upon user selection of at least one item of data in one of the views.

Classification Codes (CPC)

Cooperative Patent Classification codes for this invention. Click any code to explore related patents in that topic.

Patent Metadata

Filing Date

June 2, 2014

Publication Date

November 21, 2017

Want to explore more patents?

Browse 5M+ US patents with plain-English claim translations and AI-generated analysis.

Citation & reuse

Analysis on this page is generated by Patentable — an AI-powered patent intelligence platform. AI-generated summaries, explanations, and analysis may be reused with attribution and a visible link back to the canonical URL below. Patent abstracts and claims are USPTO public domain.

Cite as: Patentable. “Computer implemented method and device for accessing a data set” (US-9824160). https://patentable.app/patents/US-9824160

© 2026 Patentable. All rights reserved.

Patentable is a research and drafting-assistant tool, not a law firm, and does not provide legal advice. Documents we generate are drafts for review by a licensed patent attorney.