Inverted Files, Parsing, Discovery, and Clustering

Last time, I told you that inverted files and Boolean are basic to most or all search. So, what we have is the inverted file – that big alphabetical list of all the terms and where they came from – and a way to combine them. At the end of the day, no matter which direction you come from on the high end and even higher presentation levels, you are depending on these two to make it work. If you add that taxonomy to that, you are strengthening the use of your terms.

Read More

T-Minus: One Week and Counting!

We’re busy putting the final preparations on the Ninth Annual Data Harmony Users Group Meeting. This year’s meeting is gearing up to be our biggest and best meeting yet! If we can assist you as you plan your trip to Albuquerque, please contact Heather Kotula or Susan Burritt at any time.

Read More

Access Innovations Partners with Leading Scientific Organizations for the Launch of a New Thesaurus Created for Astronomy Community

Access Innovations, Inc., a leader in semantic enrichment, has teamed up with the American Institute of Physics (AIP) and IOP Publishing (IOP) to create a new astronomy thesaurus called the Unified Astronomy Thesaurus (UAT) for the American Astronomical Society (AAS) that will help improve future information discovery for researchers.

Read More

Kinds of Search

There are search systems that are advertised as keyword search. There are ones that say that they are Bayesian. There are ones that are Boolean. There are ones that are primarily ranking algorithms.

Read More

Measuring Accuracy in Search

When people talk about how accurate the search is, there are lots of different ways to measure that. This list indicates some of the ways that we talk about measuring accuracy.

Read More

Don’t Miss the Data Harmony Users Group Meeting on February 18-20, 2013

On Tuesday evening, join us for dinner at the Albuquerque Aquarium, part of the ABQ BioPark, along the majestic Rio Grande in Albuquerque, New Mexico. Our dinner site is next to the 285,000-gallon shark tank, which is home to five species of sharks, several sea turtles, dozens of manta rays and even a few barracudas. (Fear not! Our dinner menu will be different from the inhabitants of the tank! We do not plan to have fish on the menu Tuesday night!)

Read More

How Search Works

Search has many parts. The parts of search are moving parts and every system does it a little differently. There’s the search software, based on one of two major camps. Then, there’s the computer network that it’s riding on. Then, there’s the way that the text is parsed, which is not always the same.

Read More

Don’t Miss the Data Harmony Users Group Meeting on February 18-20, 2013

Another exciting case study…
MAI in Amazon Cloud using Amazon Web Services
Nigel Kerr, JSTOR, and Bob Kasenchak, Access Innovations, Inc., will describe the revolutionary way MAI is being implemented by the JSTOR collection, which numbers about 8.5 million academic articles. Over the course of the JSTOR thesaurus project, we have to use the emergent thesaurus to index these articles several times for testing and implementation. One significant challenge is: how we can index and re-index this vast volume of data without slowing down the thesaurus assembly efforts? To meet this challenge we are using Amazon Web Services (aka “the Amazon cloud”), in which we can “rent” processing time and storage on-demand. This allows us use a large number of computer resources for a very short period of time to get the indexing done. Just another example of how MAI is being adapted to meet our clients needs!

Read More