Feeds

IBM open sources enterprise search

Takes unstructured approach

  • alert
  • submit to reddit

Internet Security Threat Report 2014

IBM is open sourcing a jointly developed search architecture with a view to creating a common industry approach to querying unstructured enterprise data.

The company is today expected to announce plans to open source the Unstructured Information Management Architecture (UIMA) used in the company's WebSphere Information Integrator OmniFind Edition, WebSphere Portal Server and Lotus Workplace. IBM is also open sourcing the UIMA toolkit.

UIMA provides a framework for software tools and services capable of conducting context-based searches across millions of unstructured records, databases, content repositories and email systems.

UIMA was developed by IBM Research with "significant" input from the Defense Advanced Research Projects Agency (DARPA), along with other contributors.

Nelson Mattos, IBM distinguished engineer and vice president of information integration, said UIMA could return hundreds of relevant documents from a search query compared to a key word search-based approach that would return millions of documents.

According to Mattos, IBM hopes to create an industry standard for text analytics through the release of the code. "The key goal is to create a forum for other research institutions to contribute to and develop the framework without having to depend purely on IBM to support it," he said.

IBM also hopes to attract buy-in from the commercial sector. As such, IBM is today also expected to announce 15 companies will use UIMA as the framework for planned search and text analysis tools.

Open sourcing UIMA is the first-step in a process that could see IBM adopt existing industry standards for use with the architecture. IBM said it would investigate use of the Object Management Group's (OMG's) Unified Modeling Language (UML), eCore, and XML Metadata Interchange (XMI) with the UIMA's Common Analysis Structure (CAS) specification later this year. CAS handles data exchange between UIMA's various components. ®

Related stories

IBM 'really committed' to Java community
IBM and Google find each other in desktop search
Search pioneers join Yahoo! - but is the web beyond search?

Security for virtualized datacentres

More from The Register

next story
Microsoft WINDOWS 10: Seven ATE Nine. Or Eight did really
Windows NEIN skipped, tech preview due out on Wednesday
Business is back, baby! Hasta la VISTA, Win 8... Oh, yeah, Windows 9
Forget touchscreen millennials, Microsoft goes for mouse crowd
Apple: SO sorry for the iOS 8.0.1 UPDATE BUNGLE HORROR
Apple kills 'upgrade'. Hey, Microsoft. You sure you want to be like these guys?
ARM gives Internet of Things a piece of its mind – the Cortex-M7
32-bit core packs some DSP for VIP IoT CPU LOL
Microsoft on the Threshold of a new name for Windows next week
Rebranded OS reportedly set to be flung open by Redmond
Lotus Notes inventor Ozzie invents app to talk to people on your phone
Imagine that. Startup floats with voice collab app for Win iPhone
'Google is NOT the gatekeeper to the web, as some claim'
Plus: 'Pretty sure iOS 8.0.2 will just turn the iPhone into a fax machine'
prev story

Whitepapers

Forging a new future with identity relationship management
Learn about ForgeRock's next generation IRM platform and how it is designed to empower CEOS's and enterprises to engage with consumers.
Storage capacity and performance optimization at Mizuno USA
Mizuno USA turn to Tegile storage technology to solve both their SAN and backup issues.
The next step in data security
With recent increased privacy concerns and computers becoming more powerful, the chance of hackers being able to crack smaller-sized RSA keys increases.
Security for virtualized datacentres
Legacy security solutions are inefficient due to the architectural differences between physical and virtual environments.
A strategic approach to identity relationship management
ForgeRock commissioned Forrester to evaluate companies’ IAM practices and requirements when it comes to customer-facing scenarios versus employee-facing ones.