SUBSCRIBE
SUBSCRIBE
EXPLORE +
  • About infoDOCKET
  • Academic Libraries on LJ
  • Research on LJ
  • News on LJ
  • Skip to primary navigation
  • Skip to main content
  • Skip to primary sidebar
  • Libraries
    • Academic Libraries
    • Government Libraries
    • National Libraries
    • Public Libraries
  • Companies (Publishers/Vendors)
    • EBSCO
    • Elsevier
    • Ex Libris
    • Frontiers
    • Gale
    • PLOS
    • Scholastic
  • New Resources
    • Dashboards
    • Data Files
    • Digital Collections
    • Digital Preservation
    • Interactive Tools
    • Maps
    • Other
    • Podcasts
    • Productivity
  • New Research
    • Conference Presentations
    • Journal Articles
    • Lecture
    • New Issue
    • Reports
  • Topics
    • Archives & Special Collections
    • Associations & Organizations
    • Awards
    • Funding
    • Interviews
    • Jobs
    • Management & Leadership
    • News
    • Patrons & Users
    • Preservation
    • Profiles
    • Publishing
    • Roundup
    • Scholarly Communications
      • Open Access

February 26, 2014 by Gary Price

Research From Germany: Software Maps Ambiguous Names in Texts to the Right Person

February 26, 2014 by Gary Price

From University Saarland:

Computer scientists at the Max Planck Institute for Informatics in Saarbrücken have developed software that resolves the ambiguity of names within texts automatically. This mapping between mentions and actual entities like persons not only improves search engines, but also makes it possible to analyze huge amounts of text efficiently.
If a name is ambiguous and given without context, even humans struggle. When reading the last name “Merkel”, people do not know if it refers to the Chancellor of Germany Angela Merkel or the famous soccer coach Max Merkel. It is a drawback for web search, too. Up to now, the programs can capture character strings like “Angela Merkel”, but they do not pay attention to attributes like “German Chancellor” or “Germany’s First Lady” at all.
Even worse, after the word “Merkel” is entered, the search engines provide information about a lot of people with the same last name. Researchers at the Max Planck Institute for Informatics have now developed a program that enables accurate disambiguation of named entities by analyzing them with the help of the free Internet encyclopedia Wikipedia.
Their software named AIDA establishes connections between the mentions in the text and potential persons or places. “The more references exist between a mention and a specific person in Wikipedia, the more words of the person’s Wikipedia article can also be found in the input text, and the higher the score the mention-entity edge receives. AIDA checks this score and selects the mention-entity edge with the highest score as the accurate mapping,” explains Johannes Hoffart, who co-developed AIDA at the Max Planck Institute for Informatics. 
To demonstrate their novel technique, the researchers have implemented a search engine based on their approach.
The search engine makes it possible not only to combine the search for strings with the search for specific objects like persons and locations, but also to search on categories. In this way, the search for “Angela Merkel + phone call + Ukrainian politicians” results in texts dealing with the German Chancellor within the context of Ukrainian politicians like “Yulia Tymoshenko” and the string “phone call”. Currently the researchers use AIDA to analyze the text corpus of the German National Library to combine the search for keywords with the search for specific objects. “The search results are more precise this way”, Hoffart points out.
“With our new technique we can not only build better search engines, but also make computers understand texts almost as a human does, in an efficient way,” explains Gerhard Weikum, Scientific Director at the Max Planck Institute for Informatics in Saarbrücken. The approach also opens new possibilities for automatically generated recommendations and the analysis of datasets, says Weikum, who also does research at the Cluster of Excellence “Multimodal Computing and Interaction” in Saarbrücken. “Whoever is a fan of the soccer coach Merkel will receive recommendations for his books. Those more interested in the Chancellor get referred to books dealing with her and her way of governing Germany,” Weikum explains.
Both the software AIDA and its source code are available for the purposes of research.

Links

Demo the Technology
Access the Source Code (via Github)
Direct to Project Homepage

Filed under: Data Files, Libraries, Maps, National Libraries, News

SHARE:

About Gary Price

Gary Price (gprice@gmail.com) is a librarian, writer, consultant, and frequent conference speaker based in the Washington D.C. metro area. He earned his MLIS degree from Wayne State University in Detroit. Price has won several awards including the SLA Innovations in Technology Award and Alumnus of the Year from the Wayne St. University Library and Information Science Program. From 2006-2009 he was Director of Online Information Services at Ask.com. Gary is also the co-founder of infoDJ an innovation research consultancy supporting corporate product and business model teams with just-in-time fact and insight finding.

ADVERTISEMENT

Archives

Job Zone

ADVERTISEMENT

Related Infodocket Posts

NY Times: "New York Public Library Acquires Joan Didion’s Papers"

From The NY Times: When [Joan] Didion died in 2021 at age 87, the news set off an outpouring of tributes to a writer who fused penetrating insight and idiosyncratic personal voice, ...

University of North Carolina at Chapel Hill: María Estorino Named Vice Provost for University Libraries and University Librarian

Below, Find the Full Text of a Letter Sent to the Carolina Community From Kevin M. Guskiewicz University of North Carolina at Chapel Hill Chancellor Kevin M. Guskiewicz and J. ...

Boston Public Library Celebrates Black History Month with Annual “Black Is…” Booklist & Special Events

From the Boston Public Library: The Boston Public Library is proud to contribute to the celebration of Black History Month with its annual “Black Is…” booklist. The booklist aims to commemorate ...

Research Resources: New Online Tool Provides Health Snapshot of All 435 U.S. Congressional Districts (Congressional District Health Dashboard)

From NYU Langone: Researchers at NYU Grossman School of Medicine, in partnership with the Robert Wood Johnson Foundation, unveiled the Congressional District Health Dashboard (CDHD), a new online tool that ...

Report: "cOAlition S Confirms the End of Its Financial Support for Open Access Publishing Under Transformative Arrangements After...

From a cOAlition S  Announcement: Transformative arrangements – including Transformative Agreements and Transformative Journals – were developed to encourage subscription journals to transition to full and immediate open access within a defined timeframe (31st December 2024, ...

Library of Congress: Hannah Sommers Appointed New Associate Librarian for Researcher and Collections Services

From the Library of Congress: The Library of Congress announced today the appointment of Hannah Sommers as the new Associate Librarian for Researcher and Collections Services in the Library Collections and Services Group. In this role, Sommers will lead the future of the Library’s collections and the services it delivers to researchers and users. She will be central ...

Virginia Tech: University Libraries Dean Tyler Walters Appointed Board Chair of Academic Preservation Trust; IEEE Computer Society 2023...

As Book Bans Increase Across the Country, a Boston University Scholar is Fighting Back Core’s Library Resources & Technical Services Journal Goes Fully Open Access Digital Image Processing: It’s All ...

Funding: Library Freedom Project Receives $1 Million Grant Award From the Mellon Foundation to Advance Critical Privacy and...

Here’s the Full Text of the Library Freedom Project (LFP) Announcement:   Library Freedom Project (LFP) has been awarded $1,000,000 from the Mellon Foundation to expand the program’s work. For ...

Report: Sweden’s National Library Turns Page to AI to Parse Centuries of Data

From a NVIDIA Blog Post: For the past 500 years, the National Library of Sweden has collected virtually every word published in Swedish, from priceless medieval manuscripts to present-day pizza ...

IFLA Trend Report 2022 Released; Preprint: "The Semantic Scholar Open Data Platform"; & More Headlines

Archive for Amateur Radio Grows to 51,000 Items (via Internet Archive) Four New Appointments to the eLife New Board Members IFLA Trend Report 2022 Released (via International Federation of Library ...

Jennifer Vinopal Named HathiTrust's First Associate Director

Here’s the Full Text of Today’s HathiTrust Announcement: HathiTrust is pleased to announce that Jennifer Vinopal has been appointed HathiTrust’s first Associate Director.  Vinopal will assume a key leadership role ...

American Library Association Announces New $5.5 Million Transformational Grant From the Mellon Foundation

Here’s the Full Text of the ALA Announcement: The American Library Association (ALA) is pleased to announce a new grant in the amount of $5,515,000 from the Mellon Foundation to ...

ADVERTISEMENT

FOLLOW US ON TWITTER

Tweets by infoDOCKET

ADVERTISEMENT

This coverage is free for all visitors. Your support makes this possible.

This coverage is free for all visitors. Your support makes this possible.

Primary Sidebar

  • News
  • Reviews+
  • Technology
  • Programs+
  • Design
  • Leadership
  • People
  • COVID-19
  • Advocacy
  • Opinion
  • INFOdocket
  • Job Zone

Reviews+

  • Booklists
  • Prepub Alert
  • Book Pulse
  • Media
  • Readers' Advisory
  • Self-Published Books
  • Review Submissions
  • Review for LJ

Awards

  • Library of the Year
  • Librarian of the Year
  • Movers & Shakers 2022
  • Paralibrarian of the Year
  • Best Small Library
  • Marketer of the Year
  • All Awards Guidelines
  • Community Impact Prize

Resources

  • LJ Index/Star Libraries
  • Research
  • White Papers / Case Studies

Events & PD

  • Online Courses
  • In-Person Events
  • Virtual Events
  • Webcasts
  • About Us
  • Contact Us
  • Advertise
  • Subscribe
  • Media Inquiries
  • Newsletter Sign Up
  • Submit Features/News
  • Data Privacy
  • Terms of Use
  • Terms of Sale
  • FAQs
  • Careers at MSI


© 2023 Library Journal. All rights reserved.


© 2022 Library Journal. All rights reserved.