research-article

Semantic Annotation of Heterogeneous Data Sources: Towards an Integrated Information Framework for Service Technicians

Authors:
Sebastian Bader

Karlsruhe Institute of Technology, Karlsruhe, Germany

Karlsruhe Institute of Technology, Karlsruhe, Germany
View Profile

,
Jan Oevermann

University of Bremen & Karlsruhe University of Applied Sciences, Karlsruhe, Germany

University of Bremen & Karlsruhe University of Applied Sciences, Karlsruhe, Germany
View Profile

Semantics2017: Proceedings of the 13th International Conference on Semantic SystemsSeptember 2017Pages 73–80https://doi.org/10.1145/3132218.3132221

Published:11 September 2017Publication History

Semantics2017: Proceedings of the 13th International Conference on Semantic Systems

Pages 73–80

ABSTRACT

Service technicians in the domain of industrial maintenance require extensive technical knowledge and experience to complete their tasks. Some of the needed knowledge is made available as document-based technical manuals or reports from previous deployments. Unfortunately, due to the great amount of data, service technicians spend a considerable amount of working time searching for the correct information. Another challenge is posed by the fact that valuable insights from operation reports are not yet considered due to insufficient textual quality and content-wise ambiguity.

In this work we propose a framework to annotate and integrate these heterogeneous data sources to make them available as information units through Linked Data technologies. We use machine learning to modularize and classify information from technical manuals together with ontology-based autocompletion to enrich reports with clearly defined concepts. By combining both approaches we can provide an unified and structured interface for manual and automated querying. We verify our approach by measuring precision and recall of information for typical retrieval tasks for service technicians, and show that our framework can provide substantial improvements for service and maintenance processes.

References

2006/42/EC. 2006. Machinery directive of the European Parliament and of the Council. (2006).Google Scholar
Adobe Systems (Ed.). 2001. PDF reference: Adobe portable document format version 1.4 (3rd ed ed.). Addison-Wesley, Boston.Google Scholar
Donald F. Blumberg. 1994. Strategies for Improving Field Service Operations Productivity and Quality. The Service Industries Journal 14, 2 (1994), 262--277.Google ScholarCross Ref
AJ Brush, David Bargeron, Anoop Gupta, and Jonathan J Cadiz. 2001. Robust annotation positioning in digital documents. In Proceedings of the SIGCHI conference on Human factors in computing systems. ACM, 285--292. Google ScholarDigital Library
Fabrice Colas, Pavel Paclík, Joost N. Kok, and Pavel Brazdil. 2007. Does SVM Really Scale Up to Large Bag of Words Feature Spaces? In Advances in Intelligent Data Analysis VII, Michael R. Berthold, John Shawe-Taylor, and Nada Lavrač (Eds.). Vol. 4723. Springer, Berlin, Heidelberg, 296--307. Google ScholarDigital Library
Joachim Daiber, Max Jakob, Chris Hokamp, and Pablo N. Mendes. 2013. Improving Efficiency and Accuracy in Multilingual Entity Extraction. In Proceedings of the 9th I-SEMANTICS. ACM, New York, NY, USA, 121--124. Google ScholarDigital Library
Anastasia Dimou, Ruben Verborgh, Miel Vander Sande, Erik Mannens, and Rik Van de Walle. 2015. Machine-interpretable dataset and service descriptions for heterogeneous data access and retrieval. In Proceedings of the 11th International Conference on Semantic Systems SEMANTiCS2015. ACM, 145--152. Google ScholarDigital Library
Petra Drewer and Wolfgang Ziegler. 2011. Technische Dokumentation. Übersetzungsgerechte Texterstellung und Content-Management. Vogel, Würzburg.Google Scholar
Martin Dzbor, Enrico Motta, and John Domingue. 2004. Opening up magpie via semantic services. In International Semantic Web Conference. Springer, 635--649. Google ScholarDigital Library
Basil Ell and Andreas Harth. 2014. A language-independent method for the extraction of RDF verbalization templates. In Proceedings of the 8th International Natural Language Generation Conference (INLG). 26--34.Google ScholarCross Ref
European Association for Technical Communication - tekom e.V. 2017. iiRDS RDF Schema - First Public Working Draft, 03 April 2017. (2017).Google Scholar
European Association for Technical Communication - tekom e.V. 2017. iiRDS Specification - intelligent information Request and Delivery Standard - First Public Working Draft, 03 April 2017. (2017). https://iirds.tekom.deGoogle Scholar
Andreas Harth, Jürgen Umbrich, and Stefan Decker. 2006. Multicrawler: A pipelined architecture for crawling and indexing semantic web data. In International Semantic Web Conference, Vol. 4273. Springer, 258--271. Google ScholarDigital Library
Bernhard Haslhofer, Robert Sanderson, Rainer Simon, and Herbert van de Sompel. 2012. Open annotations on multimedia Web resources. Multimedia Tools and Applications (May 2012). Google ScholarDigital Library
Eero Hyvönen and Eetu Mäkelä. 2006. Semantic autocompletion. In Asian Semantic Web Conference. Springer, 739--751. Google ScholarDigital Library
IEC 82079-1. 2012. Preparation of Instructions for Use - Structuring, Content and Presentation. (2012).Google Scholar
Atanas Kiryakov, Borislav Popov, Ivan Terziev, Dimitar Manov, and Damyan Ognyanoff. 2004. Semantic annotation, indexing, and retrieval. Web Semantics: Science, Services and Agents on the World Wide Web 2, 1 (Dec. 2004), 49--79. Google ScholarDigital Library
Christopher D. Manning and Hinrich Schütze. 1999. Foundations of statistical natural language processing. MIT Press, Cambridge, Mass. Google ScholarDigital Library
Jan Oevermann. 2016. Reconstructing Semantic Structures in Technical Documentation with Vector Space Classification. In Posters and Demos Track of the 12th International Conference on Semantic Systems, Vol. 1695. CEUR-WS, Germany.Google Scholar
Jan Oevermann and Wolfgang Ziegler. 2016. Automated Intrinsic Text Classification for Component Content Management Applications in Technical Communication. In Proceedings of the 2016 ACM Symposium on Document Engineering. ACM Press, Vienna, Austria, 95--98. Google ScholarDigital Library
Laura Perez-Beltrachini, Rania Mohamed Sayed, and Claire Gardent. {n. d.}. Building RDF Content for Data-to-Text Generation. In Proceedings of COLING 2016, the 26th International Conference on Computational Linguistics: Technical Papers (2016). 1493--1502.Google Scholar
Eric Schweitzer and Jan C. Aurich. 2010. Continuous improvement of industrial product-service systems. CIRP Journal of Manufacturing Science and Technology 3, 2 (2010), 158--164.Google ScholarCross Ref
Marina Sokolova and Guy Lapalme. 2009. A systematic analysis of performance measures for classification tasks. Information Processing & Management 45, 4 (2009), 427--437. Google ScholarDigital Library
Axel J. Soto, Abidalrahman Mohammad, Andrew Albert, Aminul Islam, Evangelos Milios, Michael Doyle, Rosane Minghim, and Maria Cristina Ferreira de Oliveira. 2015. Similarity-Based Support for Text Reuse in Technical Writing. In Proceedings of the 2015 ACM Symposium on Document Engineering (DocEng '15). ACM, New York, NY, USA, 97--106. Google ScholarDigital Library
Steve Speicher, John Arwe, and Ashok Malhotra. 2009. Linked Data Platform 1.0. (2009). http://www.w3.org/TR/ldp/ 26 February 2015. W3C Recommendation.Google Scholar
Daniela Straub. 2016. Branchenkennzahlen für die Technische Dokumentation 2016 (Studie). tcworld GmbH, Stuttgart, Germany.Google Scholar
Victoria Uren, Philipp Cimiano, José Iria, Siegfried Handschuh, Maria Vargas-Vera, Enrico Motta, and Fabio Ciravegna. 2006. Semantic annotation for knowledge management: Requirements and a survey of the state of the art. Web Semantics: Science, Services and Agents on the World Wide Web 4, 1 (Jan. 2006), 14--28. Google ScholarDigital Library
W3C. 2017. Web Annotation Data Model - W3C Recommendation 23 February 2017. (2017). https://www.w3.org/TR/2017/REC-annotation-model-20170223/Google Scholar
W3C. 2017. Web Annotation Vocabulary - W3C Recommendation 23 February 2017. (2017). https://www.w3.org/TR/2017/REC-annotation-vocab-20170223/Google Scholar
Yutaka Yamauchi, Jack Whalen, and Daniel G. Bobrow. 2003. Information Use of Service Technicians in Difficult Cases. In Proceedings of the SIGCHI Conference on Human Factors in Computing Systems (CHI '03). ACM, 81--88. Google ScholarDigital Library
Wolfgang Ziegler and Heiko Beier. 2015. Content delivery portals: The future of modular content. tcworld e-magazine 02/2015 (2015).Google Scholar

Index Terms

Semantic Annotation of Heterogeneous Data Sources: Towards an Integrated Information Framework for Service Technicians
1. Information systems
  1. Information retrieval
    1. Document representation
      1. Content analysis and feature selection
  2. Information systems applications
    1. Enterprise information systems

Recommendations

A domain independent framework for extracting linked semantic data from tables
Search Computing

Vast amounts of information is encoded in tables found in documents, on the Web, and in spreadsheets or databases. Integrating or searching over this information benefits from understanding its intended meaning and making it explicit in a semantic ...
Read More
Learning the semantics of structured data sources

Information sources such as relational databases, spreadsheets, XML, JSON, and Web APIs contain a tremendous amount of structured data that can be leveraged to build and augment knowledge graphs. However, they rarely provide a semantic model to describe ...
Read More
Leveraging Linked Data to Discover Semantic Relations Within Data Sources
The Semantic Web – ISWC 2016
Abstract
Mapping data to a shared domain ontology is a key step in publishing semantic content on the Web. Most of the work on automatically mapping structured and semi-structured sources to ontologies focuses on semantic labeling, i.e., annotating data ...
Read More

Comments

Login options

Check if you have access through your login credentials or your institution to get full access on this article.

Full Access

Get this Publication

Published in
Semantics2017: Proceedings of the 13th International Conference on Semantic Systems
September 2017
202 pages
ISBN:9781450352963
DOI:10.1145/3132218
Editors:
Rinke Hoekstra
Elsevier B.V., Amsterdam, The Netherlands
,
Catherine Faron-Zucker
University of Nice Sophia Antipolis, France
,
Tassilo Pellegrini
University of Applied Sciences St. Poelten, Austria
,
Victor de Boer
Vrije Universiteit Amsterdam, The Netherlands
Copyright © 2017 ACM
Permission to make digital or hard copies of all or part of this work for personal or classroom use is granted without fee provided that copies are not made or distributed for profit or commercial advantage and that copies bear this notice and the full citation on the first page. Copyrights for components of this work owned by others than the author(s) must be honored. Abstracting with credit is permitted. To copy otherwise, or republish, to post on servers or to redistribute to lists, requires prior specific permission and/or a fee. Request permissions from [email protected].
Sponsors
In-Cooperation
Publisher
Association for Computing Machinery
New York, NY, United States
Publication History
- Published: 11 September 2017
Permissions
Request permissions about this article.
Request Permissions

Check for updates
Author Tags
Industrial Maintenance
Linked Data
Machine Learning
Metadata Generation
Technical Documentation
Qualifiers
- research-article
- Research
- Refereed limited
Conference

Acceptance Rates
Overall Acceptance Rate40of182submissions,22%
Funding Sources
Other Metrics
View Article Metrics

Article Metrics
- 3
  Total Citations
  View Citations
- 90
  Total Downloads
- Downloads (Last 12 months)5
- Downloads (Last 6 weeks)1
Other Metrics
View Author Metrics
Cited By
View all

PDF Format

View or Download as a PDF file.

PDF

eReader

View online with eReader.

eReader

Semantic Annotation of Heterogeneous Data Sources: Towards an Integrated Information Framework for Service Technicians

Semantics2017: Proceedings of the 13th International Conference on Semantic Systems

ABSTRACT

References

Cited By

Index Terms

Recommendations

A domain independent framework for extracting linked semantic data from tables

Learning the semantics of structured data sources

Leveraging Linked Data to Discover Semantic Relations Within Data Sources

Comments

Login options

Full Access

Published in

Sponsors

In-Cooperation

Publisher

Publication History

Permissions

Check for updates

Author Tags

Qualifiers

Conference

Acceptance Rates

Funding Sources

Other Metrics

Article Metrics

Other Metrics

Cited By

PDF Format

eReader

Digital Edition

Caption

Semantic Annotation of Heterogeneous Data Sources: Towards an Integrated Information Framework for Service Technicians

Semantics2017: Proceedings of the 13th International Conference on Semantic Systems

ABSTRACT

References

Cited By

Index Terms

Recommendations

A domain independent framework for extracting linked semantic data from tables

Learning the semantics of structured data sources

Leveraging Linked Data to Discover Semantic Relations Within Data Sources

Comments

Login options

Full Access

Published in

Sponsors

In-Cooperation

Publisher

Publication History

Permissions

Check for updates

Author Tags

Qualifiers

Conference

Acceptance Rates

Funding Sources

Article Metrics

Other Metrics

PDF Format

eReader

Digital Edition

Share this Publication link

Share on Social Media