Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design

Bioinspired design is the adaptation of methods, strategies, or principles found in nature to solve engineering problems. One formalized approach to bioinspired solution seeking is the abstraction of the engineering problem into a functional need and then seeking solutions to this function using a k...

Ausführliche Beschreibung

Gespeichert in:
Bibliographische Detailangaben
Veröffentlicht in:Journal of mechanical design (1990) 2014-11, Vol.136 (11)
Hauptverfasser: Glier, Michael W, McAdams, Daniel A, Linsey, Julie S
Format: Artikel
Sprache:eng
Schlagworte:
Online-Zugang:Volltext
Tags: Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
container_end_page
container_issue 11
container_start_page
container_title Journal of mechanical design (1990)
container_volume 136
creator Glier, Michael W
McAdams, Daniel A
Linsey, Julie S
description Bioinspired design is the adaptation of methods, strategies, or principles found in nature to solve engineering problems. One formalized approach to bioinspired solution seeking is the abstraction of the engineering problem into a functional need and then seeking solutions to this function using a keyword type search method on text based biological knowledge. These function keyword search approaches have shown potential for success, but as with many text based search methods, they produce a large number of results, many of little relevance to the problem in question. In this paper, we develop a method to train a computer to identify text passages more likely to suggest a solution to a human designer. The work presented examines the possibility of filtering biological keyword search results by using text mining algorithms to automatically identify which results are likely to be useful to a designer. The text mining algorithms are trained on a pair of surveys administered to human subjects to empirically identify a large number of sentences that are, or are not, helpful for idea generation. We develop and evaluate three text classification algorithms, namely, a Naïve Bayes (NB) classifier, a k nearest neighbors (kNN) classifier, and a support vector machine (SVM) classifier. Of these methods, the NB classifier generally had the best performance. Based on the analysis of 60 word stems, a NB classifier's precision is 0.87, recall is 0.52, and F score is 0.65. We find that word stem features that describe a physical action or process are correlated with helpful sentences. Similarly, we find biological jargon feature words are correlated with unhelpful sentences.
doi_str_mv 10.1115/1.4028167
format Article
fullrecord <record><control><sourceid>proquest_cross</sourceid><recordid>TN_cdi_proquest_miscellaneous_1651423821</recordid><sourceformat>XML</sourceformat><sourcesystem>PC</sourcesystem><sourcerecordid>1651423821</sourcerecordid><originalsourceid>FETCH-LOGICAL-a282t-ecfda6a83ff364597f0229534f5ad894383abe11ef9e9fb5f7be31522803a0ef3</originalsourceid><addsrcrecordid>eNotkDFPwzAUhCMEEqUwMLN4hCHFz44TZyylQEUlJCiz5SbPxVUSBzuB9t8T1E73dPp0endRdA10AgDiHiYJZRLS7CQagWAyzimF0-GmgsY0ydh5dBHCdjBBJmIUbea7tnLeNhsy7TtX6w5LssJdR2aVDsEaW-jOuoZ0jizq1rsfJK-4_3W-JDPn2z6QD9S--CLvGPqqC8Q4Tx6ss01orR_CHjHYTXMZnRldBbw66jj6fJqvZi_x8u15MZsuY80k62IsTKlTLbkxPE1EnhnKWC54YoQuZZ5wyfUaAdDkmJu1MNka-dCTSco1RcPH0e0hd3j1u8fQqdqGAqtKN-j6oCAVkDAuGQzo3QEtvAvBo1Gtt7X2ewVU_Y-pQB3HHNibA6tDjWrret8MLRTPBOSM_wFI8nDE</addsrcrecordid><sourcetype>Aggregation Database</sourcetype><iscdi>true</iscdi><recordtype>article</recordtype><pqid>1651423821</pqid></control><display><type>article</type><title>Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design</title><source>Alma/SFX Local Collection</source><source>ASME Transactions Journals (Current)</source><creator>Glier, Michael W ; McAdams, Daniel A ; Linsey, Julie S</creator><creatorcontrib>Glier, Michael W ; McAdams, Daniel A ; Linsey, Julie S</creatorcontrib><description>Bioinspired design is the adaptation of methods, strategies, or principles found in nature to solve engineering problems. One formalized approach to bioinspired solution seeking is the abstraction of the engineering problem into a functional need and then seeking solutions to this function using a keyword type search method on text based biological knowledge. These function keyword search approaches have shown potential for success, but as with many text based search methods, they produce a large number of results, many of little relevance to the problem in question. In this paper, we develop a method to train a computer to identify text passages more likely to suggest a solution to a human designer. The work presented examines the possibility of filtering biological keyword search results by using text mining algorithms to automatically identify which results are likely to be useful to a designer. The text mining algorithms are trained on a pair of surveys administered to human subjects to empirically identify a large number of sentences that are, or are not, helpful for idea generation. We develop and evaluate three text classification algorithms, namely, a Naïve Bayes (NB) classifier, a k nearest neighbors (kNN) classifier, and a support vector machine (SVM) classifier. Of these methods, the NB classifier generally had the best performance. Based on the analysis of 60 word stems, a NB classifier's precision is 0.87, recall is 0.52, and F score is 0.65. We find that word stem features that describe a physical action or process are correlated with helpful sentences. Similarly, we find biological jargon feature words are correlated with unhelpful sentences.</description><identifier>ISSN: 1050-0472</identifier><identifier>EISSN: 1528-9001</identifier><identifier>DOI: 10.1115/1.4028167</identifier><language>eng</language><publisher>ASME</publisher><subject>Algorithms ; Biological ; Classifiers ; Design engineering ; Design Theory and Methodology ; Mathematical models ; Searching ; Sentences ; Texts</subject><ispartof>Journal of mechanical design (1990), 2014-11, Vol.136 (11)</ispartof><lds50>peer_reviewed</lds50><woscitedreferencessubscribed>false</woscitedreferencessubscribed><citedby>FETCH-LOGICAL-a282t-ecfda6a83ff364597f0229534f5ad894383abe11ef9e9fb5f7be31522803a0ef3</citedby><cites>FETCH-LOGICAL-a282t-ecfda6a83ff364597f0229534f5ad894383abe11ef9e9fb5f7be31522803a0ef3</cites></display><links><openurl>$$Topenurl_article</openurl><openurlfulltext>$$Topenurlfull_article</openurlfulltext><thumbnail>$$Tsyndetics_thumb_exl</thumbnail><link.rule.ids>314,776,780,27901,27902,38497</link.rule.ids></links><search><creatorcontrib>Glier, Michael W</creatorcontrib><creatorcontrib>McAdams, Daniel A</creatorcontrib><creatorcontrib>Linsey, Julie S</creatorcontrib><title>Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design</title><title>Journal of mechanical design (1990)</title><addtitle>J. Mech. Des</addtitle><description>Bioinspired design is the adaptation of methods, strategies, or principles found in nature to solve engineering problems. One formalized approach to bioinspired solution seeking is the abstraction of the engineering problem into a functional need and then seeking solutions to this function using a keyword type search method on text based biological knowledge. These function keyword search approaches have shown potential for success, but as with many text based search methods, they produce a large number of results, many of little relevance to the problem in question. In this paper, we develop a method to train a computer to identify text passages more likely to suggest a solution to a human designer. The work presented examines the possibility of filtering biological keyword search results by using text mining algorithms to automatically identify which results are likely to be useful to a designer. The text mining algorithms are trained on a pair of surveys administered to human subjects to empirically identify a large number of sentences that are, or are not, helpful for idea generation. We develop and evaluate three text classification algorithms, namely, a Naïve Bayes (NB) classifier, a k nearest neighbors (kNN) classifier, and a support vector machine (SVM) classifier. Of these methods, the NB classifier generally had the best performance. Based on the analysis of 60 word stems, a NB classifier's precision is 0.87, recall is 0.52, and F score is 0.65. We find that word stem features that describe a physical action or process are correlated with helpful sentences. Similarly, we find biological jargon feature words are correlated with unhelpful sentences.</description><subject>Algorithms</subject><subject>Biological</subject><subject>Classifiers</subject><subject>Design engineering</subject><subject>Design Theory and Methodology</subject><subject>Mathematical models</subject><subject>Searching</subject><subject>Sentences</subject><subject>Texts</subject><issn>1050-0472</issn><issn>1528-9001</issn><fulltext>true</fulltext><rsrctype>article</rsrctype><creationdate>2014</creationdate><recordtype>article</recordtype><recordid>eNotkDFPwzAUhCMEEqUwMLN4hCHFz44TZyylQEUlJCiz5SbPxVUSBzuB9t8T1E73dPp0endRdA10AgDiHiYJZRLS7CQagWAyzimF0-GmgsY0ydh5dBHCdjBBJmIUbea7tnLeNhsy7TtX6w5LssJdR2aVDsEaW-jOuoZ0jizq1rsfJK-4_3W-JDPn2z6QD9S--CLvGPqqC8Q4Tx6ss01orR_CHjHYTXMZnRldBbw66jj6fJqvZi_x8u15MZsuY80k62IsTKlTLbkxPE1EnhnKWC54YoQuZZ5wyfUaAdDkmJu1MNka-dCTSco1RcPH0e0hd3j1u8fQqdqGAqtKN-j6oCAVkDAuGQzo3QEtvAvBo1Gtt7X2ewVU_Y-pQB3HHNibA6tDjWrret8MLRTPBOSM_wFI8nDE</recordid><startdate>20141101</startdate><enddate>20141101</enddate><creator>Glier, Michael W</creator><creator>McAdams, Daniel A</creator><creator>Linsey, Julie S</creator><general>ASME</general><scope>AAYXX</scope><scope>CITATION</scope><scope>7TB</scope><scope>8FD</scope><scope>F28</scope><scope>FR3</scope></search><sort><creationdate>20141101</creationdate><title>Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design</title><author>Glier, Michael W ; McAdams, Daniel A ; Linsey, Julie S</author></sort><facets><frbrtype>5</frbrtype><frbrgroupid>cdi_FETCH-LOGICAL-a282t-ecfda6a83ff364597f0229534f5ad894383abe11ef9e9fb5f7be31522803a0ef3</frbrgroupid><rsrctype>articles</rsrctype><prefilter>articles</prefilter><language>eng</language><creationdate>2014</creationdate><topic>Algorithms</topic><topic>Biological</topic><topic>Classifiers</topic><topic>Design engineering</topic><topic>Design Theory and Methodology</topic><topic>Mathematical models</topic><topic>Searching</topic><topic>Sentences</topic><topic>Texts</topic><toplevel>peer_reviewed</toplevel><toplevel>online_resources</toplevel><creatorcontrib>Glier, Michael W</creatorcontrib><creatorcontrib>McAdams, Daniel A</creatorcontrib><creatorcontrib>Linsey, Julie S</creatorcontrib><collection>CrossRef</collection><collection>Mechanical &amp; Transportation Engineering Abstracts</collection><collection>Technology Research Database</collection><collection>ANTE: Abstracts in New Technology &amp; Engineering</collection><collection>Engineering Research Database</collection><jtitle>Journal of mechanical design (1990)</jtitle></facets><delivery><delcategory>Remote Search Resource</delcategory><fulltext>fulltext</fulltext></delivery><addata><au>Glier, Michael W</au><au>McAdams, Daniel A</au><au>Linsey, Julie S</au><format>journal</format><genre>article</genre><ristype>JOUR</ristype><atitle>Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design</atitle><jtitle>Journal of mechanical design (1990)</jtitle><stitle>J. Mech. Des</stitle><date>2014-11-01</date><risdate>2014</risdate><volume>136</volume><issue>11</issue><issn>1050-0472</issn><eissn>1528-9001</eissn><abstract>Bioinspired design is the adaptation of methods, strategies, or principles found in nature to solve engineering problems. One formalized approach to bioinspired solution seeking is the abstraction of the engineering problem into a functional need and then seeking solutions to this function using a keyword type search method on text based biological knowledge. These function keyword search approaches have shown potential for success, but as with many text based search methods, they produce a large number of results, many of little relevance to the problem in question. In this paper, we develop a method to train a computer to identify text passages more likely to suggest a solution to a human designer. The work presented examines the possibility of filtering biological keyword search results by using text mining algorithms to automatically identify which results are likely to be useful to a designer. The text mining algorithms are trained on a pair of surveys administered to human subjects to empirically identify a large number of sentences that are, or are not, helpful for idea generation. We develop and evaluate three text classification algorithms, namely, a Naïve Bayes (NB) classifier, a k nearest neighbors (kNN) classifier, and a support vector machine (SVM) classifier. Of these methods, the NB classifier generally had the best performance. Based on the analysis of 60 word stems, a NB classifier's precision is 0.87, recall is 0.52, and F score is 0.65. We find that word stem features that describe a physical action or process are correlated with helpful sentences. Similarly, we find biological jargon feature words are correlated with unhelpful sentences.</abstract><pub>ASME</pub><doi>10.1115/1.4028167</doi></addata></record>
fulltext fulltext
identifier ISSN: 1050-0472
ispartof Journal of mechanical design (1990), 2014-11, Vol.136 (11)
issn 1050-0472
1528-9001
language eng
recordid cdi_proquest_miscellaneous_1651423821
source Alma/SFX Local Collection; ASME Transactions Journals (Current)
subjects Algorithms
Biological
Classifiers
Design engineering
Design Theory and Methodology
Mathematical models
Searching
Sentences
Texts
title Exploring Automated Text Classification to Improve Keyword Corpus Search Results for Bioinspired Design
url https://sfx.bib-bvb.de/sfx_tum?ctx_ver=Z39.88-2004&ctx_enc=info:ofi/enc:UTF-8&ctx_tim=2025-02-04T06%3A05%3A57IST&url_ver=Z39.88-2004&url_ctx_fmt=infofi/fmt:kev:mtx:ctx&rfr_id=info:sid/primo.exlibrisgroup.com:primo3-Article-proquest_cross&rft_val_fmt=info:ofi/fmt:kev:mtx:journal&rft.genre=article&rft.atitle=Exploring%20Automated%20Text%20Classification%20to%20Improve%20Keyword%20Corpus%20Search%20Results%20for%20Bioinspired%20Design&rft.jtitle=Journal%20of%20mechanical%20design%20(1990)&rft.au=Glier,%20Michael%20W&rft.date=2014-11-01&rft.volume=136&rft.issue=11&rft.issn=1050-0472&rft.eissn=1528-9001&rft_id=info:doi/10.1115/1.4028167&rft_dat=%3Cproquest_cross%3E1651423821%3C/proquest_cross%3E%3Curl%3E%3C/url%3E&disable_directlink=true&sfx.directlink=off&sfx.report_link=0&rft_id=info:oai/&rft_pqid=1651423821&rft_id=info:pmid/&rfr_iscdi=true