EDBT 2026 Demo / reviewers in the wild / expert
Wen-Hsiang Lu
dblp:99/3264
· DBLP profile ↗
27ranked-venue papers
8as first author
1since 2021 · last 2026
—ORCID · none
Domains — the database's venue-derived domains; a paper can count in several
Artificial intelligence and machine learning · 19 · 4 first-authorDatabases, data management, data science and information retrieval · 10 · 3 first-authorGraphics, computer vision, multimedia, augmented reality and games · 3Applied, interdisciplinary, general and emerging computing · 3 · 2 first-author · 1 since 2021
Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.
| Databases, data mining, and information retrieval
5 papers |
Information retrieval · 93% Web and social media mining · 7% | |
| Artificial intelligence
1 paper |
Machine translation · 50% Information extraction and text analysis · 50% |
Topics — the 9 heaviest of 14, each with the papers that count most for it
| Topic | Weight | Papers | Last | Evidence papers |
|---|---|---|---|---|
Information retrieval
cross-language information retrieval |
0.2 | 4 | 2004 | Anchor text mining for translation of Web queries: A transitive translation approach · ACM Trans. Inf. Syst. 2004 Translating unknown queries with web corpora for cross-language information retrieval · SIGIR 2004 Anchor Text Mining for Translation Extraction of Query Terms · SIGIR 2001 |
Information retrieval
multilingual dictionary construction |
0.0 | 1 | 2004 | Anchor text mining for translation of Web queries: A transitive translation approach · ACM Trans. Inf. Syst. 2004 |
Information retrieval › cross-language information retrieval
query translation |
0.0 | 1 | 2004 | Anchor text mining for translation of Web queries: A transitive translation approach · ACM Trans. Inf. Syst. 2004 |
Information retrieval
web search |
0.0 | 2 | 2004 | Anchor text mining for translation of Web queries: A transitive translation approach · ACM Trans. Inf. Syst. 2004 Translating unknown queries with web corpora for cross-language information retrieval · SIGIR 2004 |
Information retrieval
interactive information retrieval |
0.0 | 1 | 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web search · SIGIR 2000 |
Information retrieval
query log analysis |
0.0 | 1 | 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web search · SIGIR 2000 |
Information retrieval
query processing |
0.0 | 1 | 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web search · SIGIR 2000 |
Information retrieval › text analysis
thesaurus construction |
0.0 | 1 | 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web search · SIGIR 2000 |
Information retrieval
query understanding |
0.0 | 1 | 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web search · SIGIR 2000 |
Methods — techniques the papers use, named apart from their topics
anchor text mining · 0.1transitive translation · 0.1geographic information mining · 0.1bilingual search-result page mining · 0.1web corpus mining · 0.0transitive translation model · 0.0competitive linking · 0.0bilingual lexicon extraction · 0.0link structure mining · 0.0search term log mining · 0.0
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | A Novel Relative Distance Protein Fingerprint Algorithm for Searching DNA Mimic ProteinsabstractDNA mimic proteins are relatively obscure control factors that resemble DNA by mimicking its negatively charged distribution. They achieve this using negatively charged amino acids like aspartic acid (ASP/D) and glutamic acid (GLU/E). Known DNA mimic proteins control various cellular mechanisms, such as transcription, DNA repair, and gene regulation, by intervening in the binding of DNA to effector proteins. In addition to their biological functions, DNA mimic proteins may also be applicable in biotechnology, for example, by regulating CRISPR-Cas9 activity to enhance gene editing precision. Therefore, DNA mimic proteins warrant further research. However, most DNA mimic proteins cannot be identified using traditional bioinformatics methods owing to their unique amino acid sequences and structural features. We developed a new protein fingerprint, called relative distance protein fingerprint (RD-PFP), that can be used to analyze the distribution of amino acids on a protein surface. We optimized our RD-PFP by using machine learning and the characteristic feature of DNA mimic proteins (namely, their DNA-like negatively charged distribution) to more accurately predict DNA mimicry from protein structures. Our pioneering study contributes to the development of machine learning-based bioinformatics methods for screening DNA mimic proteins. Chia-Yen Chien, Hsin-Hung Chou, Kai-Cheng Hsu, Bo-Cheng Liao, Hao-Ching Wang, Wen-Hsiang Lu, Sun-Yuan Hsieh |
IEEE Trans. Comput. Biol. Bioinform. | 6 |
| 2016 | Analyzing depression tendency of web posts using an event-driven depression tendency warning model
Chia-Ming Tung, Wen-Hsiang Lu |
Artif. Intell. Medicine | 2 |
| 2016 | Constructing Complex Search Tasks with Coherent Subtask Search GoalsabstractNowadays, due to the explosive growth of web content and usage, users deal with their complex search tasks by web search engines. However, conventional search engines consider a search query corresponding only to a simple search task. In order to accomplish a complex search task, which consists of multiple subtask search goals, users usually have to issue a series of queries. For example, the complex search task “travel to Dubai” may involve several subtask search goals, including reserving hotel room, surveying Dubai landmarks, booking flights, and so forth. Therefore, a user can efficiently accomplish his or her complex search task if search engines can predict the complex search task with a variety of subtask search goals. In this work, we propose a complex search task model (CSTM) to deal with this problem. The CSTM first groups queries into complex search task clusters, and then generates subtask search goals from each complex search task cluster. To raise the performance of CSTM, we exploit four web resources including community question answering, query logs, search engine result pages, and clicked pages. Experimental results show that our CSTM is effective in identifying the comprehensive subtask search goals of a complex search task. Ting-Xuan Wang, Wen-Hsiang Lu |
ACM Trans. Asian Low Resour. Lang. Inf. Process. | 2 |
| 2011 | An effective tree structure for mining high utility itemsets
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
Expert Syst. Appl. | 3 |
| 2010 | Efficiently Mining High Average Utility Itemsets with a Tree Structure
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
ACIIDS (1) | 3 |
| 2010 | A two-phase fuzzy mining approachabstractIn this paper, we propose a two-phase fuzzy mining approach based on a tree structure to discover fuzzy frequent itemsets from a quantitative database. A simple tree structure called the upper-bound fuzzy frequent-pattern tree (abbreviated as UBFFP tree) is designed to help achieve the purpose. The two-phase fuzzy mining approach can easily derive the upper-bound fuzzy supports of itemsets through the tree and prune unpromising itemsets in the first phase, and then finds the actual frequent fuzzy itemsets in the second phase. Experimental results also show the good performance of the proposed approach. Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
FUZZ-IEEE | 3 |
| 2010 | Linguistic data mining with fuzzy FP-trees
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
Expert Syst. Appl. | 3 |
| 2009 | Fuzzy data mining based on the compressed fuzzy FP-treesabstractIn this paper, we design the compressed fuzzy FP-tree structure to mine fuzzy frequent itemsets from the transactions with quantitative values. It consists of two phases. In the first phase, a compressed fuzzy frequent pattern tree (CFFP tree) is constructed from the given quantitative transactions. In the second phase, a fuzzy mining approach is then proposed to mine the fuzzy frequent itemsets from the CFFP tree constructed. Experiments are also made to show the performance of the proposed approach. Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
FUZZ-IEEE | 3 |
| 2009 | The Pre-FUFP algorithm for incremental mining
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
Expert Syst. Appl. | 3 |
| 2008 | Incremental Mining with Prelarge Trees
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu, Been-Chian Chien |
IEA/AIE | 3 |
| 2008 | Using Web resources to construct multilingual medical thesaurus for cross-language medical information retrieval
Wen-Hsiang Lu, Ray Shih-Jui Lin, Yi-Che Chan, Kuan-Hsi Chen |
Decis. Support Syst. | 1 |
| 2007 | Using the Pre-FUFP Algorithm for Handling New Transactions in Incremental MiningabstractIn the past, we proposed a Fast Updated FP-tree (FUFP-tree) structure to efficiently handle new transactions and to make the tree update process become easier. In this paper, we attempt to modify the FUFP-tree construction based on the concept of pre-large itemsets. Pre-large itemsets are defined by a lower support threshold and an upper support threshold. The proposed approach can achieve a good execution time for tree construction especially when each time a small number of transactions are inserted. Experimental results also show that the proposed Pre-FUFP maintenance algorithm has a good performance for incrementally handling new transactions. Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu |
CIDM | 3 |
| 2007 | Maintenance of Fast Updated Frequent Trees for Record Deletion Based on Prelarge Concepts
Jerry Chun-Wei Lin, Tzung-Pei Hong, Wen-Hsiang Lu, Chih-Hung Wu |
IEA/AIE | 3 |
| 2007 | Improving Identification of Latent User Goals through Search-Result Snippet ClassificationabstractIn this paper, we propose an enhanced approach to improving our previous method which employs syntactic structures (verb-object pairs) to identify latent user goals. Our new approach employs a supervised-learning method to learn hint verbs and considers URL information and title information to classify snippets into three coarse categories, which are resource-seeking, informational, and navigational. Also, we propose three different models to identify three different categories of specific latent user goals from the classified snippets. Kuan-Yu He, Yao-Sheng Chang, Wen-Hsiang Lu |
Web Intelligence | 3 |
| 2006 | Overcoming Terminology Barrier Using Web Resources for Cross-Language Medical Information Retrieval
Wen-Hsiang Lu, Ray Shih-Jui Lin, Yi-Che Chan, Kuan-Hsi Chen |
AMIA | 1 |
| 2006 | Identifying User Goals from Web Search ResultsabstractWith the fast growth of the Web, users often suffer from the problem of information overload since many existing search engines response lots of non-relevant documents containing query terms based on the search mechanism of keyword matching. In fact, it is eagerly expected by both users and search engine developers to reduce overloaded information by understanding user goals clearly. In this paper, we intend to utilize Web search results to identify user goals. We propose one novel probabilistic inference model which effectively employs syntactic features to discover a variety of confined user goals Yao-Sheng Chang, Kuan-Yu He, Scott Yu, Wen-Hsiang Lu |
Web Intelligence | 4 |
| 2006 | Exploiting the Web as the multilingual corpus for unknown query translationabstractAbstract Users' cross‐lingual queries to a digital library system might be short and the query terms may not be included in a common translation dictionary (unknown terms). In this article, the authors investigate the feasibility of exploiting the Web as the multilingual corpus source to translate unknown query terms for cross‐language information retrieval in digital libraries. They propose a Web‐based term translation approach to determine effective translations for unknown query terms by mining bilingual search‐result pages obtained from a real Web search engine. This approach can enhance the construction of a domain‐specific bilingual lexicon and bring multilingual support to a digital library that only has monolingual document collections. Very promising results have been obtained in generating effective translation equivalents for many unknown terms, including proper nouns, technical terms, and Web query terms, and in assisting bilingual lexicon construction for a real digital library system. Jenq-Haur Wang, Jei-Wen Teng, Wen-Hsiang Lu, Lee-Feng Chien |
J. Assoc. Inf. Sci. Technol. | 3 |
| 2005 | Semi-Automatic Construction of the Chinese-English MeSH Using Web-BasedTerm Translation Method
Wen-Hsiang Lu, Ray Shih-Jui Lin, Yi-Che Chan, Kuan-Hsi Chen |
AMIA | 1 |
| 2004 | Creating Multilingual Translation Lexicons with Regional Variations Using Web CorporaabstractThe purpose of this paper is to automatically create multilingual translation lexicons with regional variations. We propose a transitive translation approach to determine translation variations across languages that have insufficient corpora for translation via the mining of bilingual search-result pages and clues of geographic information obtained from Web search engines. The experimental results have shown the feasibility of the proposed approach in efficiently generating translation equivalents of various terms not covered by general translation dictionaries. It also revealed that the created translation lexicons can reflect different cultural aspects across regions such as Taiwan, Hong Kong and mainland China. Pu-Jen Cheng, Wen-Hsiang Lu, Jei-Wen Teng, Lee-Feng Chien |
ACL | 2 |
| 2004 | Translating unknown queries with web corpora for cross-language information retrievalabstractIt is crucial for cross-language information retrieval (CLIR) systems to deal with the translation of unknown queries due to that real queries might be short. The purpose of this paper is to investigate the feasibility of exploiting the Web as the corpus source to translate unknown queries for CLIR. We propose an online translation approach to determine effective translations for unknown query terms via mining of bilingual search-result pages obtained from Web search engines. This approach can alleviate the problem of the lack of large bilingual corpora, translate many unknown query terms, provide flexible query specifications, and extract semantically-close translations to benefit CLIR tasks -- especially for cross-language Web search. Pu-Jen Cheng, Jei-Wen Teng, Ruey-Cheng Chen, Jenq-Haur Wang, Wen-Hsiang Lu, Lee-Feng Chien |
SIGIR | 5 |
| 2004 | Anchor text mining for translation of Web queries: A transitive translation approachabstractTo discover translation knowledge in diverse data resources on the Web, this article proposes an effective approach to finding translation equivalents of query terms and constructing multilingual lexicons through the mining of Web anchor texts and link structures. Although Web anchor texts are wide-scoped hypertext resources, not every particular pair of languages contains sufficient anchor texts for effective extraction of translations for Web queries. For more generalized applications, the approach is designed based on a transitive translation model. The translation equivalents of a query term can be extracted via its translation in an intermediate language. To reduce interference from translation errors, the approach further integrates a competitive linking algorithm into the process of determining the most probable translation. A series of experiments has been conducted, including performance tests on term translation extraction, cross-language information retrieval, and translation suggestions for practical Web search services, respectively. The obtained experimental results have shown that the proposed approach is effective in extracting translations of unknown queries, is easy to combine with the probabilistic retrieval model to improve the cross-language retrieval performance, and is very useful when the considered language pairs lack a sufficient number of anchor texts. Based on the approach, an experimental system called LiveTrans has been developed for English--Chinese cross-language Web search. Wen-Hsiang Lu, Lee-Feng Chien, Hsi-Jian Lee |
ACM Trans. Inf. Syst. | 1 |
| 2002 | A Transitive Model for Extracting Translation Equivalents of Web Queries through Anchor Text Mining
Wen-Hsiang Lu, Lee-Feng Chien, Hsi-Jian Lee |
COLING | 1 |
| 2002 | Translation of web queries using anchor text miningabstractThis article presents an approach to automatically extracting translations of Web query terms through mining of Web anchor texts and link structures. One of the existing difficulties in cross-language information retrieval (CLIR) and Web search is the lack of appropriate translations of new terminology and proper names. The proposed approach successfully exploits the anchor-text resources and reduces the existing difficulties of query term translation. Many query terms that cannot be obtained in general-purpose translation dictionaries are, therefore, extracted. Wen-Hsiang Lu, Lee-Feng Chien, Hsi-Jian Lee |
ACM Trans. Asian Lang. Inf. Process. | 1 |
| 2001 | Anchor Text Mining for Translation of Web QueriesabstractThe paper presents an approach to automatically extracting translations of Web query terms through mining of Web anchor texts and link structures. One of the existing difficulties in cross-language information retrieval (CLIR) and Web search is the lack of the appropriate translations of new terminology and proper names. Such a difficult problem can be effectively alleviated by our proposed approach, and the resource of anchor texts in the Web is proven a valuable corpus for this kind of term translation. Wen-Hsiang Lu, Lee-Feng Chien, Hsi-Jian Lee |
ICDM | 1 |
| 2001 | Anchor Text Mining for Translation Extraction of Query TermsabstractThis paper presents an approach to automatically extracting the bilingual translations of many Web query terms through mining the Web anchor texts. Some preliminary experiments are conducted on using 109,416 Web pages containing both Chinese and English anchor texts in their in-links to extract Chinese translations of 200 English queries selected from popular query terms in Taiwan. It is found that the effective translations of 75% of the popular query terms can be extracted, in which 87.2% cannot be obtained in common translation dictionaries. Wen-Hsiang Lu, Hsi-Jian Lee, Lee-Feng Chien |
SIGIR | 1 |
| 2000 | Live thesaurus construction for interactive voice-based web searchabstractSince Web users ’ queries are often too short, an accurate and interactive speech interface is believed very helpful, especially for WAP-based Web search engines. To provide high accurate speech recognition and effective interactive search, a rigid and live web thesaurus that contains users ' search terms plus a set of relations between their associated terms is highly in demand. The purpose of this paper is intended to present a log-based approach for live thesaurus construction. Based on the live thesaurus, certain kinds of users ' information behaviors could be characterized and a more effective voice-based search engine could be developed. Shui-Lung Chuang, Hsiao-Tieh Pu, Wen-Hsiang Lu, Lee-Feng Chien |
INTERSPEECH | 3 |
| 2000 | Auto-construction of a live thesaurus from search term logs for interactive Web searchabstractThe purpose of this paper is to present an on-going research that is intended to construct a live thesaurus directly from search term logs of real-world search engines. Such a thesaurus designed can contain representative search terms, their frequency in use, the corresponding subject categories, the associated and relevant terms, and the hot visiting Web sites/pages the search terms may reach. Shui-Lung Chuang, Hsiao-Tieh Pu, Wen-Hsiang Lu, Lee-Feng Chien |
SIGIR | 3 |