Jochen Dörre

dblp:80/58 · DBLP profile ↗
← Back
10ranked-venue papers
7as first author
0since 2021 · last 1999
—ORCID · none

Domains — the database's venue-derived domains; a paper can count in several

Artificial intelligence and machine learning · 8 · 5 first-authorTheory of computation · 2 · 2 first-authorDatabases, data management, data science and information retrieval · 1 · 1 first-author

Expertise — from the expertise taxonomy: the topics of the expert's papers under the CCF categories. A weight counts papers with recency: 1 for a paper about the topic, 0.3 when the topic is its context, halved every five years.

Theoretical computer science
4 papers
Logic in computer science · 32% Automata and formal languages · 31% Computational complexity · 31%
Artificial intelligence
3 papers
Information extraction and text analysis · 78% Knowledge representation and reasoning · 22%
Software engineering, system software, and programming languages
1 paper
Compilers and program optimization · 100%

Topics — the 17 heaviest of 19, each with the papers that count most for it

TopicWeightPapersLastEvidence papers
Natural language and speech › Information extraction and text analysis
document analysis
0.011999
Text Mining: Finding Nuggets in Mountains of Textual Data · KDD 1999
Logic in computer science › semantics
underspecified semantics
0.011997
Efficient Construction of Underspecified Semantics under Massive Ambiguity · ACL 1997
Automata and formal languages
categorial grammar
0.011996
Parsing for Semidirectional Lambek Grammar is NP-Complete · ACL 1996
Automata and formal languages › categorial grammar
lambek grammar
0.011996
Parsing for Semidirectional Lambek Grammar is NP-Complete · ACL 1996
Computational complexity › language complexity
parsing complexity
0.011996
Parsing for Semidirectional Lambek Grammar is NP-Complete · ACL 1996
Natural language and speech › Information extraction and text analysis
syntactic parsing
0.011995
Memoization of Coroutined Constraints · ACL 1995
Compilers and program optimization
memoization
0.011995
Memoization of Coroutined Constraints · ACL 1995
Logic in computer science › philosophical logic › non-classical logic
feature logic
0.011991
Feature Logic with Weak Subsumption Constraints · ACL 1991
Automated reasoning and model checking
satisfiability
0.011991
Feature Logic with Weak Subsumption Constraints · ACL 1991
Automata and formal languages › grammar formalisms
constraint-based grammar
0.021997
Efficient Construction of Underspecified Semantics under Massive Ambiguity · ACL 1997
Feature Logic with Weak Subsumption Constraints · ACL 1991
Logic in computer science › unification
semi-unification
0.011990
On Subsumption and Semiunification in Feature Algebras · LICS 1990
Computational complexity
undecidability
0.011990
On Subsumption and Semiunification in Feature Algebras · LICS 1990
Logic in computer science
unification
0.011990
On Subsumption and Semiunification in Feature Algebras · LICS 1990
Knowledge, reasoning and agents › Knowledge representation and reasoning › representation language › knowledge representation formalisms › knowledge representation language
feature structures
0.011988
Unification of Disjunctive Feature Descriptions · ACL 1988
Knowledge, reasoning and agents › Knowledge representation and reasoning
unification
0.011988
Unification of Disjunctive Feature Descriptions · ACL 1988
Automata and formal languages › tree languages
regular trees
0.011990
On Subsumption and Semiunification in Feature Algebras · LICS 1990
Logic in computer science › knowledge representation and reasoning › description logic
subsumption
0.011990
On Subsumption and Semiunification in Feature Algebras · LICS 1990

Methods — techniques the papers use, named apart from their topics

logic programming · 0.0chart parsing · 0.0text analysis · 0.0packed representation · 0.0constraint-based semantic construction · 0.0reduction from 3-partition · 0.0subsumption · 0.0constraint satisfaction · 0.0kasper and rounds calculus · 0.0
YearPublicationVenuePosition
1999 Text Mining: Finding Nuggets in Mountains of Textual Data
abstract
Text mining applies the same analytical functions of data mining to the domain of textual information, relying on sophisticated text analysis techniques that distill information from free-text documents.IBM's Intelligent Miner for Text provides the necessary tools to unlock the business information that is "trapped" in email, insurance claims, news feeds, or other document repositories.It has been successfully applied in analyzing patent portfolios, customer complaint letters, and even competitors' Web pages.After defining our notion of "text mining", we focus on the differences between text and data mining and describe in some more detail the unique technologies that are key to successful text mining.
Jochen Dörre, Peter Gerstl, Roland Seiffert
KDD1
1997 Efficient Construction of Underspecified Semantics under Massive Ambiguity
abstract
We investigate the problem of determining a compact underspecified semantical representation for sentences that may be highly ambiguous. Due to combinatorial explosion, the naive method of building semantics for the different syntactic readings independently is prohibitive. We present a method that takes as input a syntactic parse forest with associated constraint-based semantic construction rules and directly builds a packed semantic structure. The algorithm is fully implemented and runs in O(n4log(n)) in sentence length, if the grammar meets some reasonable 'normality' restrictions.
Jochen Dörre
ACL1
1996 Parsing for Semidirectional Lambek Grammar is NP-Complete
abstract
We study the computational complexity of the parsing problem of a variant of Lambek Categorial Grammar that we call semidirectional. In semidirectional Lambek calculus SDL there is an additional nondirectional abstraction rule allowing the formula abstracted over to appear anywhere in the premise sequent's left-hand side, thus permitting non-peripheral extraction. SDL grammars are able to generate each context-free language and more than that. We show that the parsing problem for semidirectional Lambek Grammar is NP-complete by a reduction of the 3-Partition problem.
Jochen Dörre
ACL1
1995 Memoization of Coroutined Constraints
abstract
Some linguistic constraints cannot be effectively resolved during parsing at the location in which they are most naturally introduced.This paper shows how constraints can be propagated in a memoizing parser (such as a chart parser) in much the same way that variable bindings are, providing a general treatment of constraint coroutining in memoization.Prolog code for a simple application of our technique to Bouma and van Noord's (1994) categorial grammar analysis of Dutch is provided.
Mark Johnson 0001, Jochen Dörre
ACL2
1992 On Subsumption and Semiunifaction in Feature Algebras
Jochen Dörre, William C. Rounds
J. Symb. Comput.1
1991 Feature Logic with Weak Subsumption Constraints
abstract
In the general framework of a constraint-based grammar formalism often some sort of feature logic serves as the constraint language to describe linguistic objects.We investigate the extension of basic feature logic with subsumption (or matching) constraints, based on a weak notion of subsumption.This mechanism of oneway information flow is generally deemed to be necessary to give linguistically satisfactory descriptions of coordination phenomena in such formalisms.We show that the problem whether a set of constraints is satisfiable in this logic is decidable in polynomial time and give a solution algorithm.
Jochen Dörre
ACL1
1990 Feature Logic with Disjunctive Unification
Jochen Dörre, Andreas Eisele 0001
COLING1
1990 On Subsumption and Semiunification in Feature Algebras
abstract
A generalization of term subsumption, or matching, to a class of mathematical structures called feature algebras is discussed. It is shown how these generalize both first-order terms and the feature structures used in computational linguistics. The notion of subsumption generalizes to a natural notion of homomorphism between elements of these algebras, and the authors characterize the notion, showing how it corresponds to a mapping which preserves partial information. In the setting of feature algebras, unification corresponds naturally to solving constraints involving equalities between strings of unary functions symbols, and semiunification also allows inequalities representing subsumption constraints. Their generalization allows the authors to show that the semiunification problem for finite feature algebras is undecidable. This implies that the corresponding problem for rational trees (cyclic terms) is also undecidable. Thus a partial solution to the decidability of this problem, which has been open for several years, is produced.>
Jochen Dörre, William C. Rounds
LICS1
1988 Unification of Disjunctive Feature Descriptions
abstract
The paper describes a new implementation of feature structures containing disjunctive values, which can be characterized by the following main points: Local representation of embedded disjunctions, avoidance of expansion to disjunctive normal form and of repeated test-unifications for checking consistence. The method is based on a modification of Kasper and Rounds' calculus of feature descriptions and its correctness therefore is easy to see. It can handle cyclic structures and has been incorporated successfully into an environment for grammar development.
Andreas Eisele 0001, Jochen Dörre
ACL2
1986 A Lexical Functional Grammar System in Prolog
Andreas Eisele 0001, Jochen Dörre
COLING2