VLDB 2026 Research / reviewers in the wild / expert
Claudio S. Pinhanez
dblp:53/154 · also Claudio Santos Pinhanez
· DBLP profile ↗
41ranked-venue papers
14as first author
11since 2021 · last 2026
0000-0001-6715-1290ORCID · verified
Domains — the database's venue-derived domains; a paper can count in several
Human-computer interaction and ubiquitous computing · 20 · 7 first-author · 4 since 2021Artificial intelligence and machine learning · 14 · 5 first-author · 7 since 2021Graphics, computer vision, multimedia, augmented reality and games · 10 · 5 first-author · 3 since 2021Applied, interdisciplinary, general and emerging computing · 2 · 1 since 2021Databases, data management, data science and information retrieval · 1 · 1 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | Small Models Exhibit Limited Answer Consistency in Repetition Trials of the Multiple-Choice MMLU-Redux and MedQA BenchmarksabstractThis work explores the consistency of LLMs in answering multiple times the same question. In particular, we study how known, open-source LLMs respond to 10 repetitions of questions from the multiple-choice benchmarks MMLU-Redux and MedQA, considering different inference temperatures, small (2B-10B parameters) vs. medium models (50B-80B), finetuned vs. base models, and other parameters. The paper also examines the effects of requiring answer consistency in repetitive inferences on accuracy and the trade-offs involved in deciding which model best provides both of them, for what we propose some new representations. Results show that the number of questions which can be answered consistently vary wildly among models but typically is in the 50%-85% range for small models and that accuracy among consistent answers correlates to overall accuracy at low inference temperatures. Results for medium-sized models seem to indicate much higher levels of answer consistency. Claudio S. Pinhanez, Paulo Rodrigo Cavalin, Cassia Sampaio Sanctos, Marcelo Grave |
AAAI | 1 |
| 2025 | Sentence-level Aggregation of Lexical Metrics Correlates Stronger with Human Judgements than Corpus-level AggregationabstractIn this paper we show that corpus-level aggregation hinders considerably the capability of lexical metrics to accurately evaluate machine translation (MT) systems. With empirical experiments we demonstrate that averaging individual segment-level scores can make metrics such as BLEU and chrF correlate much stronger with human judgements and make them behave considerably more similar to neural metrics such as COMET and BLEURT. We show that this difference exists because corpus- and segment-level aggregation differs considerably owing to the classical average of ratio versus ratio of averages Mathematical problem. Moreover, as we also show, such difference affects considerably the statistical robustness of corpus-level aggregation. Considering that neural metrics currently only cover a small set of sufficiently-resourced languages, the results in this paper can help make the evaluation of MT systems for low-resource languages more trustworthy. Paulo Rodrigo Cavalin, Pedro Henrique Domingues, Claudio S. Pinhanez |
AAAI | 3 |
| 2024 | Theoretical and Empirical Advantages of Dense-Vector to One-Hot Encoding of Intent Classes in Open-World ScenariosabstractThis work explores the intrinsic limitations of the popular one-hot encoding method in classification of intents when detection of out-of-scope (OOS) inputs is required. Although recent work has shown that there can be significant improvements in OOS detection when the intent classes are represented as dense-vectors based on domain-specific knowledge, we argue in this paper that such gains are more likely due to advantages of the much richer topologies that can be created with dense vectors compared to the equidistant class representation assumed by one-hot encodings. We start by demonstrating how dense-vector encodings are able to create OOS spaces with much richer topologies. Then, we show empirically, using four standard intent classification datasets, that knowledge-free, randomly generated dense-vector encodings of intent classes can yield over 20% gains over one-hot encodings, producing better systems for open-world classification tasks, mostly from improvements in OOS detection. Paulo Rodrigo Cavalin, Claudio S. Pinhanez |
LREC/COLING | 2 |
| 2024 | Disappearing without a Trace: Coverage, Community, Quality, and Temporal Dynamics of Wikipedia Articles on Endangered Brazilian Indigenous LanguagesabstractNearly half of Brazil's 180 Indigenous languages face extinction within the next 20 years. What's more concerning is that most of these languages lack a single scientific article describing them, which means they could disappear without leaving any documented evidence of their existence. This work investigates the state of articles about those languages in Wikipedia, both in the English and Portuguese versions, regarded here as indicative of the minimum world-level trace of the previous existence of these languages. Our study shows that over 30% of these languages do not have a single Wikipedia article describing them. It also highlights that the Portuguese and English editing communities are not only distinct, but have different practices, achieving similar levels of quality through different temporal dynamics. These results, although encouraging, suggest that any effort to enhance coverage comprehensiveness in both Wikipedias should consider different strategies for engaging each editing community. Marisa A. Vasconcelos, Priscila de Souza Mizukami, Claudio S. Pinhanez |
ICWSM | 3 |
| 2024 | Creating an African American-Sounding TTS: Guidelines, Technical Challenges, and Surprising EvaluationsabstractRepresentations of AI agents in user interfaces and robotics are predominantly White, not only in terms of facial and skin features, but also in the synthetic voices they use. In this paper we explore some unexpected challenges in the representation of race we found in the process of developing an U.S. English Text-to-Speech (TTS) system aimed to sound like an educated, professional, regional accent-free African American woman. The paper starts by presenting the results of focus groups with African American IT professionals where guidelines and challenges for the creation of a representative and appropriate TTS system were discussed and gathered, followed by a discussion about some of the technical difficulties faced by the TTS system developers. We then describe two studies with U.S. English speakers where the participants were not able to attribute the correct race to the African American TTS voice while overwhelmingly correctly recognizing the race of a White TTS system of similar quality. A focus group with African American IT workers not only confirmed the representativeness of the African American voice we built, but also suggested that the surprising recognition results may have been caused by the inability or the latent prejudice from non-African Americans to associate educated, non-vernacular, professionally-sounding voices to African American people. Claudio S. Pinhanez, Raul Fernandez, Marcelo Grave, Julio Nogima, Ron Hoory |
IUI | 1 |
| 2024 | Fixing Rogue Memorization in Many-to-One Multilingual Translators of Extremely-Low-Resource Languages by Rephrasing Training SamplesabstractPaulo Cavalin, Pedro Henrique Domingues, Claudio Pinhanez, Julio Nogima. Proceedings of the 2024 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies (Volume 1: Long Papers). 2024. Paulo Rodrigo Cavalin, Pedro Henrique Domingues, Claudio S. Pinhanez, Julio Nogima |
NAACL-HLT | 3 |
| 2023 | Balancing Social Impact, Opportunities, and Ethical Constraints of Using AI in the Documentation and Vitalization of Indigenous LanguagesabstractIn this paper we discuss how AI can contribute to support the documentation and vitalization of Indigenous languages and how that involves a delicate balancing of ensuring social impact, exploring technical opportunities, and dealing with ethical constraints. We start by surveying previous work on using AI and NLP to support critical activities of strengthening Indigenous and endangered languages and discussing key limitations of current technologies. After presenting basic ethical constraints of working with Indigenous languages and communities, we propose that creating and deploying language technology ethically with and for Indigenous communities forces AI researchers and engineers to address some of the main shortcomings and criticisms of current technologies. Those ideas are also explored in the discussion of a real case of development of large language models for Brazilian Indigenous languages. Claudio S. Pinhanez, Paulo Rodrigo Cavalin, Marisa A. Vasconcelos, Julio Nogima |
IJCAI | 1 |
| 2022 | Unveiling Practices of Customer Service Content Curators of Conversational AgentsabstractConversational interfaces require two types of curation: data curation by data science workers and content curation by domain experts. Recent years have seen the possibilities for content curators to instruct conversational machines in the customer service domain (i.e., Machine Teaching). The activities of curating specialized data are time-consuming. These activities have a learning curve for the domain expert, and they rely on collaborators beyond the domain experts, including product owners, technology expert curators, management, marketing, and communication employees. However, recent research has looked at making this task easier for domain experts with a lack of knowledge in the Machine Learning system, and few papers have investigated the work practices and collaborations involved in this role. This paper aims to fill this gap, presenting and unveiling practices extracted from eleven semi-structured interviews and four design workshops with experts in Banking, Technical support, Humans Resources, Telecommunications, and Automotive sectors. First, we investigate the articulation work of the content curators and tech curators in training conversational machines. Second, we inspect the curatorial and collaboration strategies they use, which are not afforded by current conversational platforms. Third, we draw the design implications and possibilities to support individual and collaboration curating practices. We reflect on how those practices rely on self and collaboration with others for curation, trust, and data tracking and ownership. Heloisa Candello, Claudio S. Pinhanez, Michael J. Muller, Mairieli Santos Wessel |
Proc. ACM Hum. Comput. Interact. | 2 |
| 2021 | Using Meta-Knowledge Mined from Identifiers to Improve Intent Recognition in Conversational SystemsabstractClaudio Pinhanez, Paulo Cavalin, Victor Henrique Alves Ribeiro, Ana Appel, Heloisa Candello, Julio Nogima, Mauro Pichiliani, Melina Guerra, Maira de Bayser, Gabriel Malfatti, Henrique Ferreira. Proceedings of the 59th Annual Meeting of the Association for Computational Linguistics and the 11th International Joint Conference on Natural Language Processing (Volume 1: Long Papers). 2021. Claudio S. Pinhanez, Paulo Rodrigo Cavalin, Victor Henrique Alves Ribeiro, Ana Paula Appel, Heloisa Candello, Julio Nogima, Mauro Pichiliani, Melina Alberio Guerra, Maíra Gatti de Bayser, Gabriel Louzada Malfatti, Henrique Ferreira |
ACL/IJCNLP (1) | 1 |
| 2021 | Integrating Machine Learning Data with Symbolic Knowledge from Collaboration Practices of Curators to Improve Conversational SystemsabstractThis paper describes how machine learning training data and symbolic knowledge from curators of conversational systems can be used together to improve the accuracy of those systems and to enable better curatorial tools. This is done in the context of a real-world practice of curators of conversational systems who often embed taxonomically-structured meta-knowledge into their documentation. The paper provides evidence that the practice is quite common among curators, that is used as part of their collaborative practices, and that the embedded knowledge can be mined by algorithms. Further, this meta-knowledge can be integrated, using neuro-symbolic algorithms, to the machine learning-based conversational system, to improve its run-time accuracy and to enable tools to support curatorial tasks. Those results point towards new ways of designing development tools which explore an integrated use of code and documentation by machines. Claudio S. Pinhanez, Heloisa Candello, Paulo Rodrigo Cavalin, Mauro Pichiliani, Ana Paula Appel, Victor Henrique Alves Ribeiro, Julio Nogima, Maíra Gatti de Bayser, Melina Alberio Guerra, Henrique Ferreira, Gabriel Louzada Malfatti |
CHI | 1 |
| 2021 | Towards a Method to Classify Language Style for Enhancing Conversational SystemsabstractChatbots have received significant attention in the last years. These systems have improved operational efficiency by reducing the cost of customer service and more and more are used in customer service channels. More recently, with the addition of speech capabilities added to those systems, bot language modifications have to be done in order to improve user experience. In this paper, we analyze the language used in those systems using datasets collected from real chatbots. For that, we first propose models that are able to identify the language style (writing, speech, and computer-mediated) using as linguist features syntactic (Biber's dimensions) and sentence embeddings as well as state-of-art datasets. Our results show that our models were able to distinguish among the three classes with an accuracy of up to 66%. Finally, we evaluate real chatbots systems, either speech- and text-based ones, using our proposed models. We found that these real chatbots are generally using a computer-mediated style, but there is indication that their developers tend to adapt the language style to be more speech-like when text-to-speech is used. Paulo Rodrigo Cavalin, Victor Henrique Alves Ribeiro, Marisa A. Vasconcelos, Claudio S. Pinhanez, Julio Nogima, Henrique Ferreira |
IJCNN | 4 |
| 2020 | Improving Out-of-Scope Detection in Intent Classification by Using Embeddings of the Word Graph Space of the ClassesabstractThis paper explores how intent classification can be improved by representing the class labels not as a discrete set of symbols but as a space where the word graphs associated to each class are mapped using typical graph embedding techniques.The approach, inspired by a previous algorithm used for an inverse dictionary task, allows the classification algorithm to take in account inter-class similarities provided by the repeated occurrence of some words in the training examples of the different classes.The classification is carried out by mapping text embeddings to the word graph embeddings of the classes.Focusing solely on improving the representation of the class label set, we show in experiments conducted in both private and public intent classification datasets, that better detection of out-of-scope examples (OOS) is achieved and, as a consequence, that the overall accuracy of intent classification is also improved.In particular, using the recently-released Larson dataset, an error of about 9.9% has been achieved for OOS detection, beating the previous state-of-the-art result by more than 31 percentage points. Paulo Rodrigo Cavalin, Victor Henrique Alves Ribeiro, Ana Paula Appel, Claudio S. Pinhanez |
EMNLP (1) | 4 |
| 2020 | Creating Corpora for Seq2Seq Tone Rephrasing Using Social Media PostsabstractWe present a methodology to use Twitter posts to create a parallel corpus which can be used to train Seq2Seq neural networks for a tone rephrasing task. Given that people tend to post texts expressing opinions or emotions of varied intensities regarding given real-world events, the main idea is to create corpus containing pairs of posts with opposite tone but about the same topic. By doing so we overcome the main limitation of current tone rephrasing methods: the lack of appropriate parallel training corpora. We explore different methods to create the datasets, including some which require some level of manual labelling. The results show that a completely automatic generation from Twitter data yields training datasets which are better than those with manual interventions, and good enough for Seq2Seq models to outperform non-Seq2Seq models trained with similar data. Paulo Rodrigo Cavalin, Marisa A. Vasconcelos, Marcelo Grave, Claudio S. Pinhanez |
IJCNN | 4 |
| 2019 | The Effect of Audiences on the User Experience with Conversational Interfaces in Physical SpacesabstractHow does the presence of an audience influence the social interaction with a conversational system in a physical space? To answer this question, we analyzed data from an art exhibit where visitors interacted in natural language with three chatbots representing characters from a book. We performed two studies to explore the influence of audiences. In Study 1, we did fieldwork cross-analyzing the reported perception of the social interaction, the audience conditions (visitor is alone, visitor is observed by acquaintances and/or strangers), and control variables such as the visitor's familiarity with the book and gender. In Study 2, we analyzed over 5,000 conversation logs and video recordings, identifying dialogue patterns and how they correlated with the audience conditions. Some significant effects were found, suggesting that conversational systems in physical spaces should be designed based on whether other people observe the user or not. Heloisa Candello, Claudio S. Pinhanez, Mauro Pichiliani, Paulo Rodrigo Cavalin, Flavio Figueiredo, Marisa A. Vasconcelos, Haylla Conde |
CHI | 2 |
| 2018 | Age density patterns in patients medical conditions: A clustering approachabstractThis paper presents a data analysis framework to uncover relationships between health conditions, age and sex for a large population of patients. We study a massive heterogeneous sample of 1.7 million patients in Brazil, containing 47 million of health records with detailed medical conditions for visits to medical facilities for a period of 17 months. The findings suggest that medical conditions can be grouped into clusters that share very distinctive densities in the ages of the patients. For each cluster, we further present the ICD-10 chapters within it. Finally, we relate the findings to comorbidity networks, uncovering the relation of the discovered clusters of age densities to comorbidity networks literature. Fahad Alhasoun, Faisal Aleissa, May Alhazzani, Luis Gregorio Moyano, Claudio S. Pinhanez, Marta C. González |
PLoS Comput. Biol. | 5 |
| 2017 | Typefaces and the Perception of Humanness in Natural Language ChatbotsabstractHow much do visual aspects influence the perception of users about whether they are conversing with a human being or a machine in a mobile-chat environment? This paper describes a study on the influence of typefaces using a blind Turing test-inspired approach. The study consisted of two user experiments. First, three different typefaces (OCR, Georgia, Helvetica) and three neutral dialogues between a human and a financial adviser were shown to participants. The second experiment applied the same study design but OCR font was substituted by Bradley font. For each of our two independent experiments, participants were shown three dialogue transcriptions and three typefaces counterbalanced. For each dialogue typeface pair, participants had to classify adviser conversations as human or chatbot-like. The results showed that machine-like typefaces biased users towards perceiving the adviser as machines but, unexpectedly, handwritten-like typefaces had not the opposite effect. Those effects were, however, influenced by the familiarity of the user to artificial intelligence and other participants' characteristics. Heloisa Candello, Claudio S. Pinhanez, Flavio Figueiredo |
CHI | 2 |
| 2017 | Design Methods for Personified Interfaces
Claudio S. Pinhanez |
CHIRA | 1 |
| 2013 | Large-Scale Multi-agent-Based Modeling and Simulation of Microblogging-Based Online Social Network
Maíra Gatti de Bayser, Paulo Rodrigo Cavalin, Samuel Martins Barbosa Neto, Claudio S. Pinhanez, Cícero Nogueira dos Santos, Daniel Gribel, Ana Paula Appel |
MABS | 4 |
| 2009 | Multimedia Chat for Helpdesks: A Requirements Study, a Practical SOA Architecture, and an Interface PrototypeabstractIn this work we study the actual requirements of a multimedia chat system for helpdesks by quantifying the verbal imagery communication in 115 real transcripts. The data shows that in chats which use verbal imagery, it corresponds to about 25% of the total time, most of it (75%) coming from the agent. These results clearly show that multimedia support for contact center chat systems is called for. But unlike in personal chat systems, the existing 3rdparty chat services for helpdesk centers are text-based only. To circumvent this issue, an innovative, practical SOA-based chat system architecture is proposed in this paper that creates multimedia chat services by decomposition and re-composition of the services from an existing 3rdparty text-based helpdesk chat product. Our SOA architecture implements uniquely a session-based interactive mash up mechanism and bundles it to the previous operations in the same chat session. However, to be effective in the context of technical support, a multimedia chat system must be provided with user and agent interfaces which address the distinctive aspects of technical support communication and imagery use in the conversations. We present an interface prototype that address those issues through the use of persistent imagery, linking instructions and imagery, use of common images,and easy screen sharing. Zon-Yin Shae, Tony Bergstrom, Claudio S. Pinhanez, Mark Podlaseck |
ISORC | 3 |
| 2007 | Editorial PerCom 2007 special issue
Thomas La Porta, Matt W. Mutka, Claudio S. Pinhanez, Peter Steenkiste |
Pervasive Mob. Comput. | 3 |
| 2005 | A study on the manipulation of 2D objects in a projector/camera-based augmented reality environmentabstractAre the object manipulation techniques traditionally used in head-mounted displays (HMDs) applicable to augmented reality based projection systems? This paper examines the differences between HMD- and projector/camera-based AR interfaces in the light of a manipulation task involving documents and applications projected on common office surfaces such as tables, walls, cabinets, and floor. We report a Wizard of Oz study where subjects were first asked to create gesture/voice commands to move 2D objects on those surfaces and then exposed to gestures created by the authors. Among the options, subjects could select the object to be manipulated using voice command; touching, pointing, and grabbing gesture; or a virtual mouse. The results show a strong preference for a manipulation interface based on pointing gestures using small hand movements and involving minimal body movement. Direct touching of the object was also common when the object being manipulated was within the subjects' arm reach. Based on these results, we expect that the preferred interface resembles, in many ways, the egocentric model traditionally used in AR. Stephen Voida, Mark Podlaseck, Rick Kjeldsen, Claudio S. Pinhanez |
CHI | 4 |
| 2005 | To Frame or Not to Frame: The Role and Design of Frameless Displays in Ubiquitous Applications
Claudio S. Pinhanez, Mark Podlaseck |
UbiComp | 1 |
| 2005 | Using Symbiotic Displays to View Sensitive Information in PublicabstractSymbiotic displays are intelligent network devices that can be enlisted on demand by mobile devices, offering higher resolutions and larger viewing areas. Users of small mobile devices can consequently map various kinds of information to the displays that are most appropriate for comfortably viewing them. The availability of these displays in public areas, where passersby may glimpse another's sensitive information, introduces a privacy issue. To address this, we suggest a technique that combines the blurring of sensitive words on the public symbiotic display with a user interface that allows these words to be read on the display of a personal mobile device. We describe three ways in which such interactions can be enabled and discuss the architectural trade-offs associated with these approaches. We also describe a prototype implementation of an email reading application utilizing an IBM WatchPad and an Everywhere Display projector. This prototype was demonstrated at PerCom'04 Stefan Berger, Rick Kjeldsen, Chandrasekhar Narayanaswami 0001, Claudio S. Pinhanez, Mark Podlaseck, Mandayam T. Raghunath |
PerCom | 4 |
| 2004 | Dynamically reconfigurable vision-based user interfaces
Rick Kjeldsen, Anthony Levas, Claudio S. Pinhanez |
Mach. Vis. Appl. | 3 |
| 2003 | An Architecture and Framework for Steerable Interface Systems
Anthony Levas, Claudio S. Pinhanez, Gopal Sarma Pingali, Rick Kjeldsen, Mark Podlaseck, Noi Sukaviriya |
UbiComp | 2 |
| 2003 | Dynamically Reconfigurable Vision-Based User Interfaces
Rick Kjeldsen, Anthony Levas, Claudio S. Pinhanez |
ICVS | 3 |
| 2003 | Embedding Interactions in a Retail Store Environment: The Design and Lessons Learned
Noi Sukaviriya, Mark Podlaseck, Rick Kjeldsen, Anthony Levas, Gopal Sarma Pingali, Claudio S. Pinhanez |
INTERACT | 6 |
| 2003 | Creating touch-screens anywhere with interactive projected displaysabstractWe demonstrate a system that combines steerable projection and computer vision technologies to create "touch-screen" style interactive displays on any flat surface in a space. A high-end version of the system -- the Everywhere Display (ED) -- combines an LCD projector with motorized focus and zoom, a computer controlled pan-tilt mirror, and a pan-tilt zoom camera to enable steering of interactive projections around space. A low-end version (ED-lite) enables creation of interactive displays using a portable projector and camera attached to a laptop computer. Unlike traditional augmented reality systems, the ED systems enable delivery of interactive multimedia content on ordinary objects without requiring users to wear head mounted displays or carry special input devices. Claudio S. Pinhanez, Rick Kjeldsen, Lijun Tang, Anthony Levas, Mark Podlaseck, Noi Sukaviriya, Gopal Sarma Pingali |
ACM Multimedia | 1 |
| 2003 | Steerable Interfaces for Pervasive Computing SpacesabstractThis paper introduces a new class of interactive interfaces that can be moved around to appear on ordinary objects and surfaces anywhere in a space. By dynamically adapting the form, function, and location of an interface to suit the context of the user, such steerable interfaces have the potential to offer radically new and powerful styles of interaction in intelligent pervasive computing spaces. We propose defining characteristics of steerable interfaces and present the first steerable interface system that combines projection, gesture recognition, user tracking, environment modeling and geometric reasoning components within a system architecture. Our work suggests that there is great promise and rich potential for further research on steerable interfaces. Gopal Sarma Pingali, Claudio S. Pinhanez, Anthony Levas, Rick Kjeldsen, Mark Podlaseck, Noi Sukaviriya |
PerCom | 2 |
| 2003 | Interval scripts: a programming paradigm for interactive environments and agents
Claudio S. Pinhanez, Aaron F. Bobick |
Pers. Ubiquitous Comput. | 1 |
| 2002 | User-following displaysabstractTraditionally, a user has positioned himself/herself to be in front of a display in order to access information from it. In this information age, life at work and even at home is often confined to be in front of a display device that is the source of information or entertainment. The paper introduces another paradigm where the display follows the user rather than the user being tied to the display. We demonstrate how steerable projection and people tracking can be combined to achieve a display that automatically follows the user. Gopal Sarma Pingali, Claudio S. Pinhanez, Anthony Levas, Rick Kjeldsen, Mark Podlaseck |
ICME (1) | 2 |
| 2002 | That's Entertainment! Designing Streaming, Multimedia Web ExperiencesabstractThis article investigates the use of streaming multimedia narratives in Web entertainment. Based on experience gained during the user-centered design of a Web site for art and culture, evidence is provided that users want and like "less clicking, more watching" Web experiences where the point of view of experts, artists, or celebrities is presented in a narrative form. A study was conducted where users evaluated 2 prototypes of cultural tours that stream continuously for several minutes unless the user chooses to exercise control over the flow or to explore hotlinks that lead to extra information. Those tours were positively evaluated as both entertaining and engaging. By analyzing mouse activity, it was determined that users who interacted more tended to report less entertainment and engagement. It was also found that such "watchable" experiences are not necessarily a solitary experience and can be enjoyed by groups of people. Finally, users see the Web experiences as a highly enriching and accessible way to augment the cultural experiences and performances they enjoy in brick-and-mortar cultural institutions around the world, rather than as a substitute for them. Clare-Marie Karat, John Karat, John Vergo, Claudio S. Pinhanez, Doug Riecken, Thomas Cofino |
Int. J. Hum. Comput. Interact. | 4 |
| 2002 | BlueSpace: personalizing workspace through awareness and adaptability
Jennifer C. Lai, Anthony Levas, Paul B. Chou, Claudio S. Pinhanez, Marisa S. Viveros |
Int. J. Hum. Comput. Stud. | 4 |
| 2002 | HyperMask - projecting a talking head onto a real object
Tatsuo Yotsukura, Shigeo Morishima, Frank Nielsen, Kim Binsted, Claudio S. Pinhanez |
Vis. Comput. | 5 |
| 2001 | The Everywhere Displays Projector: A Device to Create Ubiquitous Graphical Interfaces
Claudio S. Pinhanez |
UbiComp | 1 |
| 2001 | Less Clicking, More Watching: Results of the Iterative Design and Evaluation of Entertaining Web Experiences
Clare-Marie Karat, Claudio S. Pinhanez, John Karat, Renee Arora, John Vergo |
INTERACT | 2 |
| 1999 | Using Computer Vision to Control a Reactive Computer Graphics Character in a Theater Play
Claudio S. Pinhanez, Aaron F. Bobick |
ICVS | 1 |
| 1998 | Human Action Detection Using PNF Propagation of Temporal ConstraintsabstractIn this paper we develop a representation for the temporal structure inherent in human actions and demonstrate an effective method for using that representation to detect the occurrence of actions. The temporal structure of the action, sub-actions, events, and sensor information is described using a constraint network based on Allen's interval algebra. We map these networks onto a simpler, S-valued domain (past, now, fut) network-a PNF-network-to allow fast detection of actions and sub-actions. The occurrence of an action is computed by considering the minimal domain of its PNF-network, under constraints imposed by the current state of the sensors and the previous states of the network. We illustrate the approach with examples, showing that a major advantage of PNF propagation is the detection and removal of in-consistent situations. Claudio S. Pinhanez, Aaron F. Bobick |
CVPR | 1 |
| 1997 | Interval Scripts: a Design Paradigm for Story-Based Interactive SystemsabstractA system to manage human interaction in immersive environments was designed and implemented.The interaction is defined by an interval scripi which describes the relationships between the time intervals which command actuators or gather information from sensors.With this formalism, reactive, linear, and tree-like interaction can be equally described, as well as less regular story and interaction patterns.Control of actuators and sensors is accomplished using PNF-restriction, a calculus which propagates the sensed information through the interval script determining which intervals are or should be happening at each moment.The prototype was used in an immersive, story-based interactive environment called SingSong where a user or a performer tries to conduct four computer character singers in spite of the hostility of one of them. Claudio S. Pinhanez, Kenji Mase, Aaron F. Bobick |
CHI | 1 |
| 1997 | Controlling view-based algorithms using approximate world models and action informationabstractMost view-based vision algorithms are based on strong assumptions about the disposition of the objects in the image. To safely apply those algorithms in real world image sequences, we propose that a vision system should be divided into two components. The first component contains an approximate world model of the scene-a low accuracy, coarse description of the objects and actions in the world. Approximate world models are constructed and updated by simple vision routines and by the use of action information provided by an external source. The second component employs view-based algorithms to perform required perceptual tasks; the selection and control of the view-based methods are determined by the information provided by the approximate world model. We demonstrate the approximate world model approach in a project to control cameras in a TV studio where the external context is provided by a script. Aaron F. Bobick, Claudio S. Pinhanez |
CVPR | 2 |
| 1994 | Behavior-Based Active VisionabstractA vision system was built using a behavior-based model, the subsumption architecture. The so-called active eye moves the camera’s axis through the environment, detecting areas with high concentration of edges, with the help of a kind of saccadic movement. The design and implementation process is detailed in the article, paying particular attention to the fovea-like sensor structure which enables the active eye to efficiently use local information to control its movements. Numerical measures for the eye’s behavior were developed, and applied to evaluate the incremental building process and the effects of the saccadic movements on the whole system. A higher level behavior was also implemented, with the purpose of detecting long straight edges in the image, producing pictures similar to hand drawings. Robustness and efficiency problems are addressed at the end of the paper. The results seem to prove that interesting behaviors can be achieved using simple vision methods and algorithms, if their results are properly interconnected and timed. Claudio S. Pinhanez |
Int. J. Pattern Recognit. Artif. Intell. | 1 |