VLDB 2026 Research / reviewers in the wild / expert
Scott A. Carter
dblp:59/4765
· DBLP profile ↗
45ranked-venue papers
13as first author
10since 2021 · last 2026
0000-0002-2942-972XORCID · reported
Domains — the database's venue-derived domains; a paper can count in several
Human-computer interaction and ubiquitous computing · 24 · 7 first-author · 6 since 2021Databases, data management, data science and information retrieval · 8 · 1 first-authorGraphics, computer vision, multimedia, augmented reality and games · 8 · 5 first-authorArtificial intelligence and machine learning · 5 · 4 since 2021Applied, interdisciplinary, general and emerging computing · 3 · 3 since 2021
| Year | Publication | Venue | Position |
|---|---|---|---|
| 2026 | A Framework for Adapting In-Car Touchscreen Interfaces to Driver Behaviors, Perception, and Cognition
Seokhyun Hwang, Xiyuan Shen, Alex Filipowicz, Andrew Best, Jean Marcel dos Reis Costa, Scott A. Carter, James Fogarty, Jacob O. Wobbrock |
CHI | 6 |
| 2025 | Empathy Prediction from Diverse PerspectivesabstractFrancine Chen, Scott Carter, Tatiana Lau, Nayeli Suseth Bravo, Sumanta Bhattacharyya, Kate Sieck, Charlene C. Wu. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers). 2025. Francine Chen 0001, Scott A. Carter, Tatiana Lau, Nayeli Bravo, Sumanta Bhattacharyya, Katharine Sieck, Charlene C. Wu |
ACL (1) | 2 |
| 2025 | Touchscreens in Motion: Quantifying the Impact of Cognitive Load on Distracted Drivers
Xiyuan Shen, Seokhyun Hwang, Junhan Kong, Alex Filipowicz, Andrew Best, Jean Marcel dos Reis Costa, Scott A. Carter, James Fogarty, Jacob O. Wobbrock |
UIST | 7 |
| 2023 | More human than human: LLM-generated narratives outperform human-LLM interleaved narrativesabstractNarrative story generation has gained emerging interest in the field of large language models. The present paper aims to compare stories generated by an LLM only (non-interleaved) with those generated by interleaving human-generated and LLM-generated text (interleaved). The study’s hypothesis is that interleaved stories would perform better than non-interleaved stories. To verify this hypothesis, we conducted two tests with roughly 500 participants each. Participants were asked to rate stories of each type, including an overall score or preference and four facets—logical soundness, plausibility, understandability, and novelty. Our findings indicate that interleaved stories were in fact less preferred than non-interleaved stories. The result has implications for the design and implementation of our story generators. This study contributes new insights into the potential uses and restrictions of interleaved and non-interleaved systems regarding generating narrative stories, which may help to improve the performance of such story generators. Kexin Zhao 0010, Sophie Song, Bridget Duah, Jamie C. Macbeth, Scott A. Carter, Monica P. Van, Nayeli Bravo, Matthew Klenk 0001, Katharine Sieck, Alex Filipowicz |
Creativity & Cognition | 5 |
| 2023 | Understanding People's Perception and Usage of Plug-in Electric HybridsabstractElectrification is an important first step toward reducing the greenhouse emissions of passenger vehicles. However, how drivers drive, charge, and operate their electrified vehicles can have a large impact on their emissions, particularly for Plug-in Hybrid Electric vehicles (PHEVs) that combine all-electric driving with an internal combustion engine. In this paper, we investigate how and why drivers use their PHEVs and uncover design opportunities for interfaces that can support the efficient use of PHEVs. We used a mixed-method approach combining quantitative, qualitative, and concept elicitation methods with PHEV owners in the US. While past findings indicate that PHEV drivers are not motivated to charge regularly, our work contradicts this with evidence of (1) regular charging with home infrastructure, (2) high cost sensitivity, and (3) preference for driving in all-electric mode. Our results indicate that the most critical problem is inadequate user support for navigating poor charging infrastructure. Matthew L. Lee, Scott A. Carter, Rumen Iliev, Nayeli Bravo, Monica P. Van, Laurent Denoue, Everlyne Kimani, Alex Filipowicz, David A. Shamma, Katharine Sieck, Candice Hogan, Charlene C. Wu |
CHI | 2 |
| 2023 | Save A Tree or 6 kg of CO2? Understanding Effective Carbon Footprint Interventions for Eco-Friendly Vehicular ChoicesabstractFrom ride-hailing to car rentals, consumers are often presented with eco-friendly options. Beyond highlighting a “green” vehicle and CO2 emissions, CO2 equivalencies have been designed to provide understandable amounts; we ask which equivalencies will lead to eco-friendly decisions. We conducted five ride-hailing scenario surveys where participants picked between regular and eco-friendly options, testing equivalencies, social features, and valence-based interventions. Further, we tested a car-rental embodiment to gauge how an individual (needing a car for several days) might behave versus the immediate ride-hailing context. We find that participants are more likely to choose green rides when presented with additional information about emissions; CO2 by weight was found to be the most effective. Further, we found that information framing—be it individual or collective footprint, positive or negative valence—had an impact on participants’ choices. Finally, we discuss how our findings inform the design of effective interventions for reducing car-based carbon-emissions. Vikram Mohanty, Alex Filipowicz, Nayeli Bravo, Scott A. Carter, David A. Shamma |
CHI | 4 |
| 2023 | Machine learning-based measure of cognitive complexity explains variance in rank-ordered preference
Shabnam Hakimi, Yan-Ying Chen, Monica P. Van, Scott A. Carter, Emily S. Sumner, Nayeli Bravo, Kalani Murakami, Charlene C. Wu, Matthew Klenk 0001 |
CogSci | 4 |
| 2023 | Can Behavioral Experts Predict Outcome Heterogeneity?
Rumen Iliev, Alex Filipowicz, Emily S. Sumner, Francine Chen 0001, Nikos Aréchiga, Scott A. Carter, Totte Harinen, Katharine Sieck, Charlene C. Wu |
CogSci | 6 |
| 2022 | You Complete Me: Human-AI Teams and Complementary ExpertiseabstractPeople consider recommendations from AI systems in diverse domains ranging from recognizing tumors in medical images to deciding which shoes look cute with an outfit. Implicit in the decision process is the perceived expertise of the AI system. In this paper, we investigate how people trust and rely on an AI assistant that performs with different levels of expertise relative to the person, ranging from completely overlapping expertise to perfectly complementary expertise. Through a series of controlled online lab studies where participants identified objects with the help of an AI assistant, we demonstrate that participants were able to perceive when the assistant was an expert or non-expert within the same task and calibrate their reliance on the AI to improve team performance. We also demonstrate that communicating expertise through the linguistic properties of the explanation text was effective, where embracing language increased reliance and distancing language reduced reliance on AI. Qiaoning Zhang, Matthew L. Lee, Scott A. Carter |
CHI | 3 |
| 2022 | Familiarity plays a unique role in increasing preferences for battery electric vehicle adoption
Alex Filipowicz, Charlene C. Wu, Matthew L. Lee, David A. Shamma, Shabnam Hakimi, Scott A. Carter, Rumen Iliev, Totte Harinen, Emily S. Sumner, Candice Hogan |
CogSci | 6 |
| 2019 | Use Your Head! Exploring Interaction Modalities for Hat TechnologiesabstractAs our landscape of wearable technologies proliferates, we find more devices situated on our heads. However, many challenges hinder them from widespread adoption - from their awkward, bulky form factor (today's AR and VR goggles) to their socially stigmatized designs (Google Glass) and a lack of a well-developed head-based interaction design language. In this paper, we explore a socially acceptable, large, head-worn interactive wearable - a hat. We report results from a gesture elicitation study with 17 participants, extract a taxonomy of gestures, and define a set of design concerns for interactive hats. Through this lens, we detail the design and fabrication of three hat prototypes capable of sensing touch, head movements, and gestures, and including ambient displays of several types. Finally, we report an evaluation of our hat prototype and insights to inform the design of future hat technologies. Christine Dierk, Scott A. Carter, Patrick Chiu, Anthony Dunnigan, Don Kimber |
Conference on Designing Interactive Systems | 2 |
| 2019 | Documenting Physical Objects with Live Video and Object DetectionabstractResponding to requests for information from an application, a remote person, or an organization that involve documenting the presence and/or state of physical objects can lead to incomplete or inaccurate documentation. We propose a system that couples information requests with a live object recognition tool to semi-automatically catalog requested items and collect evidence of their current state. Scott A. Carter, Laurent Denoue, Daniel Avrahami |
ACM Multimedia | 1 |
| 2019 | CamaLeon: Smart Camera for Conferencing in the WildabstractDespite work on smart spaces, nowadays a lot of knowledge work happens in the wild: at home, in coffee places, trains, buses, planes, and of course in crowded open office cubicles. Conducting web conferences in these settings creates privacy issues, and can also distract participants, leading to a perceived lack of professionalism from the remote peer(s). To solve this common problem, we implemented CamaLeon, a browser-based tool that uses real-time machine vision powered by deep learning to change the webcam stream sent by the remote peer. Specifically, CamaLeon dynamically changes the "wild" background into one that resembles that of the office workers. In order to detect the background in disparate settings, we designed and trained a fast UNet model on head and shoulder images. CamaLeon also uses a face detector to determine whether it should stream the person's face, depending on its location (or lack of presence). It uses face recognition to make sure it streams only a face that belongs to the user who connected to the meeting. We tested the system during a few real video conferencing calls at our company in which two workers are remote. Both parties felt a sense of enhanced co-presence, and the remote participants felt more professional with their background replaced. Laurent Denoue, Scott A. Carter, Chelhwon Kim |
ACM Multimedia | 2 |
| 2018 | FormYak: Converting forms to conversationsabstractHistorically, people have interacted with companies and institutions through telephone-based dialogue systems and paper-based forms. Now, these interactions are rapidly moving to web- and phone-based chat systems. While converting traditional telephone dialogues to chat is relatively straightforward, converting forms to conversational interfaces can be challenging. In this work, we introduce methods and interfaces to enable the conversion of PDF and web-based documents that solicit user input into chat-based dialogues. Document data is first extracted to associate fields and their textual descriptions using metadata and lightweight visual analysis. The field labels, their spatial layout, and associated text are further analyzed to group related fields into natural conversational units. These correspond to questions presented to users in chat interfaces to solicit information needed to complete the original documents and downstream processes they support. This user supplied data can be inserted into the source documents and/or in downstream databases. User studies of our tool show that it streamlines form-to-chat conversion and produces conversational dialogues of at least the same quality as a purely manual approach. Scott A. Carter, Laurent Denoue, Matthew Cooper 0002, Jennifer Marlow |
DocEng | 1 |
| 2018 | SlideDiff: Animating Textual and Media Changes in SlidesabstractSlideDiff is a system that automatically creates an animated rendering of textual and media differences between two versions of a slide presentation. While previous work focused on either textual or image data, SlideDiff integrates both text and media changes, as well as their interactions, for example when adding an image forces nearby text boxes to shrink. Given two versions of a slide (not the full history of edits), SlideDiff detects the textual and image differences, and then animates the changes by mimicking what a user might have done, such as moving the cursor, typing text, resizing image boxes, adding images. This editing metaphor is well known to most users, helping them better understand what has changed, and fosters a sense of connection between remote workers, derived from communicating both the revision process as well as its results. After detection of text and image differences, the animations are rendered in HTML and CSS, including mouse cursor motion, text and image box selection and resizing, text deletion and insertion with its cursor. We discuss strategies for animating changes, in particular the importance of starting with large changes and finishing with smaller edits, and provide details of the implementation using modern HTML and CSS. Laurent Denoue, Scott A. Carter, Matthew Cooper 0002 |
DocEng | 2 |
| 2017 | DocHandles: Linking Document Fragments in Messaging AppsabstractIn this paper, we describe DocHandles, a novel system that allows users to link to specific document parts in their chat applications. As users type a message, they can invoke the tool by referring to a specific part of a document, e.g., "@fig1 needs revision". By combining text parsing and document layout analysis, DocHandles can find and present all the figures "1" inside previously shared documents, allowing users to explicitly link to the relevant "document handle". In this way, Ddocuments become first-class citizens inside the conversation stream where users can seamlessly integrate documents in their text-centric messaging application. Laurent Denoue, Scott A. Carter, Jennifer Marlow, Matthew Cooper 0002 |
DocEng | 2 |
| 2017 | I Should Listen More: Real-time Sensing and Feedback of Non-Verbal Communication in Video TelehealthabstractVideo telehealth is growing to allow more clinicians to see patients from afar. As a result, clinicians, typically trained for in-person visits, must learn to communicate both health information and non-verbal affective signals to patients through a digital medium. We introduce a system called ReflectLive that senses and provides real-time feedback about non-verbal communication behaviors to clinicians so they can improve their communication behaviors. A user evaluation with 10 clinicians showed that the real-time feedback helped clinicians maintain better eye contact with patients and was not overly distracting. Clinicians reported being more aware of their non-verbal communication behaviors and reacted positively to summaries of their conversational metrics, motivating them to want to improve. Using ReflectLive as a probe, we also discuss the benefits and concerns around automatically quantifying the "soft skills" and complexities of clinician-patient communication, the controllability of behaviors, and the design considerations for how to present real-time and summative feedback to clinicians. Heather A. Faucett, Matthew L. Lee, Scott A. Carter |
Proc. ACM Hum. Comput. Interact. | 3 |
| 2016 | Beyond Talking Heads: Multimedia Artifact Creation, Use, and Sharing in Distributed MeetingsabstractDistributed meetings can be messy, particularly when the task requires collaboration around multimedia artifacts. Teams must not only share a variety of materials related to the work in real time, but also need to refer back to information after a meeting ends. While video tools make it relatively easy to have conversations at a distance, they are less adept at sharing and archiving multimedia content. We conducted a survey of and interviews with members of distributed teams to investigate how they create, use, and share multimedia content before, during, and after distributed meetings. Our findings shed light on decisions made and rationales used in selecting technologies to prepare for, conduct, and archive the results of a video-mediated distributed meeting. The results suggest a need for flexible interfaces for information sharing in multiple meeting contexts so content can be both easily referred to in the moment and also found again later. Jennifer Marlow, Scott A. Carter, Nathaniel Good, Jung-Wei Chen |
CSCW | 2 |
| 2016 | DocuGram: Turning Screen Recordings into DocumentsabstractIn this paper we describe DocuGram, a novel tool to capture and share documents originating from any application. As users scroll through pages of their document inside the native application (Word, Google Docs, web browser), the system captures and analyses in real-time the rendered video frames and reconstitutes the original document pages into an easy to view HTML-based representation. In addition to detecting and regenerating the document pages, a DocuGram also includes the interactions users had over them, e.g. mouse motions and voice comments. A DocuGram allows users to flexibly share enhanced documents across applications. Laurent Denoue, Scott A. Carter, Matthew Cooper 0002 |
DocEng | 2 |
| 2016 | Bringing mobile into meetings: enhancing distributed meeting participation on smartwatches and mobile phonesabstractMost teleconferencing tools treat users in distributed meetings monolithically: all participants are meant to be interconnected in more-or-less the same manner. In practice, people connect to meetings in different contexts, sometimes sitting in front of a laptop or tablet giving their full attention, but at other times mobile and concurrently involved in other tasks or as a liminal participant in a larger group meeting. In this paper, we present the design and evaluation of two applications, MixMeetWear and MixMeetMate, to help users in non-standard contexts flexibly participate in meetings. Scott A. Carter, Jennifer Marlow, Aki Komori, Ville Mäkelä |
MobileHCI | 1 |
| 2016 | WorkCache: Salvaging siloed knowledgeabstractThe proliferation of workplace multimedia collaboration applications has meant on one hand more opportunities for group work but on the other more data locked away in proprietary interfaces. We are developing new tools to capture and access multimedia content from any source. In this demo, we focus primarily on new methods that allow users to rapidly reconstitute, enhance, and share document-based information. Scott A. Carter, Laurent Denoue, Matthew Cooper 0002 |
ACM Multimedia | 1 |
| 2015 | Searching Live Meeting Documents "Show me the Action"abstractLive meeting documents require different techniques for effectively retrieving important pieces of information. During live meetings, people share web sites, edit presentation slides, and share code editors. A simple approach is to index with Optical Character Recognition (OCR) the video frames, or key-frames, being shared and let user retrieve them. Here we show that a more useful approach is to look at what actions users take inside the live document streams. Based on observations of real meetings, we focus on two important signals: text editing and mouse cursor motion. We describe the detection of text and cursor motion, their implementation in our WebRTC (Web Real-Time Communication)-based system, and how users are better able to search live documents during a meeting based on these extracted actions. Laurent Denoue, Scott A. Carter, Matthew Cooper 0002 |
DocEng | 2 |
| 2015 | Searching and Browsing Live, Web-based MeetingsabstractEstablishing common ground is one of the key problems for any form of communication. The problem is particularly pronounced in remote meetings, in which participants can easily lose track of the details of dialogue for any number of reasons. In this demo we present a web-based tool, MixMeet, that allows teleconferencing participants to search the contents of live meetings so they can rapidly retrieve previously shared content to get on the same page, correct a misunderstanding, or discuss a new idea. Scott A. Carter, Laurent Denoue, Matthew Cooper 0002 |
ACM Multimedia | 1 |
| 2014 | Building digital project rooms for web meetingsabstractDistributed teams must co-ordinate a variety of tasks. To do so they need to be able to create, share, and annotate documents as well as discuss plans and goals. Many workflow tools support document sharing, while other tools support videoconferencing. However, there exists little support for connecting the two. In this work, we describe a system that allows users to share and markup content during web meetings. This shared content can provide important conversational props within the context of a meeting; it can also help users review archived meetings. Users can also extract content from meetings directly into their personal notes or other workflow tools. Laurent Denoue, Scott A. Carter, Andreas Girgensohn, Matthew Cooper 0002 |
ACM Symposium on Document Engineering | 2 |
| 2013 | Content-based copy and paste from video documentsabstractUnlike text, copying and pasting parts of video documents is challenging. Yet, the abundance of video documents now available including how-to tutorials requires simpler tools that allow users to easily copy and paste fragments of video materials into new documents. We describe new direct video manipulation techniques enabling users to quickly copy and paste content from video documents into a user's own multimedia document. While the video plays, users interact with the video canvas to select text regions, scrollable regions, slide sequences built up across many frames, or semantically meaningful regions such as dialog boxes. Instead of relying on the timeline to accurately select sub-parts of the video document, users navigate using familiar selection techniques such as mouse-wheel to scroll back and forward over a video shot in which the content scrolls, double-clicks over rectangular regions to select them, or clicks and drags over textual regions of the video canvas to select them. We describe the video processing techniques that run in real-time in modern web browsers using HTML5 and JavaScript; and show how they help users quickly copy and paste video fragments into new documents, allowing them to efficiently reuse video documents for authoring or note-taking. Laurent Denoue, Scott A. Carter, Matthew Cooper 0002 |
ACM Symposium on Document Engineering | 2 |
| 2013 | SmartDCap: semi-automatic capture of higher quality document images from a smartphoneabstractPeople frequently capture photos with their smartphones, and some are starting to capture images of documents. However, the quality of captured document images is often lower than expected, even when an application that performs post-processing to improve the image is used. To improve the quality of captured images before post-processing, we developed the Smart Document Capture (SmartDCap) application that provides real-time feedback to users about the likely quality of a captured image. The quality measures capture the sharpness and framing of a page or regions on a page, such as a set of one or more columns, a part of a column, a figure, or a table. Using our approach, while users adjust the camera position, the application automatically determines when to take a picture of a document to produce a good quality result. We performed a subjective evaluation comparing SmartDCap and the Android Ice Cream Sandwich (ICS) camera application; we also used raters to evaluate the quality of the captured images. Our results indicate that users find SmartDCap to be as easy to use as the standard ICS camera application. Also, images captured using SmartDCap are sharper and better framed on average than images using the ICS camera application. Francine Chen 0001, Scott A. Carter, Laurent Denoue, Jayant Kumar |
IUI | 2 |
| 2012 | Understanding screen contents for building a high performance, real time screen sharing systemabstractFaithful sharing of screen contents is an important collaboration feature. Prior systems were designed to operate over constrained networks. They performed poorly even without such bottlenecks. To build a high performance screen sharing system, we empirically analyzed screen contents for a variety of scenarios. We showed that screen updates were sporadic with long periods of inactivity. When active, screens were updated at far higher rates than was supported by earlier systems. The mismatch was pronounced for interactive scenarios. Even during active screen updates, the number of updated pixels were frequently small. We showed that crucial information can be lost if individual updates were merged. When the available system resources could not support high capture rates, we showed ways in which updates can be effectively collapsed. We showed that Zlib lossless compression performed poorly for screen updates. By analyzing the screen pixels, we developed a practical transformation that significantly improved compression rates. Our system captured 240 updates per second while only using 4.6 Mbps for interactive scenarios. Still, while playing movies in fullscreen mode, our approach could not achieve higher capture rates than prior systems; the CPU remains the bottleneck. A system that incorporates our findings is deployed within the lab. Surendar Chandra, Jacob T. Biehl, John S. Boreczky, Scott A. Carter, Lawrence A. Rowe |
ACM Multimedia | 4 |
| 2011 | DiG: a task-based approach to product searchabstractWhile there are many commercial systems to help people browse and compare products, these interfaces are typically product centric. To help users identify products that match their needs more efficiently, we instead focus on building a task centric interface and system. Based on answers to initial questions about the situations in which they expect to use the product, the interface identifies products that match their needs, and exposes high-level product features related to their tasks, as well as low-level information including customer reviews and product specifications. We developed semi-automatic methods to extract the high-level features used by the system from online product data. These methods identify and group product features, mine and summarize opinions about those features, and identify product uses. User studies verified our focus on high-level features for browsing products and low-level features and specifications for comparing products. Scott A. Carter, Francine Chen 0001, Aditi S. Muralidharan, Jeremy Pickens |
IUI | 1 |
| 2011 | ARA: the active reading applicationabstractThe Active Reading Application (ARA) brings the familiar experience of writing on paper to the tablet. The application augments paper-based practices with audio, the ability to review annotations, and sharing. It is designed to make it easier to review, annotate, and comment on documents by individuals and groups. ARA incorporates several patented technologies and draws on several years of research and experimentation. Gene Golovchinsky, Scott A. Carter, Anthony Dunnigan |
ACM Multimedia | 2 |
| 2010 | Let's go from the whiteboard: supporting transitions in work through whiteboard capture and reuseabstractThe use of whiteboards is pervasive across a wide range of work domains. But some of the qualities that make them successful--an intuitive interface, physical working space, and easy erasure--inherently make them poor tools for archival and reuse. If whiteboard content could be made available in times and spaces beyond those supported by the whiteboard alone, how might it be appropriated? We explore this question via ReBoard, a system that automatically captures whiteboard images and makes them accessible through a novel set of user-centered access tools. Through the lens of a seven week workplace field study, we found that by enabling new workflows, ReBoard increased the value of whiteboard content for collaboration. Stacy M. Branham, Gene Golovchinsky, Scott A. Carter, Jacob T. Biehl |
CHI | 3 |
| 2010 | FormCracker: interactive web-based form fillingabstractFilling out document forms distributed by email or hosted on the Web is still problematic and usually requires a printer and scanner. Users commonly download and print forms, fill them out by hand, scan and email them. Even if the document is form-enabled (PDFs with FDF information), to read the file users still have to launch a separate application which may not be available, especially on mobile devices. Laurent Denoue, John Adcock, Scott A. Carter, Patrick Chiu, Francine Chen 0001 |
ACM Symposium on Document Engineering | 3 |
| 2010 | NudgeCam: toward targeted, higher quality media captureabstractNudgeCam is a mobile application that can help users capture more relevant, higher quality media. To guide users to capture media more relevant to a particular project, third-party template creators can show users media that demonstrates relevant content and can tell users what content should be present in each captured media using tags and other meta-data such as location and camera orientation. To encourage higher quality media capture, NudgeCam provides real time feedback based on standard media capture heuristics, including face positioning, pan speed, audio quality, and many others. We describe an implementation of NudgeCam on the Android platform as well as field deployments of the application. Scott A. Carter, John Adcock, John Doherty, Stacy M. Branham |
ACM Multimedia | 1 |
| 2009 | DICE: designing conference rooms for usabilityabstractOne of the core challenges now facing smart rooms is supporting realistic, everyday activities. While much research has been done to push forward the frontiers of novel interaction techniques, we argue that technology geared toward widespread adoption requires a design approach that emphasizes straightforward configuration and control, as well as flexibility. We examined the work practices of users of a large, multi-purpose conference room, and designed DICE, a system to help them use the room's capabilities. We describe the design process, and report findings about the system's usability and about people's use of a multi-purpose conference room. Gene Golovchinsky, Pernilla Qvarfordt, William van Melle, Scott A. Carter, Anthony Dunnigan |
CHI | 4 |
| 2009 | Kartta: extracting landmarks near personalized points-of-interest from user generated contentabstractMost mobile navigation systems focus on answering the question, "I know where I want to go, now can you show me exactly how to get there?" While this approach works well for many tasks, it is not as useful for unconstrained situations in which user goals and spatial landscapes are more fluid, such as festivals or conferences. In this paper we describe the design and iteration of the Kartta system, which we developed to answer a slightly different question: "What are the most interesting areas here and how do I find them?" Arttu Perttula, Scott A. Carter, Laurent Denoue |
Mobile HCI | 2 |
| 2008 | PicNTell: a camcorder metaphor for screen recordingabstractPicNTell is a new technique for generating compelling screencasts where users can quickly record desktop activities and generate videos that are embeddable on popular video sharing distributions such as YouTube®. While standard video editing and screen capture tools are useful for some editing tasks, they have two main drawbacks: (1) they require users to import and organize media in a separate interface, and (2) they do not support natural (or camcorder-like) screen recording, and instead usually require the user to define a specific region or window to record. In this paper we review current screen recording use, and present the PicNTell system, pilot studies, and a new six degree-of-freedom tracker we are developing in response to our findings. Scott A. Carter, Laurent Denoue |
ACM Multimedia | 1 |
| 2008 | Exiting the Cleanroom: On Ecological Validity and Ubiquitous ComputingabstractOver the past decade and a half, corporations and academies have invested considerable time and money in the realization of ubiquitous computing. Yet design approaches that yield ecologically valid understandings of ubiquitous computing systems, which can help designers make design decisions based on how systems perform in the context of actual experience, remain rare. The central question underlying this article is, What barriers stand in the way of real-world, ecologically valid design for ubicomp? Using a literature survey and interviews with 28 developers, we illustrate how issues of sensing and scale cause ubicomp systems to resist iteration, prototype creation, and ecologically valid evaluation. In particular, we found that developers have difficulty creating prototypes that are both robust enough for realistic use and able to handle ambiguity and error and that they struggle to gather useful data from evaluations because critical events occur infrequently, because the level of use necessary to evaluate the system is difficult to maintain, or because the evaluation itself interferes with use of the system. We outline pitfalls for developers to avoid as well as practical solutions, and we draw on our results to outline research challenges for the future. Crucially, we do not argue for particular processes, sets of metrics, or intended outcomes, but rather we focus on prototyping tools and evaluation methods that support realistic use in realistic settings that can be selected according to the needs and goals of a particular developer or researcher. Scott A. Carter, Jennifer Mankoff, Scott R. Klemmer, Tara Matthews |
Hum. Comput. Interact. | 1 |
| 2007 | Momento: support for situated ubicomp experimentationabstractWe present the iterative design of Momento, a tool that providesintegrated support for situated evaluation of ubiquitouscomputing applications. We derived requirements for Momento from a user-centered design process that includedinterviews, observations and field studies of early versionsof the tool. Motivated by our findings, Momento supportsremote testing of ubicomp applications, helps with participantadoption and retention by minimizing the need for newhardware, and supports mid-to-long term studies to addressinfrequently occurring data. Also, Momento can gather logdata, experience sampling, diary, and other qualitative data. Scott A. Carter, Jennifer Mankoff, Jeffrey Heer |
CHI | 1 |
| 2007 | Defining, Designing, and Evaluating Peripheral Displays: An Analysis Using Activity TheoryabstractPeripheral displays are an important class of ubiquitous computing applications. However, the field has suffered from a lack of clear, consistent terminology surrounding peripheral display research. We present an Activity Theory analysis of peripheral displays in order to establish common terminology and meaning for peripheral displays. We also present an Activity Theory-based approach for designing and evaluating peripheral displays Tara Matthews, Tye Rattenbury, Scott A. Carter |
Hum. Comput. Interact. | 3 |
| 2006 | Dynamically adapting GUIs to diverse input devicesabstractMany of today's desktop applications are designed for use with a pointing device and keyboard. Someone with a disability, or in a unique environment, may not be able to use one or both of these devices. We have developed an approach for automatically modifying desktop applications to accommodate a variety of input alternatives as well as a demonstration implementation, the Input Adapter Tool (IAT). Our work is differentiated from past work by our focus on input adaptation (such as adapting a paint program to work without a pointing device) rather than output adaptation (such as adapting web pages to work on a cellphone). We present an analysis showing how different common interactive elements and navigation techniques can be adapted to specific input modalities. We also describe IAT, which supports a subset of these adaptations, and illustrate how it adapts different inputs to two applications, a paint program and a form entry program. Scott A. Carter, Amy Hurst, Jennifer Mankoff |
ASSETS | 1 |
| 2006 | Scribe4Me: Evaluating a Mobile Sound Transcription Tool for the Deaf
Tara Matthews, Scott A. Carter, Carol Pai, Janette Fong, Jennifer Mankoff |
UbiComp | 2 |
| 2005 | When participants do the capturing: the role of media in diary studiesabstractIn this paper, we investigate how the choice of media for capture and access affects the diary study method. The diary study is a method of understanding participant behavior and intent in situ that minimizes the effects of observers on participants. We first situate diary studies within a framework of field studies and review related literature. We then report on three diary studies we conducted that involve photographs, audio recordings, location information and tangible artifacts. We then analyze our findings, specifically addressing the following questions: How do context information and episodic memory prompts captured by participants vary with media? In what way do different media "jog" memory? How do different media affect the diary study process? These questions are particularly important for diary studies because they can be especially useful as compared to other methods when a participant intends to do an action but does not or when actions are particularly difficult to sense. We also built and tested a tool based on participant and researcher frustrations with the method. Our contribution includes suggested modifications to traditional diary techniques that enable annotation and review of captured media; a new variation on the diary study appropriate for researchers using digital capture media; and a lightweight tool to support it, motivated by past work and findings from our studies. Scott A. Carter, Jennifer Mankoff |
CHI | 1 |
| 2004 | A toolkit for managing user attention in peripheral displaysabstractTraditionally, computer interfaces have been confined to conventional displays and focused activities. However, as displays become embedded throughout our environment and daily lives, increasing numbers of them must operate on the periphery of our attention. Peripheral displays can allow a person to be aware of information while she is attending to some other primary task or activity. We present the Peripheral Displays Toolkit (PTK), a toolkit that provides structured support for managing user attention in the development of peripheral displays. Our goal is to enable designers to explore different approaches to managing user attention. The PTK supports three issues specific to conveying information on the periphery of human attention. These issues are abstraction of raw input, rules for assigning notification levels to input, and transitions for updating a display when input arrives. Our contribution is the investigation of issues specific to attention in peripheral display design and a toolkit that encapsulates support for these issues. We describe our toolkit architecture and present five sample peripheral displays demonstrating our toolkit's capabilities. Tara Matthews, Anind K. Dey, Jennifer Mankoff, Scott A. Carter, Tye Rattenbury |
UIST | 4 |
| 2004 | Building Connections among Loosely Coupled Groups: Hebb's Rule at Work
Scott A. Carter, Jennifer Mankoff, P. Goddi |
Comput. Support. Cooperative Work. | 1 |
| 2002 | Distributed mediation of ambiguous context in aware environmentsabstractMany context-aware services make the assumption that the context they use is completely accurate. However, in reality, both sensed and interpreted context is often ambiguous. A challenge facing the development of realistic and deployable context-aware services, therefore, is the ability to handle ambiguous context. In this paper, we describe an architecture that supports the building of context-aware services that assume context is ambiguous and allows for mediation of ambiguity by mobile users in aware environments. We illustrate the use of our architecture and evaluate it through three example context-aware services, a word predictor system, an In/Out Board, and a reminder tool. Anind K. Dey, Jennifer Mankoff, Gregory D. Abowd, Scott A. Carter |
UIST | 4 |
| 2000 | Blind source separation of multichannel neuromagnetic responses
Akaysha C. Tang, Barak A. Pearlmutter, Michael Zibulevsky, Scott A. Carter |
Neurocomputing | 4 |