697 Matching Annotations
  1. Jul 2026
    1. We process this data in a three-stage pipeline (Figure 6). In the first stage, Sentence Segmentation and Categorization, abstracts are split into individual sentences using the NLTK package, and each sentence is classified into one of the five pre-defined aspects as listed in Section 4.1.1. Classification is performed by prompting an LLM (see prompt used in Appendix D.1) with the sentence and its full abstract.

      sentence describing how analysis was performed on data collected by the authors of this paper

    2. Then, we segment sentences within each aspect into grammar-preserving chunks (see prompt used in Appendix D.2). This results in grammatically coherent chunks that are the basis of structure patterns. After identifying chunk boundaries, we again prompt an LLM to generate labels for chunks in a human-in-the-loop approach: starting from an initial set of labels for chunk roles, when a new label is generated, a researcher from the research team examines the new label and merges it with existing labels if appropriate, controlling for the total number of labels.

      sentence relating to methodology

    3. After obtaining an expanded set of high-level chunk labels, we assign them to each of the sentence chunks by using LLMs in a multiclass classification few-shot learning task, with the initial labels and assignment as examples (see prompt used in Appendix D.3).

      sentence describing how analysis was performed on data collected by the authors of this paper

    4. Then, we segment sentences within each aspect into grammarpreserving chunks (see prompt used in Appendix D.2). This results in grammatically coherent chunks that are the basis of structure patterns. After identifying chunk boundaries, we again prompt an LLM to generate labels for chunks in a human-in-the-loop approach: starting from an initial set of labels for chunk roles, when a new label is generated, a researcher from the research team examines the new label and merges it with existing labels if appropriate, controlling for the total number of labels.

      sentence describing how analysis was performed on data collected by the authors of this paper

    5. We conducted a qualitative analysis of user study transcripts and survey responses using a Grounded Theory approach [8]. First, the lead researcher collected a list of participants' behaviors, approaches, reflections on their experience, and feedback about the interface. The researcher then systematically coded this data, revisiting the data multiples times and refining the codes to ensure consistency and coherence. Through this process, high-level themes were identified and organized using affinity diagramming. Once the thematic structure was finalized, the researcher gathered supporting evidence for each theme and synthesized the findings, which were reviewed by the research team to ensure agreement on the results.

      sentence describing how analysis was performed on data collected by the authors of this paper

    6. Interviews were video and audio recorded. We transcribed the audio using OpenAI's Whisper automatic speech recognition system and anonymized the transcript before analysis. We analyzed the interview data using thematic analysis [1]. First, two members of the research team independently coded four (25% of collected data) randomly chosen participant data to generate low-level codes. The inter-coder reliability between the coders was 0.88 using Krippendorff's alpha [37]. The two coders then met together to cross-check, resolve coding conflicts, and consolidate the codes into a codebook across two sessions. Using the codebook, the two coders analyzed six randomly selected participant data each. The research team then met, discussed the analysis outcomes, and finalized themes over three sessions.

      sentence describing how analysis was performed on data collected by the authors of this paper

    7. Future work could explore more seamless ways of preserving context, such as allowing users to navigate through every sentence of an abstract directly within the Cross-Sentence Relationship pane, fostering a more cohesive understanding of the content.

      any sentence that describes explicit design implications

    8. In this sense, AbstractExplorer enables dialectical activities that users may otherwise have found to be too tedious or difficult to engage with.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    9. In this work, we introduce a new paradigm for exploring a large corpus of small documents by identifying roles at the phrasal and sentence levels, then slice on, reify, group, and/or align the text itself on those roles, with sentences left intact.

      any sentence that describes explicit design implications

    10. Our work demonstrates that designs informed by Structure-Mapping Theory can support users in navigating, making use of, and engaging with variation present in information. In this sense, AbstractExplorer enables dialectical activities that users may otherwise have found to be too tedious or difficult to engage with.

      any sentence that describes explicit design implications

    11. Like prior Structural Mapping Theory (SMT)-informed work in text corpora representation, AbstractExplorer's features have enabled some users to see more of both the overview and the details at the same time, facilitating abstraction without losing context.

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    12. Like prior Structural Mapping Theory (SMT)-informed work in text corpora representation, AbstractExplorer's features have enabled some users to see more of both the overview and the details at the same time, facilitating abstraction without losing context.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    13. We posit that our approach can generalize to other domains such as journalism, code synthesis, and social media analytics where visual alignment of text can enable meaningful comparisons of underlying patterns to identify relational clarity.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    14. In this work, we introduce a new paradigm for exploring a large corpus of small documents by identifying roles at the phrasal and sentence levels, then slice on, reify, group, and/or align the text itself on those roles, with sentences left intact.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    15. Dialectical activities cannot be done on a user's behalf by AI; with variation affordances, AI is supporting the user's engagement with the data themselves.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    16. The ablation and summative studies verified the value of Abstract-Explorer, specifically showing that all three components of the Structural Mapping Engine—color coding, sentence ordering, and vertical alignment—are crucial for facilitating comparative close reading at scale.

      sentence relating to testing

    17. We posit that our approach can generalize to other domains such as journalism, code synthesis, and social media analytics where visual alignment of text can enable meaningful comparisons of underlying patterns to identify relational clarity.

      any sentence that describes explicit design implications

    18. We demonstrate how slicing sentences according to roles and visually aligning them can help readers perceive cross-document relationships in a coherent manner.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    19. Our work demonstrates that designs informed by Structure-Mapping Theory can support users in navigating, making use of, and engaging with variation present in information.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    20. pre-computing and reifying cross-document analogous relationships make it psychologically possible for users to engage—if they are willing to be guided by it. (Lower NFC users are more likely to fall into this category.)

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    21. Activity log data, which revealed how participants actually used the interface, echoed the above findings. According to the log data, participants spent most of their reading time (66.31%) with vertical alignment on the second element in structure pairs, followed by alignment on the first element (29.19%), and left-justified alignment (5.13%). Highlighting usage showed a similar preference: 91.13% of time with all chunks highlighted, 8.25% with partial highlighting, and minimal time (0.63%) without highlights.

      sentence describing how analysis was performed on data collected by the authors of this paper

    22. In this section, we present findings on how AbstractExplorer supports comparative close reading at scale by integrating quantitative survey responses and log data with qualitative analysis of transcripts and open-ended responses. The qualitative analysis process is described in detail in Appendix H.

      sentence describing how analysis was performed on data collected by the authors of this paper

    23. Throughout the two tasks, we also collected detailed interaction logs including counts of user-defined aspects created, duration of highlighting usage, and time allocation across the three possible alignment options.

      sentence describing how analysis was performed on data collected by the authors of this paper

    24. After the ablation study validated the effectiveness of all three SMT-inspired features together (especially for lower NFC users), we completed the implementation of AbstractExplorer and eval-uated its impact on researchers’ reading and sensemaking of a corpus of all ∼1000 paper abstracts from ACM CHI 2024.

      sentence relating to testing

    25. The most popular condition had all three features enabled, i.e., 11 out of 24 participants (≈ 50%) preferred Figure 7C, as shown in the “Preferred” columns of Table 1. The remaining participants were roughly evenly split between the no-features baseline (6 par-ticipants) and the without-alignment ablation condition (5 partic-ipants). One participant each liked the without-highlighting and without-ordering ablation conditions most, respectively.

      sentence relating to testing

    26. Using a two-tailed Mann-Whitney U Test, we found that participants who reported their lowest perceived cognitive load when all three features were enabled had significantly lower NFC than participants who reported their lowest cognitive load level when skimming with no features enabled—in the baseline interface (p=0.03).

      sentence describing how analysis was performed on data collected by the authors of this paper

    27. Both gaze data and the semi-structured interviews revealed that lower NFC participants were more willing to be guided by the three features and took advantage of them consciously.

      sentence describing how analysis was performed on data collected by the authors of this paper

    28. For simplicity of analysis, we denote participants with NFC scores above the overall participants' median NFC of 5.42 (IQR = 0.583) as higher NFC, and lower NFC otherwise.

      sentence describing how analysis was performed on data collected by the authors of this paper

    29. The study concluded with a 15-minute semi-structured interview. During the interview, participants saw screenshots from the three conditions and were asked which they preferred and disliked, why, what they wished the interface had, what influenced their skimming, and how they normally skimmed texts.

      sentence describing any interview procedures

    30. Lower NFC participants were generally guided by emergent visual patterns created by the interactions between features, especially blocks of color spanning multiple sentences created when all three features are turned on.

      statements that draw general conclusions about humans, computers, and/or human-computer interaction based on the results of the specific experiment done in the paper.

    31. The study concluded with a 15-minute semi-structured interview. During the interview, participants saw screenshots from the three conditions and were asked which they preferred and disliked, why, what they wished the interface had, what influenced their skimming, and how they normally skimmed texts.

      sentence relating to testing

    32. The most preferred condition (all three features enabled) was tied with the baseline no-features-enabled condition for lowest reported cognitive load. Specifically, 11 participants reported their lowest raw NASA-TLX scores8 in the all-three-features condition, and a different 11 participants reported their lowest raw NASA-TLX scores in the baseline condition.

      sentence relating to testing

    33. In this study, we allowed participants to experience views of same-aspect sentences (Section 4.1.1) with different combinations of highlighting, ordering, and alignment (as described in Section 4.1.2 and Section 4.1.4) enabled or not, in order to understand which and/or what combinations most effectively supported users' ability to skim and read laterally across documents.

      sentence relating to methodology

    34. We collected 80 sentences from our abstracts dataset labeled by our system as "Methodology/Contribution." Participants viewed the same 80 sentences in each condition—often with a different subset of sentences initially visible due to ordering changes—but only had two minutes to look at them in each condition.

      sentence describing how analysis was performed on data collected by the authors of this paper

    35. The specific research questions for this study were: (1) How do highlighting, alignment, and ordering affect reading patterns, user experience, and cognitive load? (2) How do participants’ valuation of these features relate to their Need for Cognition? (3) Does each feature provide value on its own, or only in conjunction with one or more of the other two features?

      sentence relating to testing

    36. To contrast participants' gaze patterns in each condition, we used a Tobii Pro Spark eye-tracker placed below the desktop monitor used by all subjects; Tobii Pro Lab software recorded each participant's gaze over time in each condition.

      sentence describing how analysis was performed on data collected by the authors of this paper

    37. In this study, we allowed participants to experience views of same-aspect sentences (Section 4.1.1) with different combinations of high-lighting, ordering, and alignment (as described in Section 4.1.2 and Section 4.1.4) enabled or not, in order to understand which and/or what combinations most effectively supported users’ ability to skim and read laterally across documents.

      sentence relating to testing

    38. Structural mappings between objects are part of the cognitive process of comparison according to the Structure-Mapping Theory [17], and juxtaposition can facilitate humans in recognizing particular possible structural mappings between objects [75].

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    39. Inspired by GP-TSM [24], AbstractExplorer first segments sentences into grammar-preserving chunks—segments that respect grammatical boundaries, i.e., an LLM judges that the sentence can be truncated at that chunk boundary without breaking the grammatical integrity of the preceding text. Each chunk is then classified by an LLM as having one of nine pre-defined roles, each of which has its own assigned color.

      sentence relating to methodology

    40. We consider common sequences of chunk roles to be alignable structures that could be used to support users in identifying structural similarities and differences across sentences in different abstracts, in line with Structure-Mapping Theory [17].

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    41. This ordering prioritizes dominant structural patterns (largest groups first) while exposing fine-grained variations (via length-sorted triplets), mirroring how humans compare sentences, if SMT is an accurate description in this domain of comparative close reading.

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    42. In SMT terminology, rendering and arranging according to corresponding chunks reify "commonalities in structure," while variation within corresponding chunks are "alignable differences" that users are predicted to notice.

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    43. SMT posits that visual alignment helps people perceive relational similarities and differences more clearly, thereby improving their ability to make meaningful comparisons and understand underlying patterns [28, 38, 47].

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    44. In the first part of the session, we asked participants about their strategies for selecting publication venues for their manuscript submissions, how they identify and synthesize information from venues, their approaches to writing manuscripts, and finally, the technology they have used to help with these processes, current technology shortcomings, and ideas for addressing these challenges.

      sentence describing any interview procedures

    45. In order to determine (1) the context in which we might offer novel views of scientific abstracts and (2) the intelligibility of various novel prototype designs for reifying cross-abstract relationships, we conducted a formative interview study with 12 active researchers (see Appendix A for participant information).

      sentence describing any interview procedures

    46. We used these mock-ups as design probes [31] to inspire ideation and elicit creative responses. Specifically, we asked participants to compare and contrast alternative mock-ups and reflect on how they could be used or improved to support their known or emerging synthesis and information-foraging goals.

      sentence describing any interview procedures

    47. The prior SMT-informed tools in Section 2.3 for both code and natural language corpora suggest that the cognitive process of comparing texts may be no exception to the cognitive processes SMT predicts.

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    48. The interview sessions were divided into two parts: an open-ended semi-structured interview about their backgrounds and practices, followed by feedback on a range of mock-ups, including novel reified relationships between analogous sentences in different abstracts (Figure 2).

      sentence describing any interview procedures

    49. Structural Mapping Theory (SMT) is a long-standing well-vetted theory from Cognitive Science that describes how humans attend to and try to compare objects by finding mental representations of them that can be structurally mapped to each other (analogies).

      sentence related to any theory

    50. These examples of text-centric lossless techniques do not abstract away or summarize; they strategically re-organize and re-render the existing text to help enhance readers' own perceptual cognition, informed by Structural Mapping Theory (SMT) [17].

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    51. A summative study (N=16) describes how these features support users in familiarizing themselves with a corpus of paper abstracts from a single large conference with over 1000 papers.

      sentence relating to testing

    52. The human perceptual, comparative mental machinery that SMT describes is part of what enables humans to form more abstract structured mental models from concrete examples, among other critical knowledge tasks.

      sentences that mention theory, explicitly or implicitly; one sentence at a time

    53. AbstractExplorer instantiates new minimally lossy2 SMT-informed techniques for skimming, reading, and reasoning about a corpus of similarly structured short documents: phrase-level role classification that drives sentence ordering, highlighting, and spatial alignment.

      sentence related to any theory

  2. Jun 2026
    1. Since histories of specific notations tends to miss detailed, direct observations around the initial creation process, we complement this "macro" analysis with occasional references to experiment-based literature from experimental semiotics, communication theory, and cognitive science into how people use notations to ground communication, largely in lab studies.
    2. we conducted a comparative historical analysis of the development of different notations which individually have been documented in prior literature. Specifically, we conduct a parallel comparative history which "seek[s] above all to demonstrate that a theory similarly holds good from case to case... [and where] differences among the cases are primarily contextual particularities against which to highlight the generality of the [theorized] processes"
    3. Given that user interfaces are a form of notation, it is not surprising that their evolution closely follows the patterns we identified.

      One or more sentences contextualizing the current work with typically uncited statements about the past.

    4. In this work, we lay some groundwork towards addressing thisquestion by presenting a parallel comparative historical analy-sis [ 129] of notation development across scientific, computing,and artistic disciplines.

      testing

    5. From our analysis, we derive a set of initial implications for the design of future systems that create new abstractions (Section 5), including that notations primarily originate through linking metaphors and most often in a social—rather than a technical—context, and that notation design decisions around what to include as "meaningful" (and thus what to exclude) are often left implicit by inventors, but could be made explicit and become manipulable objects through reification [10].
    6. Our work contributes to a longstanding dream of dynamic abstractions in HCI, where users can dynamically communicate and express themselves through notations (interfaces) that they are most comfortable with at the moment of expression, beyond ones predefined by developers [96, 143, 144, 148, 149].
    7. Our historical analysis suggests that, cognitively and socially, a notation proceeds by: (1) Enumerating dimensions of meaningful variation in the target domain, which proliferate as more situations are encountered or considered (whether by inventors or users) (2) Mapping dimensions of meaningful variation to perceptual channels of representation (3) Designing the notation to leverage perceptual affordances by visual analogy to embodied transformations like pouring cups or rotating shapes, and ensuring these "natural" manipulations hold meaning in the target domain
    8. Our work heeds calls from HCI scholars for more historicism in the field, to better contextualize technology "within dynamic temporal processes of emergence, change, continuity, decline, disappearance, or revival," and to see "the past... as a repository of design knowledge and experience"
    9. From our analysis, we derive a set of initial implications for the design of future systems that create new abstractions (Section 5), including that notations primarily originate through linking metaphors and most often in a social—rather than a technical—context, and that notation design decisions around what to include as "meaningful" (and thus what to exclude) are often left implicit by inventors, but
    10. Our analysis identifies 33 patterns of how notations are created, evolved, and formalized over time, which are largely shared across histories and loosely categorized into three social stages of development (invention/incubation, dispersion/divergence, and institutionalization/sanctification) and three functional stages (descriptive, generative, and evaluative).
    11. Yet humans collaborating together do not always defer to a pre-existing formalism. Instead, they can develop ad-hoc, new notations to ground their communication [22, 34, 143], treating notations more like malleable resources than rigid systems.
    1. SDT broadly differentiates three types of motivation [157]: Intrinsic motivation denotes activity pursued for its inherently interesting or enjoyable qualities. Extrinsic motivation refers to activity pursued for a separable outcome. Amotivation denotes the absence of intentional motivation, where a person may no longer be aware why they pursue an activity.
    2. Basic psychological needs theory (BPNT) posits three basic psychological needs that energise organismic processes: competence, the feeling of having an effect; autonomy, a sense that actions are self-endorsed and performed willingly; and relatedness, a sense of reciprocal care, value, and belonging in relation to other social figures and collectives [158].
    1. To our knowledge, the first SDT research involving videogames [18] was conducted shortly after Deci's original formulation of CET [129] and investigated whether extrinsic rewards would reduce intrinsic motivation even for 'highly intrinsically motivating' activities such as videogame play. Videogames' intrinsically motivating qualities were also examined in early research on learning [e.g., 351]; however, focused examination of other core SDT concepts such as need satisfaction largely began much later [365].
    2. Research on games and play in HCI (henceforth HCI games research), however, has continued to employ broad psychological theories as foundational work [417, 556]. One prominent example can be seen in self-determination theory (SDT) [481, 483], an influential theory of human motivation, which has provided HCI games research with propositions and concepts that can help explain motivational and experiential qualities of games and game-adjacent systems (e.g., gamification).
    3. Psychological concepts and models have long been employed in human–computer interaction (HCI) to theorise the human user [88]. However, early applications of cognitive psychological theory did not develop into a coherent foundation of knowledge about human factors [89, 109, 455]—circumstances that Rogers [456, p. 22] attribute to "the stark differences between a controlled lab setting and the messy real world setting" for which interactive artefacts and systems are designed. The deployment of broad theory in HCI has subsequently declined in the intervening years [455, 456], and this sporadic progress in theory development in domains such as usability and user experience (UX) has been identified as a cause for concern [249, 314].
    1. A father tells his daughter, “It’s time to put the tablet down.” Not wanting to stop using the tablet, butworried about the consequences of disobedience, the child finds herself in a dilemma. With a strokeof insight, she puts the tablet down on the table in front of her, and keeps playing with it. She can stilluse the tablet, and her father’s instructions were met. Technically.
    1. Alignment is a bilateral process; it refers not only to AI acting according to human intentions but also to humans better leveraging AI by understanding the mechanisms behind it [54].

      Any individual sentence that describes information designed to set the stage for the contribution of the paper.

    2. Data labeling as a cognitive task—including defining a concept or determining how two similar objects may have different labels—requires both comparison and integration [62].

      Any individual sentence that describes information designed to set the stage for the contribution of the paper.

    3. However, relying exclusively on existing examples is not ideal for tasks requiring nuanced understanding of user intentions, as these examples often fail to represent diverse and edge-case scenarios [31].

      Any individual sentence that describes information designed to set the stage for the contribution of the paper.

    4. An important challenge in interactive machine learning, particularly in subjective or ambiguous domains, is fostering bi-directional alignment between humans and models.

      Any individual sentence that describes information designed to set the stage for the contribution of the paper.

    5. Machine teaching, a part of the human-in-the-loop approach, has been used as a process in which a human expert (the "teacher") provides guidance to a machine learning model to help it learn important and robust features for decision making [57].

      An individual sentence describing the setting in which this work was done.

    6. A targeted approach in IML is machine teaching (MT) [60], an interactive framework that allows users to devise and select useful data for labeling, with the goal of teaching the model relevant features during training [7, 18].

      An individual sentence describing the setting in which this work was done.

    7. Interactive ML (IML) methods, like active learning [3], continuously apply human feedback during model training to iteratively build and refine the model [35, 42, 43].

      An individual sentence describing the setting in which this work was done.

    1. This meet-up invites CHI attendees to come together inthe META HCI community to explore how we as researchers andpractitioners reflect on our own practices – and how we might doso more intentionally.
    2. Importantly, reflection happens on multiple levels: as individuals questioning assumptions and choices, as groups working together in projects or labs, and as a community negotiating shared values, norms, and directions.

      sentences that describe the concept/practice of reflection

    3. Structures — reflection on the structures that condition HCI and our own standings within them: societal constructs (positions, values, power) shaping what problems are visible and whose knowledge is legitimised.

      sentences that describe the concept/practice of reflection

    4. Reflection has been a recurring theme in HCI – from Schön's reflective practitioner [24] to Sengers et al.'s reflective design [25]. However, it is seldom centred in our collective conversations [2].

      sentences that describe the concept/practice of reflection

    1. As the community around augmented reading broadens and as possibilities continue to unfold, it is the purpose of this workshop to set up our community to drive innovation in a productive, desirable, and responsible way.

      Sentence that describes the setting in which the paper's contribution is relevant or intended.

    2. The landscape of technology for consuming information is changing rapidly. One mode of information consumption, reading, stands to see profound changes due to its ubiquity and frequency as a cognitive task.

      Sentence that describes the setting in which the paper's contribution is relevant or intended.

    3. Recent changes in the technological landscape are significantly changing the reading experience. AI has introduced many new possibilities for interfaces to augment or transform text to be more rapidly scanned, navigated, understood, and compared to other texts.

      Sentence that describes the setting in which the paper's contribution is relevant or intended.

  3. May 2026
    1. But the ear-lier we go in development, the less able children are to comprehend verbal explanationsof abstract ideas. In contrast, there is evidence that analogical comparison and abstractionprocesses are present in 7–9-month-old infants, and even earlier (Anderson, Chang, Hes-pos, & Gentner, under review; Ferry, Hespos, & Gentner, 2015).
    1. In HCI, evaluation refers to the application of some systematic methodology to attribute human-related values to an artifact, prototype, system, or process. Examples of such attributes include performance, experience, safety, and ethical aspects, such as the avoidance of bias or harm.