Showing posts with label cognition. Show all posts
Showing posts with label cognition. Show all posts

Thursday, April 30, 2009

ID FMP: Distributed Cognition Models

Distributed cognition models conceptualize cognitive phenomena as happening across multiple individuals, objects, and internal and external representations of knowledge.  In contrast to the Information Processing Model, which is only focused on activities that happen inside the head, this model focuses on internal and external activities and encompasses External Cognitive Processes and Coordination Mechanisms described in my previous posts.

In comparison to these three frameworks, distributed cognition models provide more precise descriptions of internal and external cognitive activities. They are less abstract because their domain is limited to cognitive activities associated to specific contexts (e.g. piloting an airplane, doing taxes).

The three frameworks previously mentioned provide general descriptions of how human cognition works across all contexts. Their focus is on defining general laws that describe how our brain processes information and leverages the external world to enhance our cognitive capabilities. The distributed cognition model offers a phenomenological perspective that explores cognition as an embodied activity that takes place in specific physical and social contexts.

For example, a distributed cognition model that describes the activities that take place at an agency during creative development would differ considerably from that of a law office. They would feature many commonalities but the important thing is that the differences matter.

This perspective is important because designers need to understand how their product or service will actually fit into people’s day-to-day life. The insights that can be gleaned from the Information Processing and External Cognitive Activities Frameworks do not provide this type of understanding.  Distributed cognition models focuses on mapping these mundane day-to-day activities. They provide insight into how people actually make and share meaning and decisions within specific contexts.

A distributed cognition analysis is usually carried out as the basis for development of a distributed cognition model. Here is an overview of the main areas of examination in these types of analysis. As an example (and to work my brain just a little bit) I’ve carried out a high-level analysis of the distributed cognitive activities that take place at an advertising agency.
  • How does distributed problem solving take place? How do people work together to solve problems? In an agency environment, tasks are distributed across several departments with specific areas of expertise (e.g. client services, account & strategic planning, media, production, creative and traffic). People work together by coordinating their actions using documents (such as schedules, briefs, spec sheets and emails), events (such as meetings, phone calls, and presentations), and shared work practices (such as common vocabularies, understandings, and culture).
  • What ways does communication take place throughout the collaborative process and how is knowledge shared and accessed? Does it change as the activity progresses? Communications take place via meetings, emails and document artifacts such as presentations, briefs, schedules, conference reports, creative comps and spec sheets. The most important information is documented to facilitate sharing. Many of the document artifacts evolve as the activities progress. For example, a creative brief may be updated to reflect changes in strategy. The creative comps also change via multiple rounds of client reviews.
  • What is the role of verbal and non-verbal communication? What types of things are said or implied? Verbal communication is the primary type of communication associated to the management of projects (and communication associated to those projects). Non-verbal communication plays a fundamental important in the activities of the project itself. Layout design, videos, images, graphs, and even experiences are be used to brief creative teams regarding products or brands, and in client and internal presentations. The final creative product delivered by Agencies also employs both verbal and non-verbal communication. To elicit emotional responses from people agencies use non-verbal tools such as images, visuals, videos, sounds, interactions online, and more. In agency communication is often reinforced through by verbal and non-verbal communication.
  • What coordinating mechanisms are used? What are the rules and procedures that govern the workflow? There are several important coordination mechanisms that are used in an agency. These mechanism leverage external representations of knowledge such as schedules, job jackets, spec sheets, readers, status reports, conference reports, emails, calendars, scopes of work, etc. They also include meetings such as internal and client reviews, status meetings, and production kick-offs. Many rules and procedures are outlined in the agency’s process manual. These processes govern how work flows through the agency.
[source: Interaction Design: Beyond Human-Computer Interaction, page 129.]

** What the hell is ID FMP? **

Sunday, April 26, 2009

ID FMP: Coordination Mechanisms

Coordination is an important skill that is required to carry out activities that range from basic to complex. All collaborative activities heavily rely on the ability of individuals to coordinate their actions; all group activities require some level of coordination. Even personal activities often require coordination such as prioritization and scheduling.

So what is coordination? Coordination is the “the regulation of diverse elements into an integrated and harmonious operation” (wordnet.princeton.edu/perl/webwn). In other words, coordination refers to the phenomenon where one or more people act or interact with to accomplish a goal or complete a task. Much of the thinking related to coordination focuses on group, rather than personal, activities.

Sharp, Rogers, and Preece have identified three different types of coordination mechanisms that people use to coordinate their actions with others. I’ve modified their framework by adding one additional type of mechanism; I decided to break down their second category into two separate entities. As you will note, these coordination mechanisms are interdependent and overlapping.
  • Conventions and shared practices: Conventions and shared practices refer to the shared social and cultural understandings and beliefs that provide a foundation for coordination. Examples include cultural expectations about punctuality, shared understandings regarding meaning of activities or artifacts. These phenomena account for why it can often be harder to coordinate activities with people from different socio-cultural backgrounds. Shared conventions and practices along with verbal and non-verbal communication play a key role in enabling people to effectively use schedules, rules and shared external representations to coordinate activities.
  • Verbal and non-verbal communication: spoken and written language, and non-verbal gestures are often used as primary means of communication for the coordination of activities. Conversations are an important medium for the coordination of activities and negotiation of commitments. Written documents, such as agendas, presentations and reports, are also common tools for coordinating groups. Gestures play an especially important role in supporting the coordination of activities in situations where the conditions do not allow for users to communicate using verbal communication; examples include, a catcher using hand signs to communicate with a pitcher and a conductor using the motions of his arm and baton to lead an entire orchestra. Gestures can also help support communication between people who do not share the same language.
  • Schedules, maps and rules: Schedules, maps and rules are artifacts that document communications that outline the order of activities, conventions and shared practices. Schedules focus on organizing activities and objects across time while maps organize activities and objects across space – both are crucial tools for personal and group coordination. Rules offer descriptions of conventions, shared practices and other principles that facilitate the coordination of activities. The benefit of rules and schedules is that they enable groups of people with different practices and conventions to create a shared set of documented principles to guide their coordination and collaboration.
  • Shared external representations: Shared external representations are schedules, rules and other forms of visual or physical artifacts that are shared by a group of people. Examples vary widely across industries; in agencies like the one where I currently work, a job jacket and router is used to provide information regarding who has reviewed and commented on a given project during each round of its development. Shared online calendars, such as google calendar, offer the ability to share schedules and create shared external representations in a virtual, as opposed to physical, environment.
Activities associated to coordination are directly supported by the Cognitive Processes and External Cognitive Activities Frameworks. The mechanisms for coordination with groups encompass rely on the processes and activities outlined in these models.

Conventions and shared practices reside in the mind and are largely governed by cognitive processes such as memory, learning, and higher reasoning. While cognitive processes associated to language enable us to use verbal communication and plays an important role in our ability to create and understand external representations.

Externalizing cognitive activities is a crucial element most types of coordination mechanisms. Memory offloading is a crucial benefit provided by schedules, maps, rules, and external representations. Computational offloading is often employed using verbal communications and shared external representations. Annotating and cognitive tracing is mostly used on schedules, maps, and shared external representations.

[source: Interaction Design: Beyond Human-Computer Interaction.]

** What the hell is ID FMP? **

Wednesday, April 22, 2009

Chapter 5 Homework: What is Interaction Design

This assignment was taken from the fifth chapter of the book Interaction Design: Beyond Human-Computer Interactions, written by Helen Sharp, Jenny Preece, and Yvonne Rogers.

Assignment Questions
This assignment requires you to write a critique of the persuasive impact of a virtual agent by considering what it would take for a virtual agent to be believable, trustworthy, and convincing.

Question A: Look at a website that has a virtual assistant, e.g. Anna at Ikea or one of the case studies featured by the Digital Animations Group (DAG) at http://www.dagroupplc.com, who specialize in developing a variety of online agents, and answer the following:
  • What does the virtual agent do?
  • What type of agent is it?
  • Does it elicit an emotional response from you? If so, what kind?
  • What kind of personality does it have?
  • How is this expressed?
  • What kinds of behavior does it exhibit?
  • What are its facial expressions like?
  • What is its appearance like? Is it realistic or cartoon-like?
  • Where does it appear on the screen?
  • How does it communicate with the user (text or speech)?
  • Is the level of discourse patronizing or at right level?
  • Is the agent helpful in guiding the user towards making a purchase or finding out something?
  • Is it too pushy?
  • What gender is it? Do you think this makes sense?
  • Would you trust the agent to the extent that you would be happy to buy a product from it or follow it guidance? If not, why not?
  • What else would it take to make the agent persuasive?
Question B: Next look at an equivalent website that does not include an agent but is based on a conceptual model of browsing, e.g. Amazon.com. How does it compare with the agent-based site you have just looked at?
  • Is it easy to find information?
  • What kind of mechanism does the site use to make recommendations and guide the user in making a purchase or finding out information?
  • Is any kind of personalization used at the interface to make the user feel welcome or special?
  • Would the site be improved by having an agent? Explain your reasons either way.
Question C: Finally, discuss which site you would trust most and give your reasons for this.

Assignment Answers

Question A
Site selected: ikea.com 

What does the virtual agent do?
The virtual agent inhabits a pop-up window and is comprised of an avatar of a young blond woman who blinks and moves here head. The interface is primarily text-based, both input and output are provided in this format. The output can be enhanced with audio that sounds computer generated.

The primary function of the virtual agent is to provide help to visitors on the Ikea website. This help encompasses supporting users in all aspect of their shopping experience (it provides essentially a new interface for users to interact with the site). The agent provides support by enabling users to search for answers to common customer service queries using natural-language questions. These questions are posed through a text box. The response is provided via text and, optionally, audio (audio is available on the UK site but no on the US site). When appropriate the agent will load a relevant page on the main screen of the browser.

What type of agent is it?
The agent is a customer service representative. It is a friendly female avatar that offers a stylized representation of a human female that does not attempt to provide a realistic image of a female Ikea employee.

Does it elicit an emotional response from you? If so, what kind?
I must be upfront about my general dislike for avatar-based interfaces, with the notable exception of videogames. I often feel as though I am being patronized when I interact with an agent on a website, the Ikea agent was no exception. One of the few online agents that I found successful was Ms. Dewey [http://en.wikipedia.org/wiki/Ms._Dewey], a search-engine prototype developed by Microsoft. I can understand why it did not scale but it was pretty damn cool.

What kind of personality does it have? How is this expressed? What are its facial expressions like?
The agent has a friendly and relaxed personality. This is expressed through her facial expressions and the movement of her head. The agent is smiling all the while she opens and closes her mouth. Her large eyes blink at a natural while pace while she moves her head from side to side in a relaxed manner.

Where does it appear on the screen? What is its appearance like? Is it realistic or cartoon-like? What kinds of behavior does it exhibit?
The agent is situated in a pop-up window. Its appearance is stylized and cartoon-like. Her behavior seems for the most part fluid and natural until she responds with audio and her lips do not move. The computer-generated voice that is used only detracts from the experience because it is cold and is neither cartoon-like nor human sounding.

How does it communicate with the user (text or speech)?
The agent accepts questions via text input and is able to provide response via text and audio output.

Is the agent helpful in guiding the user towards making a purchase or finding out something? Is the level of discourse patronizing or at right level? Is it too pushy?
The Ikea agent can be helpful in guiding users towards making a purchase, or finding a product or retail location. One of the strongest features of the Ikea agent was its ability to load content that is relevant to the user’s query onto the main browser window. For example, when I searched computer desk it took me to the Ikea website’s computer solutions category.

Though I find agent-based interfaces patronizing in general, this one is much less so than most. The agent provides straightforward and short answers coupled with additional information on the main browser window. I actually found this agent to be useful, a fact that helped me overcome my initial aversion to this type of interface.

What gender is it? Do you think this makes sense?
The agent is a female. I think this makes sense largely based on my assumption that Ikea online shoppers are mostly women. I suspect that most men also prefer to deal with a female agent – especially since even the shiest guy would not be intimidated by an online agent. In the US there is a tradition of portraying customer service representatives as friendly females with a girl next door look.

Would you trust the agent to the extent that you would be happy to buy a product from it or follow it guidance? If not, why not?
I would trust the Ikea agent because she is informative, helpful and non-intrusive – she never initiates interaction with the user. The Ikea agent helps shoppers to find things and get answers to frequently asked questions regarding store and website policies.

What else would it take to make the agent persuasive?
Though I did find the agent useful, there are several things that can be done to improve its persuasiveness: improve interaction and visual design; enhance functionality; and upgrade audio interface.

Improve interaction and visual design: from an interaction standpoint the conversation with the agent should be recorded in a manner that enables the shopper to scan the queries and responses in search of answers (or a new chair). The look and feel of the agent should be upgraded to better reflect the design sense of the Ikea brand. Additional details should be added to enhance the enjoyment of users (e.g. have the rep read a book while she is waiting for the user). Since many shoppers like to go back and forth when they shop, the agent should help the user find products that they’ve looked at during their visit to the website.

Enhance functionality: additional functionality that could enhance the agent’s usefulness includes the ability to provide tips regarding other Ikea products that match pieces of furniture being viewed by the shopper. These recommendations should be provided in a non-intrusive manner.

Upgrade audio: The last thing that I would change is to upgrade the audio quality. This was one feature that I found to be very poor. Currently, the agent “speaks” in a computer-generated voice with a slight British accent. For the US version of the agent they should consider adding sound functionality, as if it is done right it can add to the user’s interactions with the agent.

Question B
Site selected: cb2.com (US furniture retailer akin to Ikea)

Is it easy to find information?
The cb2 website is pretty well organized, which makes it easy for the user to find information. Aside from the standard categorization of products by furniture type and context, they also provide lists of new and most popular products. These elements of the site help people find products through browsing. The search feature provides users with a way to shortcut the browsing process in an attempt to find a more direct route to the information they seek.

What kind of mechanism does the site use to make recommendations and guide the user in making a purchase or finding out information?
The CB2 site actual does a better job at making recommendations, though it is only equally effective at guiding users to find information regarding products, and features less compelling interactive guides. From a recommendation standpoint, the CB2 site provides shoppers with tips on other products that work with any piece that is being viewed. Though both sites differ in the way they categorize their product offerings, from a findability standpoint both the CB2 site and the Ikea site (including the agent and general information architecture) are equally effective.

Is any kind of personalization used at the interface to make the user feel welcome or special?
The CB2 site does not offer any personalization. Shopper’s are not asked to register and log-in during their visits to access special recommendations or offers. The Ikea site does provide a log-in feature, however, it has been down since I have been working on this assignment.

Would the site be improved by having an agent? Explain your reasons either way.
I don’t think an agent would have a big impact on the experience at CB2. The reason being, content on the site was easy to browse and find without the help of an agent. I believe that an agent would only improve the experience of a very small segment of the shoppers on the site. If voice-based interaction becomes more common on computers then there would be value in adding an agent to the CB2 experience. This is not an unlikely phenomenon considering that many applications now-a-days are striving to become voice-enabled to facilitate use via mobile phones (check out the new google search on iPhone and Android, cool stuff).

Question C

Finally, discuss which site you would trust most and give your reasons for this.
Both experiences were on par for one main reason: on the Ikea website the agent provides users with a supplementary interface that does not replace the traditional browsing paradigm on which the rest of the site is built. The Ikea and CB2 websites both provide well-designed information architectures that make information and products easy to find. Also, both companies have strong and respected brands that stand for modern and affordable design. I guess I have officially copped out of answering this question.

** What the hell is ID-BOOK ? **

Sunday, April 12, 2009

ID FMP: Cognitive Model of Emotions for Design

In the 90’s Don Norman, along with his colleagues Andrew Ortony and William Revelle, began to explore the role that emotions play in making products easier, more pleasurable, and effective to use. During this time period various design researchers were investigating the link between emotions (especially aesthetics) and usability. These inquiries confirmed that our affective states strongly influence our experiences, and that these states can be induced by design of product systems.


The model they developed aims to explain how different levels of our brain govern our emotions and behaviors. At the visceral level our brain is pre-wired to rapidly respond to events in the physical world by triggering physiological responses in our body. The behavioral level controls our everyday behaviors, including learned routines such as walking and talking. Lastly, the reflective level is responsible for cognitive processes related to contemplation and planning.

Emotions can arise at various levels and are created by a combination of physiological and behavioral responses that are influenced by reflective cognitive processes. An emotion like anger tends to be mostly visceral or behavioral in nature. However, indignation, which is a higher-level version of this emotion, is reflective in nature as well.

The main implication from this model is that our affective states have an impact on how we think. This important insight applies to thinking about the user’s affective state when using the product, and to how a user’s affective state will be impacted by use of the product. In regards to the former consideration, designers can take leverage an understanding regarding common physiological and emotional responses to stressful situations in order to design products that can be successfully used in such contexts.

A users’ experience with a product itself can also have impact on their affective states. High- and low-level emotions can influence all levels of cognitive activity, which is why a one’s visceral response to a product’s aesthetics can impact our behavior. On the other hand, the one’s higher cognitive functions control one’s lower level functions, which is why we can overcome our initial emotional responses if a product is effective enough.

The most common way that designers apply this model is by exploring the design considerations associated to each of the three levels. Visceral design encompasses considerations such as the aesthetics of the look, feel, smell and sound of the product. Behavioral design refers to considerations associated to the product’s usability. Reflective design is concerned with the meaning and value that a product provides within the context of a specific culture.

I consider this model to be an evolution of modes of cognition framework. The main change is that in the Emotional Model the “experiential” mode of cognition has been divided into distinct types of cognition: visceral and behavior. This revision enables the model to reflect the important role played by our emotional response to a product’s aesthetics.

[source: Interaction Design: Beyond Human-Computer Interaction; Don Norman’s book Emotional Design.]

** What the hell is ID FMP? **

Sunday, March 29, 2009

ID FMP: External Cognitive Activities

People often leverage artifacts and characteristics from their environment to reduce cognitive load and enhance their cognitive capabilities. External cognition refers to the activities that people use to support their cognitive efforts. These activities rely on: a wide range of artifacts such as computers, watches, pens and papers; characteristics of the environment such as visible landmarks, and signs; and other people. There are three main types of external cognition activities.

These three types of activities are heavily inter-dependent. In the diagram above they are listed from broadest to most specific. The externalization of memory load is the most basic external cognitive activity. It is involved in all types of external cognitive activities.

Computational offloading leverages memory externalization for the specific purpose of performing computational tasks. It is the next most basic external cognitive activity.

Annotation and cognitive tracing can be used to support both types of distributed cognitive activities mentioned above. This type of distributed cognition involves the manipulation or modification of memory and computational externalizations that impact the meaning of the externalizations themselves.

External cognitive activities are used to support experiential and reflective modes of cognition [more info on cognitive modes]. These activities rely on and support all types cognitive processes defined in my earlier post – attention, perception, memory, language, learning, and higher reasoning [more info on cognitive process types].

This framework of external cognitive activities complements the Information Processing model by identifying how people leverage their external environment to enhance and support their cognitive capabilities [more info on information processing model].

It also complements the model of interaction by providing additional insights regarding how people interact with the world (or system images) to support and enhance their cognitive capabilities. However, it does not provide insight into how people interact with systems for non-cognitive pursuits, such as physical and communication ones [more info on model of interaction].

[source: Interaction Design: Beyond Human-Computer Interaction]

** What the hell is ID FMP? **

Friday, March 27, 2009

ID FMP: Information Processing Cognitive Model

One of the most prevalent metaphors used in cognitive psychology compares the mind to an information processor. According to this perspective, information enters the mind and is processed through four linear stages that enable users to choose an appropriate response.


Though this model offers insights into how people process information, it is limited by its exclusive focus on activities that happen in the mind. Most of our cognitive activities involve interactions with people, objects, and other aspects of the environment around us. In other words, cognition does not take place only in the mind.

In my next ID FMP post I will cover the external cognition framework that describes external cognitive activities; and distributed cognition models that attempt to map all internal and external activities. Here’s how this model aligns to the frameworks, models, and principles that I have explored over the past several weeks.

The cognitive activities modeled by Information Processing framework above can be mapped to the mental activities outlined in Norman’s Model of Interaction. At a high level, Norman’s model provides additional insights regarding the mental activities that take place and it features the external environment as an important, though unexplored, element. Here is a brief overview of how the phases from this model relates to the interaction one:
  • “Input encoding” maps to “perception”
  • “comparison” encompasses “interpretation” and “evaluation”
  • “response selection” corresponds to “intention” and “action specification”
  • “response execution” maps to “execution
The “goals” phase of the Interaction Model crosses over between the “comparison” and “response selection” phases of the Information Processing framework – so do the additional phases of the modified Model of Interaction.

Here is how this model aligns with the framework regarding the relationship between a designer’s conceptual model and a user’s mental model. The focus of the Information Processing model is on the cognitive processes that occur in the user’s mind when they are interacting with the world. These processes are closely related to mental models in two ways:
  • First, mental models provide the foundation for people to understand their interactions with the world and select appropriate responses.
  • Second, mental models evolve as people evaluate the impact of their own actions and other events on the world.
[Note: by “world” I refer to any physical, virtual and social entities with which people can interact.]

The Conversation Turn Taking Model is related to the Information processing model in a broad sense only. The turn taking framework focuses on explaining an external phenomenon related to language and communication that is driven by the cognitive functions described in the Information Processing model. They do not contradict one another nor do they directly support each other.

The Information Processing model can be applied to both reflective and experiential modes of cognition, though the phases involved in each mode differ. Reflective cognition tends to be active during the “comparison” and “action selection” phases. On the other hand, experiential cognition can be active across all phases depending on the type of interaction.

The chart below provides an overview regarding which cognitive process types are involved with each phase of the Information Processing model.


[source: Interaction Design: Beyond Human-Computer Interaction]

** What the hell is ID FMP? **

Sunday, March 22, 2009

ID FMP: Model of Interaction

There are many theories that attempt to describe the cognitive processes that govern users’ interactions with products systems. Here I will focus on a model developed by Don Norman, which was outlined in his book Design of Everyday Things. This framework breaks down the process of interaction between a human and a product into seven distinct phases.

Seven Phases of Interaction with a Product System
  1. Forming the goal
  2. Forming the intention
  3. Specifying an action
  4. Executing the action
  5. Perceiving the state of the world
  6. Interpreting the state of the world
  7. Evaluating the outcome
Two important concepts related to Norman’s Theory of Action are the gulf of execution and evaluation. The gulf of execution refers to the gap between how the user wants to act and how the system allows the user to take action. The gulf of evaluation corresponds to the gap between how the system displays data to how the user interprets this data into knowledge.

Now let’s put this theory into context with some of the concepts and models that we’ve encountered thus far. First, I want to point out that this model aligns with Don Norman’s model regarding the relationship between a designer’s conceptual model and a user’s mental model [read more here]. The focus of this framework is the interaction between the system image, the product’s interface where user interaction happens, and the user’s mental model, the user’s understanding of how the product works which governs the user’s interpretation, evaluation, goals, intention, action specification.

I’ve extended Norman’s original model to account for the reflective cognition that is also involved in peoples’ interactions with products. Reflective cognition governs peoples’ higher-level evaluations, goals and intentions that ultimately drive peoples’ experiential cognition activities. Experiential cognition governs the second-by-second evaluations, goals, and intentions involved in peoples’ interactions with products. These two different modes of cognition are explored in greater detail here.

Here is an example to distinguish and highlight the interdependencies between these two different types of cognition and interaction. Let’s consider a person’s interaction with a car. In this scenario, a person’s reflective cognitive would include setting a goal such as choosing a destination and desired time of arrival, and evaluating what route to take based on understanding of current location and traffic patterns. These activities would govern a person’s experiential interactions with a car and drive their moment-by-moment evaluations, and creation of goals and intentions. Experiential interactions would include using the steering wheel to turn a corner or switch lanes, pressing the accelerator to speed up, and stepping on the breaks to stop the car.

How does the concept of mental models relate to this framework? The mental model itself is not represented by a single phase, or grouping of phases. It refers to the understanding that a user has of how a system works. Norman’s model was developed to describe how users interact with product systems on an experiential, minute-by-minute basis. At this level of interaction a user’s mental model drives their interpretations, evaluations, setting of goals and intentions, and specification of actions.

Now let’s explore how the different cognitive types come into play during the various phases of interaction. These cognitive types have been outlined in greater detail here.
  • Attention supports all phases of interaction from perception through to action execution. This cognitive process refers to a user’s ability to focus on both external phenomena and internal thoughts.
  • Perception is clearly called out as its own phase in Don Norman’s model.
  • Memory plays an important role during all phases from interpretation through to action specification.
  • Language supports communication throughout all phases of a person’s interaction with a product. Here I refer to both verbal and visual languages.
  • Learning enables people to use new products and increase effectiveness and efficiency in their interactions with existing products. This cognitive process supports all phases between the interpretation and action specification.
  • Higher reason governs all activities related to the setting of high-level goals and intent, and driving evaluations.
[source: Interaction Design: Beyond Human-Computer Interaction; and Don Norman's Book The Design of Everyday Things]

** What the hell is ID FMP? **

Saturday, March 21, 2009

ID FMP: Conversation Turn-Taking Model

Holding a conversation is a basic human activity. It requires a large amount of coordination between participants, a fact that is often unnoticed. People need to know when to listen, when they can start talking, and when to cede the floor. Conversation mechanisms facilitate the coordination of conversations by helping people know how and when to start and stop speaking. These mechanisms enable people to effectively negotiate the turn-taking required carry out a conversation.

Harvey Sacks, Emanuel Schegloff, and Gail Jefferson have developed a model that aims to explain how people manage turn taking during conversations. The focus of their research was to create a framework that can be applied across cultures and contexts, and that can accommodate several key observations about the structure and dynamics of conversations. Here is an excerpt from the abstract of their paper The Simplest Systematics for the Organization of Turn-Taking for Conversation.

“The organization of taking turns to talk is fundamental to conversation, as well as to other speech-exchange systems. A model for the turn-taking organization for conversation is proposed, and is examined for its compatibility with a list of grossly observable facets about conversation [outlined below].”

The Foundations
Before we explore the model itself let’s take a look at its foundation. Here is a list of the “grossly observable facets about conversation” that was referred to in the quote above:
  1. Speaker changes will always occur and often recur.
  2. For most of the time only one party talks at a time.
  3. More than one person will often talk at a time, but these occurrences are brief.
  4. Most transitions occur with no gap or overlap, or with slight gap or overlap.
  5. Turn order varies throughout conversation.
  6. Turn size or length usually varies.
  7. Length of conversation is not specified.
  8. What parties say is not specified.
  9. Relative distribution of turns is not specified.
  10. Number of parties varies considerably.
  11. Talk can be continuous or not.
  12. Turn-allocation techniques are used to facilitate the conversation.
  13. Sometime turn-constructional units are used to facilitate conversation.
  14. Repair mechanisms exist for correcting turn-taking errors.
The Model
The general model that they developed, which is pictured above, is composed of the three basic rules that govern the transition of turns in a conversation. These rules are:
  1. The current speaker chooses the next speaker by asking a question or making a request.
  2. If the speaker does not choose the next speaker, then another person can self-select to start speaking.
  3. The speaker can decide to continue speaking if no other person self-selects to start speaking.
[source: Interaction Design: Beyond Human-Computer Interaction; and Harvey Sacks, Emanuel Schegloff, and Gail Jefferson’s paper The Simplest Systematics for the Organization of Turn-Taking for Conversation, 1974]

** What the hell is ID FMP? **

Sunday, March 15, 2009

ID FMP: Types of Cognitives Processes

In my last post I identified two different modes of cognition. Here I will continue my investigation into the scope of cognition by identifying six different types of cognitive processes, taken from the book Interaction Design: Beyond Human-Computer Interaction. My focus will remain on the questions: “what is cognition? And what are the main types cognitive activities?”

The six types of cognitive processes that I will describe are attention, perception, memory, language, learning, and higher reasoning. The processes are interdependent and occur simultaneously. They play a role in experiential and reflective modes of cognition. Here is a description of each process along with a few related implications.

Attention: process for selecting an object on which to concentrate. Object can be a physical or abstract one (such as an idea) that resides out in the world or in the mind.

Design implications
: make information visible when it needs attending to; avoid cluttering the interface with too much information.

Perception: process for capturing information from the environment and processing it. Enables people to perceive entities and objects in the world. Involves input from sense organs (such as eyes, ears, nose, mouth, and fingers) and the transformation of this information into perception of entities (such as objects, words, tastes, and ideas).

Design implications
: all representations of actions, events and data (whether visual, graphical, audio, physical, or a combination thereof) should be easily distinguishable by users.

Memory: process for storing, finding, and accessing knowledge. Enables people to recall and recognize entities, and to determine appropriate actions. Involves filtering new information to identify what knowledge should be stored. Context and duration of interaction are two important criteria that function as filters.

Design implications
: do not overload user’s memory; leverage recognition as opposed to recall when possible; provide a variety of different ways for users to encode information digitally.

Language: processes for understanding and communicating through language via reading, writing, speaking, and listening. Though these language-media have much in common, they differ on numerous dimensions including: permanence, scan-ability, cultural roles, use in practice, and cognitive effort requirements

Design implications
: minimize length of speech-based menus; accentuate intonation used in speech-based systems; ensure that font size and type allow for easy reading.

Learning: process for synthesizing new knowledge and know-how. Involves connecting new information and experiences with existing knowledge. Interactivity is an important element in the learning process.

Design implications
: leverage constraints to guide new users; encourage exploration by new users; link abstract concepts to concrete representations to facilitate understanding.

Higher reasoning: processes that involve reflective cognition such as problem-solving, planning, reasoning, decision-making. Most are conscious processes that require discussion, with oneself or others, and the use of artifacts such as books, and maps. Extent to which people can engage in higher reasoning is usually correlated to their level of expertise in a specific domain.

Design implications: make it easy for users with higher levels of expertise to access additional information and functionality to carry out tasks more efficiently and effectively.

[source: Interaction Design: Beyond Human-Computer Interaction]

** What the hell is ID FMP? **

ID FMP: Modes of Cognition

Cognition [define] encompasses a wide range of processes related to thinking, sensing, interpreting, evaluating, decision-making, remembering and communicating. It is important for designers to understand human cognition processes in order to design systems that are easy to learn, effective, efficient, pleasurable, and meaningful.

Here, I will first distinguish between two main modes of cognition. In my next post I will identify different categories of cognitive processes. The value of these distinctions is that different modes and types of cognition call for different technology and interaction solutions. It is important to note that both cognitive modes and multiple processes are always active simultaneously.

The focus of this, and my next, post is to explore the scope of cognition. In other words, the question being answered here is “what is cognition? And what are the main cognitive activities?” I will cover models that attempt to illustrate how cognition works at a later time; at which point the question I will address is “how does cognition function?”

The two modes of cognition identified by Don Norman are the experiential and reflective modes. Both of these are essential to human beings, and are continuously used in everyday life often in an overlapping manner. The description below and attached diagram aim to illustrate the main characteristics of each of each modes.
  • Experiential: state-of-mind associated to perception of the environment around us, and to our engagement with that environment through our actions and reactions. Contexts where an experiential mode of cognition is used include when a person is having a conversation, driving a car, or reading a book.
  • Reflective: state-of-mind associated to higher-level processing of knowledge, memory, and external information (or stimuli) through thinking, comparing, and judging. This type of cognition is needed for people to learn, create ideas, design products, and write books.
[source: Interaction Design: Beyond Human-Computer Interaction; Don Norman's book and Things That Make Us Smart.]

** What the hell is ID FMP? **

Thursday, March 12, 2009

ID FMP: Map of Relationship Between Conceptual and Mental Models

Developed by Don Norman, the model illustrated below demonstrates how relationship between a designer’s conceptual model and a user’s mental model is mediated by the system image of products or services.

So here is my explanation of what this model means: Designers develop product and service systems based on conceptual models [define] that they create or borrow. I use the terms product and service systems [define] refer to the ecosystem that encompasses products, services and their related artifacts and resources; these can include assets such as manuals and knowledge bases, and resource such as user groups and communities.

Users do not have access to the conceptual models of designers. Their understanding of how a product works is developed based on their interactions with the product itself, their previous experiences with the world, and their existing knowledge and expertise. All of these considerations affect how people interpret their experiences with a product, and the mental model [define] they create to explain how products work.

The term system [define] image refers to the way a product or service system actually appears to a user. System images are always imperfect representations of the conceptual models upon which they were built. For a product to be usable the system image needs to enable users to develop an accurate mental model of how relevant aspects of a product or service works.

An interesting feature of Norman’s 1988 model, is that designers relationship with system images is represented as a one-way phenomenon. This implies that once a product has been designed there is little opportunity for on-going improvements. During the last 20 years advances in technology and design methodology have made it possible for designers to continuously fine-tune product and service systems. This is especially true in the increasingly service-based world of software.

Bjoern Hartmann has revised Norman’s model to reflect the opportunity for designers to play an on-going role in improving the system image of the products they’ve created. The model he proposes includes a feedback loop that enables the user to communicate to the designer via the system image.

Hartmann posits that user-initiated feedback via the system will help identify mismatches between the designer’s conceptual model and the user’s model of how the system functions. Another important consideration is that offering an instantaneous feedback option in the same media on which the interaction is taking place will generate more reliable and richer data than feedback elicited later, or via a different channel.

[Sourced from Don Norman’s website, though I know Norman's framework was also featured in this book The Design of Everyday Things; Paper by Bjoern Hartmann written during graduate studies at Standford]

** What the hell is ID FMP? **