Back to results

Massachusetts Institute of Technology

Visual classification of co-verbal gestures for gesture understanding

Abstract

dc:description.abstract

A person's communicative intent can be better understood by either a human or a machine if the person's gestures are understood. This thesis project demonstrates an expansion of both the range of co-verbal gestures a machine can identify, and the range of communicative intents the machine can infer. We develop an automatic system that uses realtime video as sensory input and then segments, classifies, and responds to co-verbal gestures made by users in realtime as they converse with a synthetic character known as REA, which is being developed in parallel by Justine Cassell and her students at the MIT Media Lab. A set of 670 natural gestures, videotaped and visually tracked in the course of conversational interviews and then hand segmented and annotated according to a widely used gesture classification scheme, is used in an offline training process that trains Hidden Markov Model classifiers. A number of feature sets are extracted and tested in the offline training process, and the best performer is employed in an online HMM segmenter and classifier that requires no encumbering attachments to the user. Modifications made to the REA system enable REA to respond to the user's beat and deictic gestures as well as turntaking requests the user may convey in gesture.

Degree

thesis:*
Department dc:contributor.department
Massachusetts Institute of Technology. Dept. of Architecture. Program in Media Arts and Sciences.
Grantor dc:publisher
Massachusetts Institute of Technology
Year dc:date.issued
2001

Author and committee

dc:creator, dc:contributor.*
Author dc:creator
  • Campbell, Lee Winston
Advisor dc:contributor.advisor
  • Aaron F. Bobick.

Subjects

dc:subject × 1

Rights

dc:rights
Statement dc:rights
  • M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission.
Language dc:language.iso
eng

Identifiers

dc:identifier.*
Handle dc:identifier.uri
http://hdl.handle.net/1721.1/8707
OAI identifier oai:identifier
oai:dspace.mit.edu:1721.1/8707

Chain of custody

source
Harvested from
MIT
Base URL
dspace.mit.edu/oai/request
Last updated
2026-07-22
Source record
OAI-PMH GetRecord
citation

Campbell, Lee Winston. Visual classification of co-verbal gestures for gesture understanding. Massachusetts Institute of Technology, 2001. http://hdl.handle.net/1721.1/8707