The Experts below are selected from a list of 3405 Experts worldwide ranked by ideXlab platform

Ronald M. Baecker - One of the best experts on this subject based on the ideXlab platform.

  • CHI - Attention by proxy? issues in audience awareness for Webcasts to distributed groups
    Proceeding of the twenty-sixth annual CHI conference on Human factors in computing systems - CHI '08, 2008
    Co-Authors: Jeremy Birnholtz, Saul Greenberg, Ronald M. Baecker
    Abstract:

    Instructor/student interaction in e-learning environments can positively impact both student learning and instructor satisfaction. In online webcast lectures, however, interaction can be difficult because instructors lack basic awareness information about their remote students. Our goal is to better understand the kinds of awareness information that instructors should have if they are to interact frequently and effectively with their students in e-learning environments. We conducted an exploratory study -- via interviews and observations -- of instructor attention in face-to-face classrooms at a large university. Our results imply that a webcast system should provide instructors with overview and detailed data about their students, but that this detailed information should not be displayed publicly.

  • attention by proxy issues in audience awareness for Webcasts to distributed groups
    Human Factors in Computing Systems, 2008
    Co-Authors: Jeremy Birnholtz, Saul Greenberg, Ronald M. Baecker
    Abstract:

    Instructor/student interaction in e-learning environments can positively impact both student learning and instructor satisfaction. In online webcast lectures, however, interaction can be difficult because instructors lack basic awareness information about their remote students. Our goal is to better understand the kinds of awareness information that instructors should have if they are to interact frequently and effectively with their students in e-learning environments. We conducted an exploratory study -- via interviews and observations -- of instructor attention in face-to-face classrooms at a large university. Our results imply that a webcast system should provide instructors with overview and detailed data about their students, but that this detailed information should not be displayed publicly.

  • automatic speech recognition for Webcasts how good is good enough and what to do when it isn t
    International Conference on Multimodal Interfaces, 2006
    Co-Authors: Cosmin Munteanu, Gerald Penn, Ronald M. Baecker, Yuecheng Zhang
    Abstract:

    The increased availability of broadband connections has recently led to an increase in the use of Internet broadcasting (webcasting). Most Webcasts are archived and accessed numerous times retrospectively. One challenge to skimming and browsing through such archives is the lack of text transcripts of the webcast's audio channel. This paper describes a procedure for prototyping an Automatic Speech Recognition (ASR) system that generates realistic transcripts of any desired Word Error Rate (WER), thus overcoming the drawbacks of both prototype-based and Wizard of Oz simulations. We used such a system in a user study showing that transcripts with WERs less than 25% are acceptable for use in webcast archives. As current ASR systems can only deliver, in realistic conditions, Word Error Rates (WERs) of around 45%, we also describe a solution for reducing the WER of such transcripts by engaging users to collaborate in a "wiki" fashion on editing the imperfect transcripts obtained through ASR.

  • CHI - The effect of speech recognition accuracy rates on the usefulness and usability of webcast archives
    Proceedings of the SIGCHI conference on Human Factors in computing systems - CHI '06, 2006
    Co-Authors: Cosmin Munteanu, Gerald Penn, Ronald M. Baecker, Elaine G Toms, David F. James
    Abstract:

    The widespread availability of broadband connections has led to an increase in the use of Internet broadcasting (webcasting). Most Webcasts are archived and accessed numerous times retrospectively. In the absence of transcripts of what was said, users have difficulty searching and scanning for specific topics. This research investigates user needs for transcription accuracy in webcast archives, and measures how the quality of transcripts affects user performance in a question-answering task, and how quality affects overall user experience. We tested 48 subjects in a within-subjects design under 4 conditions: perfect transcripts, transcripts with 25% Word Error Rate (WER), transcripts with 45% WER, and no transcript. Our data reveals that speech recognition accuracy linearly influences both user performance and experience, shows that transcripts with 45% WER are unsatisfactory, and suggests that transcripts having a WER of 25% or less would be useful and usable in webcast archives.

  • enhancing interactivity in Webcasts with voip
    Human Factors in Computing Systems, 2006
    Co-Authors: Ronald M. Baecker, Kelly Rankin, Clarence Chan, Melanie Baran, Jeremy Birnholtz, Joe Laszlo, Russ Schick, Peter Wolf
    Abstract:

    This Interactivity demonstration presents a novel coupling of webcasting (streaming) with audioconferencing in which Voice over Internet Protocol (VoIP) communication is used in Webcasts to enhance interactivity, engagement, and the sense of presence among viewers and presenters.

Cosmin Munteanu - One of the best experts on this subject based on the ideXlab platform.

  • Collaborative editing for improved usefulness and usability of transcript-enhanced Webcasts
    Conference on Human Factors in Computing Systems - Proceedings, 2008
    Co-Authors: Cosmin Munteanu, Ron Baecker, Gerald Penn
    Abstract:

    One challenge in facilitating skimming or browsing through archives of on-line recordings of webcast lectures is the lack of text transcripts of the recorded lecture. Ideally, transcripts would be obtainable through Automatic Speech Recognition (ASR). However, current ASR systems can only deliver, in realistic lecture conditions, a Word Error Rate of around 45% -- above the accepted threshold of 25%. In this paper, we present the iterative design of a webcast extension that engages users to collaborate in a wiki-like manner on editing the ASR-produced imperfect transcripts, and show that this is a feasible solution for improving the quality of lecture transcripts. We also present the findings of a field study carried out in a real lecture environment investigating how students use and edit the transcripts.

  • Usable speech recognition: toward improved access to webcast lectures
    2008
    Co-Authors: Cosmin Munteanu
    Abstract:

    A growing number of lecture Webcasts are archived after being delivered live. In the absence of transcripts, users are faced with increased difficulty in performing tasks easily achieved with text documents (retrieval, browsing, skimming). Unfortunately, speech recognition systems do not perform satisfactorily when transcribing lectures. In this paper, we present an overview of the ePresence lecture transcription project, whose goal is to improve the usefulness and usability of automaticallygenerated transcripts of webcast lectures. We achieve this by integrating novel speech recognition techniques specifically addressed at increasing the accuracy of webcast transcriptions with the development of an interactive collaborative interface that facilitates users' contribution to the improvement of machine-generated transcripts. We conclude by discussing the challenges (and possible solutions) to successfully integrate transcripts into archives of webcast lectures.

  • automatic speech recognition for Webcasts how good is good enough and what to do when it isn t
    International Conference on Multimodal Interfaces, 2006
    Co-Authors: Cosmin Munteanu, Gerald Penn, Ronald M. Baecker, Yuecheng Zhang
    Abstract:

    The increased availability of broadband connections has recently led to an increase in the use of Internet broadcasting (webcasting). Most Webcasts are archived and accessed numerous times retrospectively. One challenge to skimming and browsing through such archives is the lack of text transcripts of the webcast's audio channel. This paper describes a procedure for prototyping an Automatic Speech Recognition (ASR) system that generates realistic transcripts of any desired Word Error Rate (WER), thus overcoming the drawbacks of both prototype-based and Wizard of Oz simulations. We used such a system in a user study showing that transcripts with WERs less than 25% are acceptable for use in webcast archives. As current ASR systems can only deliver, in realistic conditions, Word Error Rates (WERs) of around 45%, we also describe a solution for reducing the WER of such transcripts by engaging users to collaborate in a "wiki" fashion on editing the imperfect transcripts obtained through ASR.

  • CHI - The effect of speech recognition accuracy rates on the usefulness and usability of webcast archives
    Proceedings of the SIGCHI conference on Human Factors in computing systems - CHI '06, 2006
    Co-Authors: Cosmin Munteanu, Gerald Penn, Ronald M. Baecker, Elaine G Toms, David F. James
    Abstract:

    The widespread availability of broadband connections has led to an increase in the use of Internet broadcasting (webcasting). Most Webcasts are archived and accessed numerous times retrospectively. In the absence of transcripts of what was said, users have difficulty searching and scanning for specific topics. This research investigates user needs for transcription accuracy in webcast archives, and measures how the quality of transcripts affects user performance in a question-answering task, and how quality affects overall user experience. We tested 48 subjects in a within-subjects design under 4 conditions: perfect transcripts, transcripts with 25% Word Error Rate (WER), transcripts with 45% WER, and no transcript. Our data reveals that speech recognition accuracy linearly influences both user performance and experience, shows that transcripts with 45% WER are unsatisfactory, and suggests that transcripts having a WER of 25% or less would be useful and usable in webcast archives.

  • Measuring the acceptable word error rate of machine-generated webcast transcripts
    Proceedings of the Annual Conference of the International Speech Communication Association INTERSPEECH, 2006
    Co-Authors: Cosmin Munteanu, Gerald Penn, Ron Baecker, Elaine Toms, David James
    Abstract:

    The increased availability of broadband connections has recently led to an increase in the use of Internet broadcasting (webcasting). Most Webcasts are archived and accessed numerous times retrospectively. One of the hurdles users face when browsing and skimming through archives is the lack of text transcripts of the audio channel of the webcast archive. In this paper, we proposed a procedure for prototyping an Automatic Speech Recognition (ASR) system that generates realistic transcripts of any desired Word Error Rate (WER), thus overcoming the drawbacks of both prototypebased and Wizard of Oz simulations. We used such a system in a study where human subjects perform question-answering tasks using archives of webcast lectures, and showed that their performance and perception of transcript quality is linearly affected by WER, and that transcripts of WER equal or less than 25% would be acceptable for use in webcast archives.

Michael Böhm - One of the best experts on this subject based on the ideXlab platform.

  • Highlights of the hotline sessions presented at the scientific sessions 2008 of the American Heart Association
    Clinical Research in Cardiology, 2009
    Co-Authors: Helge Möllmann, Michael Böhm, Ulrich Laufs
    Abstract:

    Summaries and commentaries on trials presented at the hotline sessions of the scientific sessions 2008 of the American Heart Association in New Orleans have been generated from the oral presentations and the Webcasts of the American Heart Association. The following papers are discussed: APPROACH, ATLAS, BACH, BICC, HF-ACTION, I-PRESERVE, JPAD, JUPITER, Mass-DAC, Physicians’ Health Study II, SEARCH, tailored clopidogrel loading to prevent stent thrombosis, and TIMACS.

  • Clinical Trial Updates and Hotline Sessions presented at the Scientific Session 2007 of the American heart association
    Clinical Research in Cardiology, 2008
    Co-Authors: Ulrich Laufs, Helge Möllmann, Florian Custodis, Michael Böhm
    Abstract:

    This article provides information and commentaries on trials which were presented at Clinical Trial Updates and Hotline Sessions presented at the Scientific Sessions 2007 of the American Heart Association in Orlando, Florida. The comprehensive summaries have been generated from the oral presentations and the Webcasts of the American Heart Association. Most reports have not been published as full papers and therefore have to be considered as preliminary data, as the analysis may change in the final publications. The following papers are discussed: TRITON TIMI-38, EVA-AMI, BRIEF-PCI, RACE, MASS Stent, HF-ART, STITCH, CORONA, ILLUMINATE, CORE-64, OAT Substudy, AFCHF, MASCOT, RETHINQ, MASTER I, POISE, COUMA-GEN, HIJ-CREATE, PROVIDENCE I, CAUSMIC, IC-BMC, IC/IM BMCs.

  • Clinical Trial Updates and Hotline Sessions presented at the European Society of Cardiology Congress 2007
    Clinical Research in Cardiology, 2007
    Co-Authors: Michael Kindermann, Oliver Adam, Nikos Werner, Michael Böhm
    Abstract:

    This article provides information and commentaries on trials which were presented at the Hotline and Clinical Trial Update Sessions at the European Society of Cardiology Congress 2007 in Vienna. The key presentations were performed by leading experts in the field with relevant positions in the trials or registries. It is important to note that unpublished reports should be considered as preliminary data, as the analysis may change in the final publications. The comprehensive summaries have been generated from the oral presentation and the Webcasts of the European Society of Cardiology and should provide the readers with the most comprehensive information of relevant publications.

Ulrich Laufs - One of the best experts on this subject based on the ideXlab platform.

  • Highlights of the hotline sessions presented at the scientific sessions 2008 of the American Heart Association
    Clinical Research in Cardiology, 2009
    Co-Authors: Helge Möllmann, Michael Böhm, Ulrich Laufs
    Abstract:

    Summaries and commentaries on trials presented at the hotline sessions of the scientific sessions 2008 of the American Heart Association in New Orleans have been generated from the oral presentations and the Webcasts of the American Heart Association. The following papers are discussed: APPROACH, ATLAS, BACH, BICC, HF-ACTION, I-PRESERVE, JPAD, JUPITER, Mass-DAC, Physicians’ Health Study II, SEARCH, tailored clopidogrel loading to prevent stent thrombosis, and TIMACS.

  • Clinical Trial Updates and Hotline Sessions presented at the Scientific Session 2007 of the American heart association
    Clinical Research in Cardiology, 2008
    Co-Authors: Ulrich Laufs, Helge Möllmann, Florian Custodis, Michael Böhm
    Abstract:

    This article provides information and commentaries on trials which were presented at Clinical Trial Updates and Hotline Sessions presented at the Scientific Sessions 2007 of the American Heart Association in Orlando, Florida. The comprehensive summaries have been generated from the oral presentations and the Webcasts of the American Heart Association. Most reports have not been published as full papers and therefore have to be considered as preliminary data, as the analysis may change in the final publications. The following papers are discussed: TRITON TIMI-38, EVA-AMI, BRIEF-PCI, RACE, MASS Stent, HF-ART, STITCH, CORONA, ILLUMINATE, CORE-64, OAT Substudy, AFCHF, MASCOT, RETHINQ, MASTER I, POISE, COUMA-GEN, HIJ-CREATE, PROVIDENCE I, CAUSMIC, IC-BMC, IC/IM BMCs.

Qingming Huang - One of the best experts on this subject based on the ideXlab platform.

  • Using webcast text for semantic event detection in broadcast sports video
    IEEE Transactions on Multimedia, 2008
    Co-Authors: Changsheng Xu, Hanqing Lu, Yong Rui, Guangyu Zhu, Yi-fan Zhang, Qingming Huang
    Abstract:

    Sports video semantic event detection is essential for sports video summarization and retrieval. Extensive research efforts have been devoted to this area in recent years. However, the existing sports video event detection approaches heavily rely on either video content itself, which face the difficulty of high-level semantic information extraction from video content using computer vision and image processing techniques, or manually generated video ontology, which is domain specific and difficult to be automatically aligned with the video content. In this paper, we present a novel approach for sports video semantic event detection based on analysis and alignment of Webcast text and broadcast video. Webcast text is a text broadcast channel for sports game which is co-produced with the broadcast video and is easily obtained from the Web. We first analyze Webcast text to cluster and detect text events in an unsupervised way using probabilistic latent semantic analysis (pLSA). Based on the detected text event and video structure analysis, we employ a conditional random field model (CRFM) to align text event and video event by detecting event moment and event boundary in the video. Incorporation of Webcast text into sports video analysis significantly facilitates sports video semantic event detection. We conducted experiments on 33 hours of soccer and basketball games for Webcast analysis, broadcast video analysis and text/video semantic alignment. The results are encouraging and compared with the manually labeled ground truth.