880 resultados para Pen-based video annotation
Resumo:
Dissertação para obtenção do Grau de Doutor em Informática
Resumo:
This paper presents a semi-parametric Algorithm for parsing football video structures. The approach works on a two interleaved based process that closely collaborate towards a common goal. The core part of the proposed method focus perform a fast automatic football video annotation by looking at the enhance entropy variance within a series of shot frames. The entropy is extracted on the Hue parameter from the HSV color system, not as a global feature but in spatial domain to identify regions within a shot that will characterize a certain activity within the shot period. The second part of the algorithm works towards the identification of dominant color regions that could represent players and playfield for further activity recognition. Experimental Results shows that the proposed football video segmentation algorithm performs with high accuracy.
Resumo:
Dissertação para obtenção do Grau de Mestre em Engenharia Informática
Resumo:
The Iowa Department of Transportation is committed to improved management systems, which in turn has led to increased automation to record and manage construction data. A possible improvement to the current data management system can be found with pen-based computers. Pen-based computers coupled with user friendly software are now to the point where an individual's handwriting can be captured and converted to typed text to be used for data collection. It would appear pen-based computers are sufficiently advanced to be used by construction inspectors to record daily project data. The objective of this research was to determine: (1) if pen-based computers are durable enough to allow maintenance-free operation for field work during Iowa's construction season; and (2) if pen-based computers can be used effectively by inspectors with little computer experience. The pen-based computer's handwriting recognition was not fast or accurate enough to be successfully utilized. The IBM Thinkpad with the pen pointing device did prove useful for working in Windows' graphical environment. The pen was used for pointing, selecting and scrolling in the Windows applications because of its intuitive nature.
Resumo:
This paper presents a Robust Content Based Video Retrieval (CBVR) system. This system retrieves similar videos based on a local feature descriptor called SURF (Speeded Up Robust Feature). The higher dimensionality of SURF like feature descriptors causes huge storage consumption during indexing of video information. To achieve a dimensionality reduction on the SURF feature descriptor, this system employs a stochastic dimensionality reduction method and thus provides a model data for the videos. On retrieval, the model data of the test clip is classified to its similar videos using a minimum distance classifier. The performance of this system is evaluated using two different minimum distance classifiers during the retrieval stage. The experimental analyses performed on the system shows that the system has a retrieval performance of 78%. This system also analyses the performance efficiency of the low dimensional SURF descriptor.
Resumo:
We study state-based video communication where a client simultaneously informs the server about the presence status of various packets in its buffer. In sender-driven transmission, the client periodically sends to the server a single acknowledgement packet that provides information about all packets that have arrived at the client by the time the acknowledgment is sent. In receiver-driven streaming, the client periodically sends to the server a single request packet that comprises a transmission schedule for sending missing data to the client over a horizon of time. We develop a comprehensive optimization framework that enables computing packet transmission decisions that maximize the end-to-end video quality for the given bandwidth resources, in both prospective scenarios. The core step of the optimization comprises computing the probability that a single packet will be communicated in error as a function of the expected transmission redundancy (or cost) used to communicate the packet. Through comprehensive simulation experiments, we carefully examine the performance advances that our framework enables relative to state-of-the-art scheduling systems that employ regular acknowledgement or request packets. Consistent gains in video quality of up to 2B are demonstrated across a variety of content types. We show that there is a direct analogy between the error-cost efficiency of streaming a single packet and the overall rate-distortion performance of streaming the whole content. In the case of sender-driven transmission, we develop an effective modeling approach that accurately characterizes the end-to-end performance as a function of the packet loss rate on the backward channel and the source encoding characteristics.
Resumo:
Colorectal cancer (CRC) is the third largest cause of cancer death in the United States. While the disease burden is high, there are proven methods to screen for CRC and detect it at a stage that is amenable to cure. Patients with low health literacy have difficulty navigating the health care system and are at increased risk to not receive preventive care services such as colorectal cancer screening (CRCS). To address this need, an exam-room based video was developed to be played for patients in the privacy of the exam room, while they are waiting to be seen by their medical provider. In roughly 2 minutes, the video informs the patient about CRC and CRCS and how they can successfully complete CRCS. One of the key barriers to completing CRCS is the need to increase patients' knowledge and improve attitudes surrounding CRCS. This study examines the impact of the video on patients' knowledge and attitudes about CRC and CRCS in a medically underserved patient population in Houston, Texas. ^ Sixty-one patients presenting for routine medical care were enrolled in the study. Depending on their randomization, the patients either received routine information about CRC and CRCS or they watched the video. We found that the patients who did watch the video did have improvements in their knowledge and improved attitudes about CRC and CRCS. Future studies will be needed to examine whether the video improves the patients' completion of CRCS.^
Resumo:
We have designed and tested an Internet-based video-phone suitable for use in the homes of families in need of paediatric palliative care services. The equipment uses an ordinary telephone line and includes a PC, Web camera and modem housed in a custom-made box. In initial field testing, six clinical consultations were conducted in a one-month trial of the videophone with a family in receipt of palliative care services who were living in the outer suburbs of Brisbane. Problems with variability in call quality-namely audio and video freezing, and audio break-up-prompted further laboratory testing. We completed a programme of over 250 test calls. Fixing modem connection parameters to use the V.34 modulation protocol at a set bandwidth of 24 kbit/s improved connection stability and the reliability of the video-phone. In subsequent field testing 47 of 50 calls (94%) connected without problems. The freezes that did occur were brief (with greatly reduced packet loss) and had little effect on the ability to communicate, unlike the problems arising in the home testing. The low-bandwidth Internet-based video-phone we have developed provides a feasible means of doing telemedicine in the home.
Resumo:
In order to address problems of information overload in digital imagery task domains we have developed an interactive approach to the capture and reuse of image context information. Our framework models different aspects of the relationship between images and domain tasks they support by monitoring the interactive manipulation and annotation of task-relevant imagery. The approach allows us to gauge a measure of a user's intentions as they complete goal-directed image tasks. As users analyze retrieved imagery their interactions are captured and an expert task context is dynamically constructed. This human expertise, proficiency, and knowledge can then be leveraged to support other users in carrying out similar domain tasks. We have applied our techniques to two multimedia retrieval applications for two different image domains, namely the geo-spatial and medical imagery domains. © Springer-Verlag Berlin Heidelberg 2007.
Resumo:
An increasing interest in “bringing actors back in” and gaining a nuanced understanding of their actions and interactions across a variety of strands in the management literature, has recently helped ethnography to unknown prominence in the field of organizational studies. Yet, calls remain that ethnography should “play a much more central role in the organization and management studies repertoire than it currently does” (Watson, 2011: 202). Ironically, those organizational realities that ethnographers are called to examine have at the same time become less and less amenable to ethnographic study. In this paper, we respond to these calls for innovative ethnographic methods in two ways. First, we report on the practices and ethnographic experiences of conducting a year-long team-based video ethnography of reinsurance trading in Lloyd’s of London. Second, drawing on these experiences, we propose an initial framework for systematizing new approaches to organizational ethnography and visualizing the ways in which they are ‘expanding’ ethnography as it was traditionally practiced.
Resumo:
Automatic video segmentation plays a vital role in sports videos annotation. This paper presents a fully automatic and computationally efficient algorithm for analysis of sports videos. Various methods of automatic shot boundary detection have been proposed to perform automatic video segmentation. These investigations mainly concentrate on detecting fades and dissolves for fast processing of the entire video scene without providing any additional feedback on object relativity within the shots. The goal of the proposed method is to identify regions that perform certain activities in a scene. The model uses some low-level feature video processing algorithms to extract the shot boundaries from a video scene and to identify dominant colours within these boundaries. An object classification method is used for clustering the seed distributions of the dominant colours to homogeneous regions. Using a simple tracking method a classification of these regions to active or static is performed. The efficiency of the proposed framework is demonstrated over a standard video benchmark with numerous types of sport events and the experimental results show that our algorithm can be used with high accuracy for automatic annotation of active regions for sport videos.
Resumo:
Personalised video can be achieved by inserting objects into a video play-out according to the viewer's profile. Content which has been authored and produced for general broadcast can take on additional commercial service features when personalised either for individual viewers or for groups of viewers participating in entertainment, training, gaming or informational activities. Although several scenarios and use-cases can be envisaged, we are focussed on the application of personalised product placement. Targeted advertising and product placement are currently garnering intense interest in the commercial networked media industries. Personalisation of product placement is a relevant and timely service for next generation online marketing and advertising and for many other revenue generating interactive services. This paper discusses the acquisition and insertion of media objects into a TV video play-out stream where the objects are determined by the profile of the viewer. The technology is based on MPEG-4 standards using object based video and MPEG-7 for metadata. No proprietary technology or protocol is proposed. To trade the objects into the video play-out, a Software-as-a-Service brokerage platform based on intelligent agent technology is adopted. Agencies, libraries and service providers are represented in a commercial negotiation to facilitate the contractual selection and usage of objects to be inserted into the video play-out.
Resumo:
Dissertação para obtenção do Grau de Mestre em Engenharia Informática
Resumo:
Currently the world swiftly adapts to visual communication. Online services like YouTube and Vine show that video is no longer the domain of broadcast television only. Video is used for different purposes like entertainment, information, education or communication. The rapid growth of today’s video archives with sparsely available editorial data creates a big problem of its retrieval. The humans see a video like a complex interplay of cognitive concepts. As a result there is a need to build a bridge between numeric values and semantic concepts. This establishes a connection that will facilitate videos’ retrieval by humans. The critical aspect of this bridge is video annotation. The process could be done manually or automatically. Manual annotation is very tedious, subjective and expensive. Therefore automatic annotation is being actively studied. In this thesis we focus on the multimedia content automatic annotation. Namely the use of analysis techniques for information retrieval allowing to automatically extract metadata from video in a videomail system. Furthermore the identification of text, people, actions, spaces, objects, including animals and plants. Hence it will be possible to align multimedia content with the text presented in the email message and the creation of applications for semantic video database indexing and retrieving.