Showing posts with label video. Show all posts
Showing posts with label video. Show all posts

Tuesday, February 13, 2018

MPEG news: a report from the 121st meeting, Gwangju, Korea

The original blog post can be found at the Bitmovin Techblog and has been updated here to focus on and highlight research aspects. Additionally, this version of the blog post will be also posted at ACM SIGMM Records.

The MPEG press release comprises the following topics:
  • Compact Descriptors for Video Analysis (CDVA) reaches Committee Draft level
  • MPEG-G standards reach Committee Draft for metadata and APIs
  • MPEG issues Calls for Visual Test Material for Immersive Applications
  • Internet of Media Things (IoMT) reaches Committee Draft level
  • MPEG finalizes its Media Orchestration (MORE) standard
At the end I will also briefly summarize what else happened with respect to DASH, CMAF, OMAF as well as discuss future aspects of MPEG.

Compact Descriptors for Video Analysis (CDVA) reaches Committee Draft level

The Committee Draft (CD) for CDVA has been approved at the 121st MPEG meeting, which is the first formal step of the ISO/IEC approval process for a new standard. This will become a new part of MPEG-7 to support video search and retrieval applications (ISO/IEC 15938-15).

Managing and organizing the quickly increasing volume of video content is a challenge for many industry sectors, such as media and entertainment or surveillance. One example task is scalable instance search, i.e., finding content containing a specific object instance or location in a very large video database. This requires video descriptors which can be efficiently extracted, stored, and matched. Standardization enables extracting interoperable descriptors on different devices and using software from different providers, so that only the compact descriptors instead of the much larger source videos can be exchanged for matching or querying. The CDVA standard specifies descriptors that fulfil these needs and includes (i) the components of the CDVA descriptor, (ii) its bitstream representation and (iii) the extraction process. The final standard is expected to be finished in early 2019.

CDVA introduces a new descriptor based on features which are output from a Deep Neural Network (DNN). CDVA is robust against viewpoint changes and moderate transformations of the video (e.g., re-encoding, overlays), it supports partial matching and temporal localization of the matching content. The CDVA descriptor has a typical size of 2–4 KBytes per second of video. For typical test cases, it has been demonstrated to reach a correct matching rate of 88% (at 1% false matching rate).

Research aspects: There are probably endless research aspects in the visual descriptor space ranging from validation of the achieved to results so far to further improving informative aspects with the goal to increase correct matching rate (and consequently decreasing the false matching rating). In general, however, the question is whether there's a need for descriptors in the era of bandwidth-storage-computing over-provisioning and the raising usage of artificial intelligence techniques such as machine learning and deep learning.

MPEG-G standards reach Committee Draft for metadata and APIs

In my previous report I introduced the MPEG-G standard for compression and transport technologies of genomic data. At the 121st MPEG meeting, metadata and APIs reached CD level. The former - metadata - provides relevant information associated to genomic data and the latter - APIs - allow for building interoperable applications capable of manipulating MPEG-G files. Additional standardization plans for MPEG-G include the CDs for reference software (ISO/IEC 23092-4) and conformance (ISO/IEC 23092-4), which are planned to be issued at the next 122nd MPEG meeting with the objective of producing Draft International Standards (DIS) at the end of 2018.

Research aspects: Metadata typically enables certain functionality which can be tested and evaluated against requirements. APIs allow to build applications and services on top of the underlying functions, which could be a driver for research projects to make use of such APIs.

MPEG issues Calls for Visual Test Material for Immersive Applications

I have reported about the Omnidirectional Media Format (OMAF) in my previous report. At the 121st MPEG meeting, MPEG was working on extending OMAF functionalities to allow the modification of viewing positions, e.g., in case of head movements when using a head-mounted display, or for use with other forms of interactive navigation. Unlike OMAF which only provides 3 degrees of freedom (3DoF) for the user to view the content from a perspective looking outwards from the original camera position, the anticipated extension will also support motion parallax within some limited range which is referred to as 3DoF+. In the future with further enhanced technologies, a full 6 degrees of freedom (6DoF) will be achieved with changes of viewing position over a much larger range. To develop technology in these domains, MPEG has issued two Calls for Test Material in the areas of 3DoF+ and 6DoF, asking owners of image and video material to provide such content for use in developing and testing candidate technologies for standardization. Details about these calls can be found at https://mpeg.chiariglione.org/.

Research aspects: The good thing about test material is that it allows for reproducibility, which is an important aspect within the research community. Thus, it is more than appreciated that MPEG issues such a call and let's hope that this material will become publicly available. Typically this kind of visual test material targets coding but it would be also interesting to have such test content for storage and delivery.

Internet of Media Things (IoMT) reaches Committee Draft level

The goal of IoMT is is to facilitate the large-scale deployment of distributed media systems with interoperable audio/visual data and metadata exchange. This standard specifies APIs providing media things (i.e., cameras/displays and microphones/loudspeakers, possibly capable of significant processing power) with the capability of being discovered, setting-up ad-hoc communication protocols, exposing usage conditions, and providing media and metadata as well as services processing them. IoMT APIs encompass a large variety of devices, not just connected cameras and displays but also sophisticated devices such as smart glasses, image/speech analyzers and gesture recognizers. IoMT enables the expression of the economic value of resources (media and metadata) and of associated processing in terms of digital tokens leveraged by the use of blockchain technologies.

Research aspects: The main focus of IoMT is APIs which provides easy and flexible access to the underlying device' functionality and, thus, are an important factor to enable research within this interesting domain. For example, using these APIs to enable communicates among these various media things could bring up new forms of interaction with these technologies.

MPEG finalizes its Media Orchestration (MORE) standard

MPEG "Media Orchestration" (MORE) standard reached Final Draft International Standard (FDIS), the final stage of development before being published by ISO/IEC. The scope of the Media Orchestration standard is as follows:
  • It supports the automated combination of multiple media sources (i.e., cameras, microphones) into a coherent multimedia experience.
  • It supports rendering multimedia experiences on multiple devices simultaneously, again giving a consistent and coherent experience.
  • It contains tools for orchestration in time (synchronization) and space.
MPEG expects that the Media Orchestration standard to be especially useful in immersive media settings. This applies notably in social virtual reality (VR) applications, where people share a VR experience and are able to communicate about it. Media Orchestration is expected to allow synchronizing the media experience for all users, and to give them a spatially consistent experience as it is important for a social VR user to be able to understand when other users are looking at them.

Research aspects: This standard enables the social multimedia experience proposed in literature. Interestingly, the W3C is working on something similar referred to as timing object and it would be interesting to see whether these approaches have some commonalities.

What else happened at the MPEG meeting?

DASH is fully in maintenance mode and we are still waiting for the 3rd edition which is supposed to be a consolidation of existing corrigenda and amendments. Currently only minor extensions are proposed and conformance/reference software is being updated. Similar things can be said for CMAF where we have one amendment and one corrigendum under development. Additionally, MPEG is working on CMAF conformance. OMAF has reached FDIS at the last meeting and MPEG is working on reference software and conformance also. It is expected that in the future we will see additional standards and/or technical reports defining/describing how to use CMAF and OMAF in DASH.

Regarding the future video codec, the call for proposals is out since the last meeting as announced in my previous report and responses are due for the next meeting. Thus, it is expected that the 122nd MPEG meeting will be the place to be in terms of MPEG’s future video codec. Speaking about the future, shortly after the 121st MPEG, Leonardo Chiariglione published a blog post entitled “a crisis, the causes and a solution”, which is related to HEVC licensing, Alliance for Open Media (AOM), and possible future options. The blog post certainly caused some reactions within the video community at large and I think this was also intended. Let’s hope it will galvanice the video industry -- not to push the button -- but to start addressing and resolving the issues. As pointed out in one of my other blog posts about what to care about in 2018, the upcoming MPEG meeting in April 2018 is certainly a place to be. Additionally, it highlights some conferences related to various aspects also discussed in MPEG which I'd like to republish here:
  • QoMEX -- Int'l Conf. on Quality of Multimedia Experience -- will be hosted in Sardinia, Italy from May 29-31, which is THE conference to be for QoE of multimedia applications and services. Submission deadline is January 15/22, 2018.
  • MMSys -- Multimedia Systems Conf. -- and specifically Packet Video, which will be on June 12 in Amsterdam, The Netherlands. Packet Video is THE adaptive streaming scientific event 2018. Submission deadline is March 1, 2018.
  • Additionally, you might be interested in ICME (July 23-27, 2018, San Diego, USA), ICIP (October 7-10, 2018, Athens, Greece; specifically in the context of video coding), and PCS (June 24-27, 2018, San Francisco, CA, USA; also in the context of video coding).
  • The DASH-IF academic track hosts special events at MMSys (Excellence in DASH Award) and ICME (DASH Grand Challenge).
  • MIPR -- 1st Int'l Conf. on Multimedia Information Processing and Retrieval -- will be in Miami, Florida, USA from April 10-12, 2018. It has a broad range of topics including networking for multimedia systems as well as systems and infrastructures.

Tuesday, November 18, 2014

ACM International Conference on Interactive Experiences for Television & Online Video


*** TVX 2015 ***
ACM International Conference on Interactive Experiences for Television & Online Video
3rd – 5th June 2015

Hosted by iMinds Digital Society Department at the Crowne Plaza Hotel, Brussels, Belgium

Jointly organised with the International Symposium on Media Innovations (ISMI) and the Private Television Conference

IMPORTANT DATES:
  •  November 15, 2014: Course and Workshop proposals
  • January 12, 2015: Full and Short Paper submissions
  • March 2, 2015: WiP, TVX in Industry, Demo, Doctoral Consortium
TVX is the ACM International Conference on Interactive Experiences for Television and Online Video. TVX is the leading international conference for presentation and discussion of research into online video and TV interaction and user experience. The conference brings together international researchers and practitioners from a wide range of disciplines, ranging from human-computer interaction, multimedia engineering and design to media studies, media psychology and sociology. In addition to standard research paper presentations the conference includes a wide range of formats for presentation and discussion of research, including Industry Papers, Demos, Works-in-Progress, and also provides the opportunity to participate in the Doctoral Consortium and to run and attend courses and workshops on specialist topics in TV and online video interaction and user experience.

Topics of interest include (but are not limited to):
  • Content Production: traditional & novel content production for the new media landscape, including cross-platform services and interactive storytelling, and personalisation.
  • Systems & Infrastructures: system designs and architectures and their evaluation, including delivery, transmission, and synchronization of media.
  • Interaction Technologies & techniques: including gestural and multi-sensory, multi-display systems, and interaction for device ecosystems.
  • Experience Design & Evaluation: TV and online video design and evaluation research, including social and shared experiences.
  • Media Studies: including consumption practices and theoretical and practical ethical, regulatory, and policy issues.
  • Empirical Methods: novel methods for evaluating TV and online video experience, and audience measurement.
  • Data Science for TV & Online Video: advances in techniques for collaborative filtering, interactive/synchronous environments, collective intelligence and crowd-sourcing, and location-based and context-aware applications and services.
  • Business Models & Marketing: research and practice around novel business models and marketing strategies for the new media landscape of television and online video. Studies around novel ways of advertisement models and strategies.
  • Innovation & Visions of Future TV & Online Video: research on innovative design strategies, new concepts, and prototype experiences for television and online video, including case studies and media artworks and performances.
SUBMISSION GUIDELINES

Contributions must describe unpublished original work, emphasizing completed or advanced research, and a parallel submission to other venues should be clearly indicated to the program committee. Research paper submissions are double-blind and will be reviewed by at least three program committee members.

For detailed submission guidelines, including the use of the ACM SIGCHI PCS submission system see: http://tvx2015.com/participation

MENTORING PROGRAM

Along the inclusion and accessibility strategy at TVX2015, we also provide a mentoring opportunity. During the TVX submission process, we can provide the opportunity to bring the experience of established researchers to new researchers. In the mentorship program we get you in contact with a specific member of the community who will provide feedback and support for your submission.

Please find more information on how to ask for mentoring, becoming a mentor, and our existing mentors: http://tvx2015.com/inclusion-accessibility/mentoring-program/

CONTACT

For up-to-date information and further details visit: http://tvx2015.com/

For questions, please contact us on: info@tvx2015.com

ORGANISATION COMMITTEE

GENERAL CHAIRS

David Geerts, iMinds / KU Leuven, Belgium

Lieven De Marez, iMinds / UGent, Belgium

Caroline Pauwels, iMinds / VUB, Belgium

PROGRAM CHAIRS

Frank Bentley, Yahoo Labs, USA

Christian Timmerer, Alpen-Adria-Universität Klagenfurt, Austria

WORK IN PROGRESS CHAIRS

Hokyoung Blake Ryu, Hanyang University, South Korea

Jeroen Vanattenhoven, iMinds / KU Leuven, Belgium

WORKSHOP CHAIRS

Rene Kaiser, Joanneum Research, Austria

Noor Ali-Hasan, Google, USA

COURSE CHAIRS

Pedro Almeida, University of Aveiro, Portugal

Santosh Basapur, Illinois Institute of Technology, USA

DOCTORAL CONSORTIUM CHAIRS

Marian Ursu, University of York, UK

Teresa Chambel, University of Lisbon, Portugal

TVX IN INDUSTRY CHAIRS

Katia Aerts, Mike Matton, VRT, Belgium

DEMO CHAIRS

Tom Bartindale, Culture Lab, Newcastle University, UK

Rinze Leenheer, iMinds / KU Leuven, Belgium

INCLUSION AND ACCESSIBILITY CHAIR

Reuben Kirkham, Culture Lab, Newcastle University, UK

Tom Evens, iMinds / UGent, Belgium

LOCAL PRODUCTION CHAIR

Jonathan Huyghe, iMinds / KU Leuven, Belgium

Thursday, June 26, 2014

VideoNext: Design, Quality and Deployment of Adaptive Video Streaming


The workshop co-located with CoNEXT 2014
December 2, 2014
Sydney, Australia

Submission deadline changed: August 29, 2014 (no further extensions)

Call for Papers

As we continue to develop our ability to generate, process, and display video at increasingly higher quality, we confront the challenge of streaming the same video to the end user. Device heterogeneity in terms of size and processing capabilities combined with the lack of timing guarantees of packet switching networks is forcing the industry to adopt streaming solutions capable of dynamically adapting the video quality in response to resource variability in the end-to-end transport chain. For example, many vendors and providers are already trialing their own proprietary adaptive video streaming platforms while MPEG has recently ratified a standard, called Dynamic Adaptive Streaming over HTTP (DASH), to facilitate widespread deployment of such technology. However, how to best adapt the video to ensure highest user quality of experience while consuming the minimum network resources poses many fundamental challenges, which is attracting the attention of researchers from both academia and industry. The goal of this workshop is to bring together researchers and developers working on all aspects of adaptive video streaming with special emphasis on innovative concepts backed up by experimental evidence.

Specific areas of interest include, but are not limited to:
  • New metrics for measuring user quality of experience (QoE) for adaptive video streaming
  • Solutions for improving streaming QoE for high-speed user mobility
  • Analysis, modelling, and experimentation of DASH
  • Exploitation of user contexts for improving efficiency of adaptive streaming
  • Big data analytics to assess viewer experience of adaptive video
  • Efficient and fair bandwidth sharing techniques for bottleneck links supporting multiple adaptive video streams
  • Network functions to assist and improve adaptive video streaming
  • Synchronization issues in adaptive video streaming (inter-media, inter-device/destination)
  • Methods for effective simulation or emulation of large scale adaptive video streaming platforms
  • Cloud-assisted adaptive video streaming including encoding, transcoding, and adaptation in general
  • Attack scenarios and solutions for adaptive video streaming
  • Energy-efficient adaptive streaming for resource-constraint mobile devices
  • Reproducible research in adaptive video streaming: datasets, evaluation methods, benchmarking, standardization efforts, open source tools
  • Novel use cases and applications in the area of adaptive video streaming

The workshop is considered an integral part of the CoNEXT 2014 conference. All workshop papers will be published in the same set of proceedings as the main conference, and available on the ACM Digital Library. Publication at this workshop is not intended to preclude later publication of an extended version of the paper. At least one author of each accepted papers is expected to present his/her paper at the workshop.

Instructions for Authors

A submission must be no greater than 6 pages in length including all figures, tables, references, appendices, etc., and must be a PDF file of less than 10MB. The review process is single-blind.

Follow the same formatting guidelines as the CoNEXT conference, except VideoNext has a 6 page limit and a 10MB file size limit. See the “Formatting Guidelines” section. Submissions that deviate from these guidelines will be rejected without consideration.

Then use the paper submission site to submit your paper by 8:59 pm Pacific Standard Time (PDT), August 29, 2014.
Important dates
  • Paper Submission: August 2229, 2014 20:59 PDT
  • Notification of Acceptance: September 30, 2014
  • Camera-ready Papers Due: October 24, 2014
  • Workshop: December 2, 2014

TPC co-chairs
  • Mahbub Hassan, University of New South Wales, Australia
  • Ali C. Begen, Cisco Canada
  • Christian Timmerer, Alpen-Adria-Universität Klagenfurt, Austria

Technical Program Committee
  • Alexander Raake, Deutsche Telecom Labs, Germany
  • Carsten Griwodz, University of Oslo/Simula, Sweden
  • Chao Chen, Qualcom, USA
  • Colin Perkins, University of Glasgow, Scotland
  • Constantine Dovrolis, Georgia Tech, USA
  • Grenville Armitage, Swinburne University of Technology, Australia
  • Imed Bouazizi, Samsung
  • Kuan-Ta Chen, Academia Sinica
  • Magda El Zarki, University of California Irvine, USA
  • Manzur Murshed, Federation University Australia, Australia
  • Pal Halvorsen, University of Oslo/Simula
  • Polychronis Koutsakis, Technical University of Crete, Greece
  • Roger Zimmerman, National University of Singapore, Singapore
  • Saverio Mascolo, University of Bari, Italy
  • Shervin Shirmohammadi ,University of Ottawa, Canada
  • Victor Leung, University of British Columbia, Canada

Tuesday, September 9, 2008

P2P'08: Video Search & Playback in Zero-Server P2P Systems

In yesterday's tutorial at P2P'08 I learned something about zero-server P2P systems, especially related to video search and playback. Unfortunately, the BBC guy did not show up and so the audience was directly confronted with the reality: standardization in DVB ;-) However, the speaker gave a good overview how P2P systems have been used in the past and especially about the trial for the European Song Contest. Next, we got input from the industry representing the consumer electronics. They want to put a P2P engine into a settop box and raised a few very interesting issues, mainly related to CPU and memory constraints (e.g., only 4 MB of free memory for metadata or so). Also, a timely standardozation is required in order to bring interoperable products on the market. The main part of the tutorial comprises an overview about Tribler, the P2P engine also used in P2P-Next project. One interesting talk was identifying the differences between traditional file sharing, VoD, and "live" streaming. However, I always wonder how muh "live" a P2P live video stream is? Just one keyword: delay!

More to come soon ...