MPEG Column: 155th MPEG Meeting

The 155th MPEG meeting took place in Geneva, Switzerland, from July 13 to 17, 2026. The official MPEG press release can be found here. This report highlights key outcomes from the meeting, with a focus on research directions relevant to the ACM SIGMM community:

  • Gaussian Splat Coding Use Cases and Current Status
  • Call for Proposals on Media Authenticity and Provenance Indication
  • Joint Call for Proposals on video compression with capability beyond VVC
  • Joint ITU-T SG 21 and ISO/IEC JTC 1/SC 29 Workshop on Media Streaming Services
  • Exploration on Systems technologies for AI-based media standards (SyfAI)

Gaussian Splat Coding Use Cases and Current Status

The previous MPEG column [MPEG154] introduced MPEG’s Gaussian Splat Coding (GSC) exploration, its two tracks (I-3DGS and A-3DGS), and a first set of 27 draft use cases. At the 155th meeting, WG 2 (Technical Requirements) approved an updated document (WG 2 N531) that drops the draft qualifier and adds two use cases.

The additions matter less for their number than their nature: both concern storage and delivery rather than compression. Use case 28 stores a static splat asset in a single file together with a cover image, thumbnail, and audio annotation, so that a capable receiver renders it interactively while a legacy receiver displays the cover image, with no modification of the file. Use case 29 is the temporal counterpart, delivering a dynamic splat track by HTTP adaptive streaming alongside synchronized audio, a pre-rendered video track serving as preview and fallback, and timed text, with representations offered per bitrate, level of detail, spherical harmonics subset, or attribute subset. The application-oriented use cases are unchanged, although the derived requirements are being refactored, notably by separating attribute subset scalability from random access.

GSC remains an exploration activity, but a busy one: 83 input contributions and nine Joint Exploration Experiments re-conducted across WG 2, WG 4 (Video Coding), WG 5 (Joint Video Experts Team, JVET), and WG 7 (3D Graphics and Haptics), with WG 1 (JPEG) now engaged on quality metrics and WG 4 examining Neural Network Coding (NNC) as a compression tool. On the fast track, GS4 (Video-based Point Cloud Compression, V-PCC, Amd. 1) and GS5 (Geometry-based Point Cloud Compression, G-PCC, Amd. 1) progress on reference software and conformance, while the JVET codec track carries five competing proposals. The Call for Proposals (CfP) still has no date, and I-3DGS planning is more mature than A-3DGS. The clearest message from the meeting is that single-frame compression is essentially solved and that the difficulty and the expected gain now lie in dynamic content.

Research aspects: Temporal coding is no longer one open question among many, but the central one, covering inter-prediction for anisotropic primitives, deformation models, and primitive correspondence across frames. Quality assessment remains unresolved: current test conditions score rendered views with peak signal-to-noise ratio (PSNR), structural similarity (SSIM), and learned perceptual image patch similarity (LPIPS) along a pose trace, a fragile proxy for artifacts such as floaters and popping; the involvement of JPEG makes this a good moment to contribute. Use case 29 gives the streaming question a concrete form: when quality varies along several orthogonal axes at once, what is the right abstraction for a representation, and what should an adaptation algorithm optimize?

Call for Proposals on Media Authenticity and Provenance Indication

At the 155th meeting, WG 2 approved a CfP on media authenticity and provenance indication with the MPEG Systems technologies (WG 2 N535), following an exploration phase in WG 3 (Systems) that produced requirements, a gap analysis, and several technical proposals (WG 3 N1842, attached to the call). The scope is the system layer: carriage and signaling of metadata that allows a receiver to verify that a media asset comes from a trusted producer and to convey provenance information, rather than the definition of a provenance language itself.

Three interfaces are addressed, namely (i) elementary streams and non-timed items, (ii) the file format level (ISO Base Media File Format (ISOBMFF) and Common Media Application Format (CMAF)), and (iii) packaged delivery (Dynamic Adaptive Streaming over HTTP (DASH), CMAF, MPEG Media Transport (MMT), and MPEG-2 Transport Stream (MPEG-2 TS)). The use cases span deepfakes and manipulated news media, forgery in insurance claims, surveillance and investigations, labeling of AI-generated content, and legitimate modifications such as editing, transcoding, ad insertion, and archival preservation. Detecting whether an asset is fake without embedded data is explicitly out of scope.

A response may be a complete solution or a single tool addressing one or more requirements, and is evaluated on requirement coverage, computational complexity, the bandwidth needed to convey the authenticity information, compatibility and backward compatibility with existing MPEG standards, and extensibility, with self-evaluation tables provided in the annexes. Proposals are submitted as input contributions to the 157th meeting in Brisbane, Australia, by January 13, 2027. Review at that meeting will produce one or more working drafts or new work item proposals, and the preliminary development plan targets Committee Draft (CD) or Committee Draft Amendment (CDAM) at the 158th meeting, Draft International Standard (DIS) or Draft Amendment (DAM) at the 159th, and Final Draft International Standard (FDIS) or Final Draft Amendment (FDAM) in early 2028.

Research aspects: The requirements make this more interesting than a signing exercise. A conventional signature breaks on any bit change, yet the call demands that verification survive transcoding, dropped scalability layers, representation switching, splicing, late binding, and loss of frames or audio, and that the unmodified remainder of a presentation stay verifiable after an authorized edit. Designing verification structures with that granularity and quantifying their overhead per segment against the added latency in live scenarios is an open problem at the intersection of cryptography and streaming systems. A second question is joint verification: establishing that a given audio track was the one the producer intended to accompany a given video, including their synchronization, cannot be achieved by signing each file separately. Finally, interoperability with provenance schemes defined elsewhere, and the question of what a receiver should do with a verification result, leave room for work on usable trust signaling.

Joint Call for Proposals on Video Compression with Capability Beyond VVC

The draft of this call was described in the previous MPEG column [MPEG154], and JVET has now issued the final version (JVET-AQ2021 [JVET-AQ2021]), approved at its 43rd meeting in Geneva in July 2026. The target is video coding technology that significantly exceeds Versatile Video Coding (VVC) in compression, implementability, applicability across content types, and features such as latency, robustness, and scalability, benchmarked against the VVC Main 10 profile. Four test cases are defined: one for improved compression without runtime limits, and three in which the aggregate encoder run time is constrained to 5x, 1x, and 0.2x that of the VVC Test Model (VTM) anchor. All are evaluated over seven categories covering (i) standard dynamic range (SDR) random access at ultra-high definition (UHD) and (ii) at high definition (HD), (iii) SDR low delay HD, (iv) high dynamic range (HDR) with perceptual quantizer (PQ) and (v) with hybrid log-gamma (HLG) transfer functions, (vi) gaming, and (vii) user-generated content. A separate track invites technology offering additional functionality, together with proposals on how its benefit should be assessed.

Formal subjective testing uses degradation category rating, with objective results reported as PSNR, multiscale SSIM (MS-SSIM), and, for PQ content, weighted PSNR. The schedule is tight: anchors have been available since May 2026, registration runs from August 1 to September 1, 2026, and the main package of bitstreams, reconstructed sequences, and binaries must reach the test coordinator on physical media by October 26, 2026. Subjective assessment then runs until late December, blind cross-checking by other proponents is mandatory, and proposals are evaluated at the 45th JVET meeting in January 2027, with an initial test model selected during 2027 and the standard targeted for October 2029. Participation is not free: up to EUR 20k per test case is charged to cover the hiring of test subjects.

Research aspects: Two design choices in the call are worth attention. First, a supplemental set of sequences is disclosed to proponents only after the decoder binaries have been submitted, and results on it are due six weeks later, which turns the call into a held-out generalization test. Read together with the ban on training on test sequences and the obligation to disclose training material, this makes out-of-domain behavior of learned coding tools measurable at scale, and reporting it well is a contribution in itself. Second, run time is aggregated as the sum over threads, which measures total compute rather than latency and therefore reads very differently for a massively parallel or GPU-resident design than for a sequential one. How to characterize the rate, distortion, and complexity trade-off fairly across such architectures, and how to value functionality such as scalability or error resilience against a plain bitrate gain, are open questions the call poses rather than answers.

Joint ITU-T SG 21 and ISO/IEC JTC 1/SC 29 Workshop on Media Streaming Services

On July 14, 2026, ITU-T SG 21 and MPEG Systems held a joint half-day workshop in Geneva, collocated with their meetings, on “Media Streaming Service, What’s next” (call for presentations, program). The premise was that after the transitions from analog to digital, enabled by MPEG-2 Systems, and from broadcast to over the top (OTT), enabled by MPEG-DASH, integrated networks, edge computing, and AI are driving another shift, and that both bodies wanted industry views before committing to new work. On challenges, Netflix spoke on the shortcomings and evolution of ISOBMFF, Bitmovin on where streaming trends meet the container, ETRI on future streaming services, and Huawei on ultra-low latency communication and streaming. On opportunities, 5G-MAG addressed standards and open source, DVB heterogeneous networks, and 3GPP SA4 delivery beyond OTT. The closing session was an open mic with questions and answers from both speakers and the audience.

The wrap-up set the existing systems standards, namely MPEG-2 TS, ISOBMFF, DASH, and CMAF as well as the volumetric, scene, and augmented reality (AR) formats, plus ongoing work on authenticity, SyfAI (see below), common metadata, and Gaussian splats, against what industry actually asked for: low latency, low overhead, AI-driven media, and open source software. The resulting agenda is an exploration of a better or new container format, an analysis of overhead, processing-friendly metadata, including JSON and ontology issues, WebCodecs integration, and ingest. Just as notable are the conclusions about process: more open access and industry involvement, faster turnaround, possibly a new home or outlet for this work, and software with interoperability testing from day one rather than reference software at the end. Three ad hoc groups were set up: (i) on MP4, DASH, and file formats, reviewing ISOBMFF against emerging transports and the delivery of AI input and output data; (ii) on vision, analyzing bottlenecks and requirements beyond ISOBMFF; and (iii) on working methods.

Research aspects: Two of these items are directly researchable. Container overhead is asserted more often than measured, and as segments shrink toward frame level for low latency, the ratio of container to media bytes grows; a careful comparison across ISOBMFF, CMAF, and object-based transports at equal latency would inform the exploration rather than follow it. Media over QUIC (MOQ) is the bigger change, since it replaces the segment with the object as the unit of delivery and pulls streaming and real-time communication into a single design space, which reopens rate adaptation, interaction with congestion control, caching, and relay behavior, and the question of what a container still contributes when the transport itself frames media. Carrying AI input and output data alongside media, finally, links this work to coding for machines and to semantic streaming.

Exploration on Systems Technologies for AI-based Media Standards (SyfAI)

Several MPEG coding standards now put a neural network inside the decoder, among them video and feature coding for machines, AI-based point cloud coding, and neural network coding. The previous MPEG column [MPEG154] described that coding side through the MPEG-AI vision document. SyfAI, an MPEG Systems exploration started at the 153rd MPEG meeting, asks the complementary and much less glamorous question: what does the surrounding infrastructure have to do so that such content can actually be stored, delivered, and played back interoperably? The underlying shift is that a media file has traditionally been self-contained, meaning that a conforming decoder and the bitstream are sufficient. Once the decoder depends on a trained model that may be selected, delivered, or updated separately, that assumption no longer holds, and the system layer has to say how a player learns which model it needs, how that model reaches it, and how both sides can be sure they are using the same one. Work so far has advanced on the most concrete piece, namely storing video coding for machines content in the MPEG file format, while proposals for carrying compressed neural networks were sent back for clearer use cases. The more interesting development is a new thread on an AI update framework, opened jointly with WG 7, which collects five topics: (i) a format for AI parameter data, (ii) a manifest for updating parameters, (iii) a repository to serve them, (iv) integrity checking, and (v) bit exactness together with conformance.

Research aspects: Each of those topics is a research problem in its own right. Distributing and updating models alongside media turns into a delivery question that looks familiar but has not been studied: when to fetch an update, how to cache and version models, what a manifest must express about capability and compatibility, and how to treat weights as a second class of asset next to the media. Conformance is harder, because a standard normally guarantees that every decoder produces identical output, whereas neural inference varies with library and hardware, so deciding what conformance means and how to test it remains open and extends the reproducibility concerns already visible in MPEG-AI. A repository of updatable weights is also an attack surface, which is why this work sits naturally beside the media authenticity call discussed above: signing and verifying a model raises the same questions one level down from signing content. And because the choice of where a model lives affects how quickly playback can start and how smoothly a player can switch, the apparently dry container questions have measurable consequences for the quality of experience (QoE).

Concluding Remarks

Taken together, the five items show MPEG engaging early with technologies whose momentum comes from outside the committee. With Gaussian splatting, it is compressing a representation that already has a renderer and a user base, which is a better starting point than earlier attempts at three-dimensional video had, even if wider adoption will depend on capture and display ecosystems as much as on coding efficiency. The authenticity call is scoped pragmatically at establishing origin, which is achievable in the near term and provides the system-level plumbing that regulation and industry are asking for. The call for video coding beyond VVC visibly reflects deployment experience, with runtime-constrained test cases treating practicality as a first-class criterion, while questions of licensing and market uptake sit largely outside MPEG itself. In systems, the joint workshop and the new ad hoc groups show a willingness to re-examine long-standing assumptions as transports such as MOQ emerge, and SyfAI stakes out the system-level questions of the AI era before they become urgent. For our community, the result is an unusually rich set of open problems in evaluation and methodology.

The 156th MPEG meeting will be held in Hangzhou, China, from October 19 to 23, 2026. Click here for more information about MPEG meetings and ongoing developments.

References

  • [MPEG154] C. Timmerer. 2026. MPEG Column: 154th MPEG Meeting. SIGMultimedia Rec. 18, 2, Article 8 (April 2026).
  • [JVET-AQ2021] J.-R. Ohm, M. Wien, F. Bossen. 2026. Joint Call for Proposals on video compression with capability beyond VVC, JVET-AQ2021 (July 2026).

The 4th Edition of Spring School on Social XR, ACM Seasonal School organised by CWI

The 4th edition of the Spring School on Social XR organised by Distributed and Interactive Systems group (DIS) at CWI in Amsterdam took place from 20 to 23 April 2026 and attracted 31 students from various disciplines, including technology, the social sciences, and the humanities. The event was organized by Silvia Rossi, Irene Viola, Thomas Röggla, and Pablo Cesar from CWI, and Omar Niamut from TNO. Also this year, it was co-sponsored by ACM SIGMM, thanks to the founding for Special initiatives, and it has been recognised also as an ACM Europe Council Seasonal School.

Students and organisers of the 4th Spring School on Social XR, 2026

Across 11 lectures (4 of them open to public) and 3 hands‑on workshops led by 12 international instructors, participants had the possibility to have cross-domain interactions on Social XR. The program covered a broad range of topics at the intersection of immersive technology, human behavior and system design: trust and safety in social virtual environments, ethical and human-centered approaches to XR design, affective computing and mental health applications, socially intelligent and AI-driven virtual humans, haptics and tactile interaction, volumetric video technologies, mobility research in XR, and the future of communications in education and research. Together, they provided a multidisciplinary perspective that went beyond knowledge transfer: bringing together early-career researchers from diverse backgrounds to create the groundwork for new collaborations and a growing community around Social XR.

Students presenting themself during first day.

List of talks and workshops:

  • “The Future of Communications and Interaction” by Gül Akcaova and Mark Cole— workshop
  • “Towards Ethical, Human-Centered XR Experiences: From Theory to Practice” by Katrien de Moor — workshop
  • “Volumetric Video Technologies for Real-Time Immersive Communication” by Guillaume Gautier and Alexandre Mercat — workshop 
  • “Studying Mobility in eXtended Reality” by Yan Feng
  • “A Brief History of Virtual Humans” by Marco Gillies
  • “Affective Social XR and Intercorporeal Regulation: A Multi-Method Approach to Supporting Mental Health and Well-being” by Alexandra Kitson
  • “Designing Inclusive XR: Experience from Day-to-Day Industrial Research” by Marta Orduna
  • “Socially Intelligent Digital Humans” by Chirag Raman
  • “Tactile Extended Reality: The Role of Haptics in Immersive Interactive Applications” by Maria Torres Vega
  • “Trust and Safety in Social XR” by D. Yvette Wohn
  • “AI-driven Animation for Virtual Humans” by Zerrin Yumak

JPEG Column: 111th JPEG Meeting

JPEG DNA reaches Draft International Standard stage at the 111th JPEG meeting

The 111th JPEG meeting was held in virtual mode from 13 to 17 April 2026.

This meeting was marked by several major achievements. JPEG DNA, the first image format that uses a quaternary representation suitable for synthesis into nucleotide sequences to be used for storage on DNA support, reached the Draft International Standard (DIS) stage. This was ensured after a successful wet-lab experiment, including DNA synthesis/sequencing that allowed a test of reliability in real conditions.

Furthermore, the DIS stage was also reached for JPEG XE, the first International standard for coding of visual events; JPEG Pleno Light Field Quality Assessment, which establishes a model for quality evaluation of encoded light fields; JPEG Trust Media Asset Watermarking, which provides watermarking support to media asset authenticity; and JPEG Trust reference software, which will provide a valuable tool for the development of applications using this family of standards. Furthermore, JPEG Pleno Light Field coding, 2nd edition, reached the Final Draft International Stage.

The following sections summarize the main highlights of the 111th JPEG meeting.

  • JPEG DNA reaches DIS stage.
  • JPEG XE reaches DIS stage.
  • JPEG Trust part 4 – Reference Software – reaches DIS stage.
  • JPEG Pleno Light Field Quality Assessment reaches DIS.
  • JPEG XS Part 2 prepares a new amendment to define a new raw Bayer profile.
  • JPEG RF explores methodologies for quality evaluation of Gaussian Splatting coding distortions.
  • JPEG AI explores machine-oriented applications in the compressed domain.

JPEG DNA

At its 111th meeting, the JPEG Committee reached a major milestone with JPEG DNA Part 1 (ISO/IEC 25508-1) advancing to the Draft International Standard (DIS) stage. This achievement marks the culmination of a multi-year standardization effort, beginning with the Final Call for Proposals at the 99th JPEG meeting and proceeding through the creation of the Verification Model at the 102nd meeting, the first Working Draft at the 103rd meeting, the Committee Draft, and the study DIS text produced at the 108th meeting. With DIS now reached, the core technical specification for the efficient coding of images into quaternary representations suitable for archival on DNA synthetic polymers is frozen for the first edition of the standard. The JPEG Committee expects publication of Part 1 an International Standard before the end of 2026.

The DIS milestone was accompanied by the successful completion of wet-lab experiments initiated at the 109th JPEG meeting and whose synthesis and sequencing phases were carried out between the 110th and 111th meetings. The sequenced results from the independent parties have now been delivered to the JPEG Committee and analyzed. The experiments confirmed that the information recorded in the DNA solution was successfully extracted and that the corresponding compressed images were correctly decoded, in accordance with the end-to-end workflow defined in the specification, including biochemical constraints, the recovery mechanism used to cope with the noise introduced by the synthesis and sequencing processes, and the procedures for reading the encoded image information back from the synthesized oligonucleotides. These results provide concrete validation that the current specification of JPEG DNA satisfies the conditions required for applications relying on the current state of the art in DNA synthesis and sequencing and represents the first end-to-end demonstration of the standard on real biological media.

JPEG XE

The joint effort between ITU-T SG21 and ISO/IEC JTC1/SC29/WG1 on JPEG XE is ongoing, with JPEG XE Part 1 ready to become the first of a series of standards for coding of visual events. During its 111th meeting, the JPEG Committee made substantial progress on the ongoing development of the additional JPEG XE Parts 2, 3, 4, and 5. Technical discussions advanced the work on Part 2, which will define Profiles and Levels, as well as a normative buffer model to ensure safe and interoperable decoder operation across implementations. Further enhancements were made to the code of the open source reference software, which will be published under Part 3 as a proof‑of‑concept and conformance testing implementation. Preparatory work commenced for Part 4, which will establish the structure and scope of formal conformance testing and will become the focus during the coming months. Finally, consensus was reached about Part 5, where the JPEG Committee converged on an ISOBMFF‑based file format design supporting efficient storage, streaming, and precise event timestamping. The JPEG Committee, mandated for the joint effort between ISO/IEC/JTC1/SG29 and ITU-T SG21, remains committed to the development of a comprehensive and industry-aligned standard that meets the growing demand for event-based vision technologies. This collaborative approach underscores a shared vision for a unified, international standard to accelerate innovation and interoperability in this emerging field. The JPEG XE joint AHG (ITU-T SG21 and ISO/IEC JTC1 SC29 WG1) was reestablished to continue the JPEG XE standards development. If you are interested, please consider joining this public joint AHG.

JPEG Trust

There have been major advances in all parts of JPEG Trust. Recently, a 2nd edition of JPEG Trust Part 1 – Core Foundation, was published to accommodate assertions for attributions (based on The Dublin Core metadata element set, ISO 15836-1:2017) and declarations of rights (based on the W3C ODRL Information Model 2.2). JPEG Trust Part 2 – Trust profiles and reports extends the framework specified in Part 1 with extended functionalities for Trust Profiles and Trust Reports; it has now entered the DIS stage. JPEG Trust Part 3 – Media asset watermarking, incorporating specific tools and associated assessment methodologies for usage scenarios that rely on invisible watermarking of media assets, has now been released for DIS balloting. The DIS for JPEG Trust Part 4 – Reference Software was also released for balloting; this Part will be extended in the future with additional implementations and test datasets.

A first JPEG Trust Summit is planned on 12-13 November 2026 in London. More information on upcoming events related to JPEG Trust can be found here.

JPEG Pleno

The JPEG Pleno Light Field activity analyzed the DoCR for the Final Draft International Standard (FDIS) of the 2nd edition of ISO/IEC 21794-2 (“Plenoptic image coding system (JPEG Pleno) Part 2: Light field coding”). This 2nd edition integrates AMD1 of ISO/IEC 21794-2 (“Profiles and levels for JPEG Pleno Light Field Coding”) and includes the specification of a third coding mode entitled “Slanted 4D Transform Mode” and its associated profile. It is expected that at the 112th JPEG meeting this new edition will become an International Standard (IS).

The JPEG committee made significant progress in JPEG Pleno Light Field Quality Assessment (Part 7) activities. Key developments were achieved, including the review and finalization of the DIS text for ISO/IEC 21794-7, which was approved for registration at this meeting. Core experiments on light field quality assessment were presented and discussed, contributing to the validation of subjective testing frameworks and the further development of objective metrics, including improvements to learning-based approaches.

JPEG XS

JPEG XS, the image and video compression format for transmitting visually lossless, high-quality pictures with minimal latency and low resource consumption, is currently in a mature state. Two amendments, JPEG XS Part 1 and Part 2, are expected to be published in July of 2026. JPEG XS AMD1 provides a new syntax to allow embedding of user-defined metadata at the slice-level in addition to the frame-level that was already possible. JPEG XS Part 2 amendment AMD1 provides additional levels and sublevels, as well as a new frame buffer level, to better support proxy-stream extraction use cases. At its 111th meeting, the JPEG Committee prepared another amendment (AMD2) for JPEG XS Part 2 to define a new raw Bayer profile, called CMainBayer, that supports reducing the implementation complexity compared to the MainBayer profile. It will be an inclusive profile, meaning that existing HighBayer and MainBayer implementations will be capable of decoding the newly proposed CMainBayer profile. A core experiment was issued to verify the proposed parameters and the impact on the coding performance and the potential resource savings. The plan is to issue the DAMD2 text at the next JPEG meeting in July 2026.

JPEG RF

At its 111th meeting, the JPEG committee continued its active exploration of technologies for radiance field representations. Significant progress was achieved, particularly in the areas of quality assessment and visualization methods for radiance fields, where an exploration study was presented and discussed, contributing to a deeper understanding of how radiance field content should be rendered and evaluated in subjective experiments. These discussions support the development of reliable and practical assessment frameworks in preparation for a future Call for Proposals. JPEG also initiated a new exploration focusing on the identification of suitable objective quality metrics and subjectively annotated datasets, as well as the evaluation of metrics’ performance for radiance field content. This work builds on prior efforts covering coding methods, subjective assessment protocols, and trajectory design, and aims to support the development of standardized evaluation methodologies.

In parallel, outreach activities continued to expand community engagement and support the preparation of dissemination materials, while work progresses toward revising the JPEG RF Common Test Conditions. The outcome of these efforts will guide future standardization in radiance field coding.

JPEG AI

At its 111th meeting, the JPEG Committee continued advancing JPEG AI. The Committee reviewed progress in key JPEG AI Core Experiments, including energy-consumption analysis, and bit-exact reconstruction, with results showing promising implementation trade-offs across CPU, GPU, and FPGA platforms. The FPGA implementation was identified as the most energy-efficient, while GPU execution provided the lowest latency. JPEG also reported progress on compressed-domain machine analysis, where JPEG AI latent representations enabled earth-observation segmentation with nearly ten times fewer model parameters. Further work will continue through new Core Experiments on energy consumption, bit-exact reconstruction, and objective and subjective evaluation. This activity reinforces JPEG AI’s role as the first International Standard for end-to-end learning-based image coding and as a technology relevant to both human visualization and machine-oriented applications.

Final Quote

“At the 111th meeting, five JPEG standardization projects reached DIS stage, a testimony to the commitment of its experts who work tirelessly towards the development of efficient standardized solutions that guarantee interoperability between emerging imaging products and services,” said Prof. Touradj Ebrahimi, the Convenor of the JPEG Committee.