logo

Standards Manage Your Business

We Manage Your Standards

ISO

ISO 28500:2017

Information and documentation — WARC file format

Standard Details

ISO 28500:2017 specifies the WARC file format:

- to store both the payload content and control information from mainstream Internet application layer protocols, such as the HTTP, DNS, and FTP;

- to store arbitrary metadata linked to other stored data (e.g. subject classifier, discovered language, encoding);

- to support data compression and maintain data record integrity;

- to store all control information from the harvesting protocol (e.g. request headers), not just response information;

- to store the results of data transformations linked to other stored data;

- to store a duplicate detection event linked to other stored data (to reduce storage in the presence of identical or substantially similar resources);

- to be extended without disruption to existing functionality;

- to support handling of overly long records by truncation or segmentation, where desired.

General Information

Status : Published
Standard Type: Main
Document No: ISO 28500:2017
Document Year: 2017
Pages: 26
Edition: 2
  • ICS:
  • 35.240.30 IT applications in information, documentation and publishing

Life Cycle

Currently Viewing

Published
ISO 28500:2017
Knowledge Corner

Expand Your Knowledge and Unlock Your Learning Potential - Your One-Stop Source for Information!

© Copyright 2024 BSB Edge Private Limited.

Enquire now +