Web harvesting services for Library of Congress digital preservation.
The Library of Congress requires preservation web harvesting services to capture and archive web content for its digital collections. The contractor will provide technical infrastructure and expertise for large-scale web crawling, data extraction, and storage. The scope includes managing the harvesting pipeline, ensuring compliance with archival standards, and delivering curated web archives. The contract is expected to be a multi-year engagement with high technical complexity, requiring experience in web archiving and digital preservation. The Library of Congress serves the U.S. Congress and the American public, and the work will be performed in Washington, DC.