The web is often treated as a durable record of institutional and social life, yet in practice it is fragile, revisable, and frequently ephemeral. Domains change, redesigns erase earlier material, institutions relocate, maintainers graduate, platforms impose silent limits, and periods of political instability can interrupt digital access entirely. This paper argues that archiving should not remain a niche activity practiced by a few specialists at the margins, but should become a proactive part of website maintenance. I motivate this claim through a case study centered on the Pakistan Embassy International School and College Tehran, whose domain, visual identity, leadership, and physical location all changed within a short period after my graduation. In response, I built and deployed a lightweight automated archival system using Python and GitHub Actions to submit pages and media from the site to the Internet Archive's Wayback Machine. The project shows both that archival preservation can be automated with modest infrastructure and that archival systems are themselves vulnerable to interruption, as illustrated by GitHub's automatic disabling of scheduled workflows after repository inactivity. Drawing on personal experience with internet shutdowns in Iran, open-source sustainability lessons from RPI's RCOS, and the operational history of the archiver, I argue that the ephemerality of the web is not an exception but a structural condition. If digital societies wish to preserve institutional memory and public history without leaving preservation to chance, proactive archiving should become a commonplace part of website maintenance.


翻译:万维网常被视为机构与社会生活的持久记录,然而在实践中,它脆弱、易变且常常转瞬即逝。域名会变更、网站改版会抹除早期内容、机构会搬迁、维护者会毕业、平台会施加静默限制,而政治不稳定时期则可能完全中断数字访问。本文认为,存档不应再是少数专家在边缘地带从事的专门活动,而应成为网站维护中的一个主动环节。我通过一个以巴基斯坦驻德黑兰使馆国际学校及学院为中心的案例研究来阐述这一论点:在我毕业后不久,该机构的域名、视觉标识、领导层及实际地点均在短期内发生了变化。为此,我利用Python和GitHub Actions构建并部署了一套轻量级自动化存档系统,将该网站中的页面与媒体提交至互联网档案的时光机。该项目表明,存档保存既可通过适度基础设施实现自动化,同时存档系统本身也易受中断影响——例如,GitHub在仓库处于非活跃状态后自动禁用预设计划工作流。结合我在伊朗经历网络中断的个人体验、来自伦斯勒理工学院RCOS项目的开源可持续性经验教训,以及该存档系统的运行历史,我论证了网页的短暂性并非例外情形,而是一种结构性状况。如果数字社会希望保存机构记忆与公共历史,而不将保存工作交由偶然性决定,那么主动存档应成为网站维护中的普遍实践。

0
下载
关闭预览

相关内容

【WWW2024教程】时间网络挖掘,附486页slides
专知会员服务
36+阅读 · 2024年5月23日
《网络态势感知问题和挑战》31页报告
专知会员服务
31+阅读 · 2023年6月18日
异质信息网络分析与应用综述,软件学报-北京邮电大学
最新《动态网络嵌入》综述论文,25页pdf
专知
37+阅读 · 2020年6月17日
亿级订单数据的访问与存储,怎么实现与优化?
码农翻身
16+阅读 · 2019年4月17日
被动DNS,一个被忽视的安全利器
运维帮
11+阅读 · 2019年3月8日
ResNet, AlexNet, VGG, Inception:各种卷积网络架构的理解
全球人工智能
20+阅读 · 2017年12月17日
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
13+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
20+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
2+阅读 · 2014年12月31日
VIP会员
最新内容
非对称防御中的自组织临界性:俄乌战争
专知会员服务
1+阅读 · 今天14:36
《战争中的大语言模型监管》
专知会员服务
2+阅读 · 今天14:26
边缘计算的军事应用
专知会员服务
8+阅读 · 8月9日
一种考虑资源机动性的武器目标分配混合算法
专知会员服务
9+阅读 · 8月8日
相关基金
国家自然科学基金
1+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
13+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2015年12月31日
国家自然科学基金
20+阅读 · 2015年12月31日
国家自然科学基金
0+阅读 · 2014年12月31日
国家自然科学基金
2+阅读 · 2014年12月31日
Top
微信扫码咨询专知VIP会员