Usenet Archive Toolkit是一个开源项目,旨在将来自多个来源的Usenet消息处理成统一的、可搜索的档案[1]。该项目由Bartosz Taudul开发,采用GNU Affero General Public License v3或更高版本许可证[1]。
项目提供了一套核心工具和功能[1]。tbrowser文本浏览器是其主要工具,用于读取生成的档案[1]。系统支持从多个数据源导入消息,包括maildir、mbox、NNTP服务器和Google Groups[1]。该工具仅需25MB内存即可处理250万条消息[1],并具备离线工作、消息去重、全UTF-8编码等特性[1]。项目仅支持64位机器[1]。
The Usenet Archive Toolkit (UAT) is an open-source project designed to process Usenet messages from multiple sources into a unified, searchable archive [1]. The toolkit addresses significant gaps left by the unavailability of Google Groups and incomplete data on Archive.org by enabling offline access to archived Usenet content [1].
The project, created by Bartosz Taudul and released under the GNU Affero General Public License v3 or later, provides a core browsing utility called tbrowser, a text-based interface for reading archived messages [1]. The toolkit accepts data from multiple sources including maildir, mbox, NNTP servers, and Google Groups [1]. Its key capabilities include offline operation, automatic deduplication of messages, full UTF-8 encoding support, and efficient memory usage—processing 2.5 million messages requires only 25 MB of memory [1]. The system is designed exclusively for 64-bit machines [1].