The method and system disclosed herein provide for detecting duplicate information contents such as emails, before storing them in the system, in a fast and reliable way. A parameter that uniquely represents each information content may be determined, and the comparison process of the information contents may be efficiently carried out on the parameters, rather than on the actual information contents.
Redundant Email Address Detection And Capture System
A mechanism for automatically detecting such unwanted messages in real time, which is referred to herein as the Redundant Email Address Detection And Capture System (READACS) is disclosed. The READACS program assumes that incoming mail files have been written to disk in a file format that is somewhat consistent and/or predictable by the programmer. The task of READACS is to identify those email message files, locate an address-of-origin within those files, identify whether the email message should be considered spam, separate spam and non-spam email messages logically, and physically move or rename (or both) those email messages as desired by the programmer.