New Issue: Orbital Catastrophe Ahead? Read Now

Automating History’s First Draft  

Computers can tell what will matter (slightly) better than humans can

Thomas Fuchs

Join Our Community of Science Lovers!

As 2019 draws to a close, prepare for endless roundups of the year's most important news stories. But few of those stories may be remembered by 2039: new research shows the difficulty of predicting which events will make the history books.

Philosopher Arthur Danto argued in 1965 that even the most informed person, an “ideal chronicler,” cannot judge a recent event's ultimate significance because it depends on chain reactions that have not happened yet. Duncan Watts, a computational social scientist at the University of Pennsylvania, had long wanted to test Danto's idea. He got his chance when Columbia University historian Matthew Connelly suggested analyzing a set of two million declassified State Department cables sent between 1973 and 1979, along with a compendium of the 0.1 percent of them that turned out to be the most historically important (compiled by historians decades after their transmission).

Connelly, Watts and their colleagues first scored each cable's “perceived contemporaneous importance” (PCI), based on metadata such as how urgent or secret it had been rated. This score corresponded only weakly with inclusion in the later compendium, they reported in September in Nature Human Behaviour: the highest-scoring cables were only four percentage points more likely to be included than the lowest-scoring ones. The most common prediction errors were false positives—cables that got high scores but later proved unimportant. “I do think there's a kind of narcissism of the present,” Connelly says. “I've been struck by how many times sports fans say, ‘That's one for the history books.’”


On supporting science journalism

If you're enjoying this article, consider supporting our award-winning journalism by subscribing. By purchasing a subscription you are helping to ensure the future of impactful stories about the discoveries and ideas shaping our world today.


Next, Watts says, to approximate an ideal chronicler, the scientists decided to “build the beefiest, fanciest machine-learning model we could and throw everything into it—all the metadata, all the text.” The resulting AI algorithm significantly outperformed humans' contemporaneous judgment. In one statistical measure of its ability to pick out cables later deemed significant, where 1 denotes no incorrect inclusions or exclusions, it scored 0.14, whereas the PCI scored 0.05. Although the algorithm's performance was far from perfect, the researchers suggest that such an “artificial archivist” could help to narrow the field of events to highlight for posterity. When tuned for this purpose, their model weeded out 96 percent of the cables while retaining 80 percent of those that wound up in the compendium.

Emily Erikson, a sociologist at Yale University, who was not involved in the new research, says that despite its use of imperfect data—compendium inclusion was up to the subjective judgment of a few historians, for example—the study offers a practical tool and addresses Danto's hypothesis. “To see a machine-learning empirical test of this conceptual puzzle is really exciting,” she says, “and just kind of fun to think through.”

Subscribe to Support Independent Journalism

Great science journalism requires human expertise, time, effort and creativity. And it costs money. That’s why I and the journalists here at Scientific American hope you’ll join our community.

When you subscribe, you are supporting staff and freelance journalists who are passionate about telling science stories that are true, important and compelling. Our editors and reporters are often experts in their fields, which means they understand the nuances of big discoveries and can untangle the breakthroughs from the hype. With a subscription, you are also supporting rigorous fact-checking to ensure the words we publish are precise and accurate. And you’re supporting original illustrations, graphics and photos that bring you closer to an advanced laboratory, an ice sheet in Antarctica or a space mission in orbit. You’re helping us craft other types of high-quality journalism as well: Our newsletters are carefully written, edited and curated by staffers you have or will come to know and love. Our Science Quickly podcast is based on original reporting, collaboration with editors and scientists and exacting production.

Subscriptions keep this engine running so we can continue to deliver thoughtful, rigorous and independent science journalism to you. In an era of viral misinformation, this work is crucial. If you value what we do, I hope you’ll consider joining us as a subscriber

Thank you,

Jeanna Bryner, Editor in Chief, Scientific American

Subscribe