← all stories

Before the Web · Era 2 · Into the Internet

1990–2003Bunyip shut down in 2003

Archie

The world's first program that searched for files across the whole internet on its own: in 1990 a student at McGill University in Montreal wrote a script that went through anonymous FTP archives around the world and compiled a list of what was stored there, a job that had been done by hand before. Two years later the authors tried to sell access to the service to companies, but with the arrival of the web and Google the program was no longer needed and died in the late 1990s.

Founders Alan Emtage (graduate student and systems administrator at McGill, author of the scripts) · Bill Heelan (co-author of the first implementation) · Peter (J.) Deutsch (Emtage's manager in the McGill IT department, later president of Bunyip)
Domains archie.mcgill.ca (the original, no longer working) · archie.serialport.org (retro reconstruction, 2024)
searchftpinfrastructure-no-monetizationacademic-to-startup

In brief

Archie was the world's first program that searched for files on the internet on its own, in place of a person: in 1990 Alan Emtage, a graduate student at McGill University in Montreal, wrote a script that went through thousands of anonymous FTP archives around the world and compiled a single list of what was stored there. Before that, the only way to find out which of hundreds of servers held the file you needed was by hand or on someone else's tip. In 1992 Emtage and his boss Peter Deutsch turned the tool into one of the first internet companies, Bunyip Information Systems, and tried to charge providers for a license. A few years later the web arrived, and with it Google and AltaVista, and indexing nothing but file names on FTP no longer made sense: Bunyip shut down in 2003. Neither Emtage nor Deutsch got rich, because back then it occurred to nobody to patent the idea.

How it started (the founders)

Alan Emtage was born in 1964 in Barbados, into a family that encouraged curiosity: his aunt, a teacher, took him to listen to BBC science programs and to explore Carlisle Bay. In 1981, on a trip to Great Britain, he got a Sinclair ZX81, his first computer. In 1983, after winning a Barbados Scholarship, he enrolled at McGill University in Montreal; he first considered meteorology and geology but chose computer science as the safer career in the middle of the 1983 recession. After earning his bachelor's degree in 1987, he stayed on for a master's and at the same time took a job as a systems administrator in the faculty's IT department, and that is where he ended up reporting to Peter Deutsch, who by then had spent several years pressing the McGill administration for a proper internet connection (the first line to Boston cost about $35,000 a year).

Emtage's job was to track useful free software on anonymous FTP servers around the world by hand; Deutsch himself jokingly called him the department's resident rat. In June 1990, by his own account, Emtage got tired of logging in to every server manually: "Rather than spending my time logging on to FTP sites... I wrote some computer scripts that would do the same thing, and much faster too." The program was named Archie, from "archive" without the "v." When Deutsch saw the internal tool, he suggested making it public; the reaction of colleagues: "Everybody was like, 'Oh my God — of course! Why didn't we think of this?'" A third co-author, Bill Heelan, worked on the first implementation together with Emtage, but it was Deutsch, as the boss, who after the fact became the public face of the project and, two years later, the president of the company.

Year-by-year timeline

Lesser-known but significant facts

  1. Emtage did not work alongside Deutsch as an equal but reported to him. The public image of three student co-authors simplifies the real hierarchy: Deutsch was the systems manager and Emtage's boss, and he gave him the task as a routine work duty rather than inviting him in as a co-founder from the start.
  2. Growth was set off by a single Usenet post. Archie was conceived as a purely internal tool; everything changed when it was mentioned in passing on an open mailing list, and a flood of requests from outsiders poured in, forcing the team to bolt on a public front-end in a hurry.
  3. The company name Bunyip was chosen out of nostalgia, not for marketing. Deutsch is half Australian; the bunyip is a creature from Aboriginal mythology and the hero of a children's book he read as a boy. The name, as he put it, simply stuck in his head.
  4. The load was designed from the start as a federation rather than as one server for the whole world. Early on, Archie servers in different countries (Finland through FUNET, Australia through AARNET) crawled only their own region and exchanged ready-made copies of the data with each other: peer replication, not a naive crawl of the entire planet by every node.
  5. The creators themselves rejected X.500, the official international directory standard, as overly complex and political, and chose instead to build their own lightweight protocol, WHOIS++, through the IETF, a deliberate bet on simplicity against the consensus heavyweight.

Legend vs. the record

The first growth lever

The growth lever was not marketing or a launch but chance: Archie was conceived as a strictly internal McGill tool, and Deutsch says so in plain words: "we mentioned it in a Usenet posting once, got flooded with postings with people asking us to do searches. So we stuck a frontend onto it, and that was sort of it." One post on an open mailing list, and the team was building out an interface in a rush to meet the demand that poured in. The same pattern of one community thread setting off growth shows up today on Reddit; in 1990 the venue simply was Usenet.

The second, less obvious lever was that growth did not bring down the infrastructure. Deutsch stresses: "we were actually very aware of this from the beginning," and from the start the team built a federated network of mirrors. Finland (FUNET) and Australia (AARNET) crawled only their own region, and data spread between nodes as copies, not through every server crawling the whole world again. Without that foresight, the viral effect of a single post could have killed a lone server in Montreal, instead of producing a worldwide network of ~25 nodes and ~100,000 queries a day by 1993.

The paired story

The pair is WAIS. Archie and WAIS solved the same problem of 1990–1991, finding something on the internet before the web existed, in opposite ways. Archie indexed only the NAMES of files on FTP servers and knew nothing about what was in them, mechanically similar to a library catalog without annotations. WAIS (Thinking Machines, 1991) searched the contents of the documents themselves and was built from the start as a partnership of Thinking Machines, Apple, Dow Jones, and KPMG Peat Marwick (see the WAIS dossier). Their commercial fates diverged as well: Bunyip sold licenses to providers and was dissolved in 2003, while WAIS Inc. was bought by AOL in 1995 for $15 million in stock, and with that money Brewster Kahle built the Internet Archive (see the WAIS dossier). The pair illustrates the fork between searching by name and searching by content well, the same choice that faces any modern indexer of other people's content (see the parallels below).

Parallels today (projects from the catalog)

What a builder can take from this in 2026

  1. An internal tool built for your own laziness is a legitimate source of a product. Emtage was not designing a search engine; he was automating a boring routine for himself, and that is exactly why the solution turned out simple and immediately useful to many people.
  2. One post in the right place can replace an entire marketing plan. Archie's growth began not with a launch but with a chance mention on Usenet: the demand already existed, and all it took was one public signal.
  3. Design for load in advance, not after the first server crash. Federated replication by region was thought through from the very start, not added in a panic after the first surge of traffic.
  4. Freely available and free are different things, and it is fine to explain that to users directly. Deutsch refused to apologize for the move to licensing: "letting people starve is a crime," while charging a reasonable fee is not.
  5. Being first does not guarantee money; execution and timing matter more. Emtage acknowledges that not patenting was a decision of its era ("nobody was making money off the Internet") and has no regrets: it is part of the same chain of decisions as the invention itself.

Discrepancies and what we could not verify

Sources (primary first)

Primary:

Secondary:

Internet History Keeper

Every historical dossier here is free and open, and it will stay that way. If you want the series to keep going (new dossiers, checks against primary sources, the English edition), support it with a subscription for $9.90 a month. It is support, not access, and you can cancel at any time.

Become a keeper — $9.90 a month →

The same thing, about today

The same breakdowns, but of projects launching right now: what the product is, where the first users came from, how they charge. The card is free, the full dossier is $5 (the dossier itself is written in Russian).

Browse the dossier catalog →