In the archive of the Pattern Recognition Association of South Africa (PRASA), the file named prasa02_01.pdf occupies a special place. It is the first PDF in the proceedings of the 2002 annual conference, held that year in Cape Town. For a researcher interested in the evolution of pattern recognition in the early 2000s, this single file offers a snapshot of both the conference's formal structure and the technical preoccupations of its contributors. Typically, the first PDF in a mid-2000s proceedings volume contained either the front matter—title page, committee list, table of contents—or the opening invited paper. In the case of PRASA 2002, it does both: the first few pages present the official cover and organisation details, followed by the first full-length technical submission, a study on feature extraction for handwritten digit recognition using Gabor filters.
The Structure of a Mid-2000s Conference PDF
Opening prasa02_01.pdf in a modern viewer reveals a document built with the tools of its time. The file's metadata indicates it was created on 5 November 2002, using Adobe Acrobat 4.0 (the Distiller version from the early 2000s). The pages are set to A4 size—standard for South African academic publishing—and the fonts are embedded as Type1 outlines, a common practice to ensure portability before widespread use of Unicode and OpenType. The document is not searchable; no hidden OCR layer exists, reflecting the pre-OCR era of PDF production.
The internal structure follows a predictable order:
- Cover page – Conference title, dates, venue (Cape Town Convention Centre), and the PRASA logo.
- Table of contents – A manually typed list of papers, grouped into sessions such as “Document Analysis”, “Biometrics”, and “Speech Recognition”.
- Organising committee – Names of the chair, programme committee, and local organisers, many from the University of Cape Town and the CSIR.
- First technical paper – “Feature Extraction for Handwritten Digit Recognition Using Gabor Filters” by authors from the University of Pretoria.
This layout was typical for PRASA proceedings of the period. A comparison with the Inside the PRASA 2003 Archive shows a nearly identical template, suggesting that the same LaTeX style file was reused year after year.

What the First Paper Reveals About Research Priorities
The technical paper that occupies the bulk of prasa02_01.pdf is indicative of the state of pattern recognition research in South Africa at the turn of the millennium. Gabor filters, a biologically inspired method for texture and feature extraction, were a popular approach in handwritten character recognition. The paper evaluates filter parameters on a small dataset of handwritten digits collected locally—a reminder that many mid-2000s studies used modest, custom corpora rather than large public benchmarks. Other papers in the same proceedings, listed in the table of contents, cover topics such as speaker identification for South African languages, automated fingerprint matching, and retinal vessel segmentation. The emphasis on applied, resource-constrained problems reflects the academic environment of the time, where computational power and labelled data were scarce.
Notably absent are any references to deep learning or neural networks with more than a few hidden layers. The dominant classifiers were support vector machines, hidden Markov models, and nearest-neighbour algorithms. This contrast with later PRASA proceedings—by 2004, papers on convolutional architectures began to appear—makes the 2002 volume a useful baseline for measuring the field's evolution.
PDF Production Workflow in 2002
Creating a camera-ready PDF like prasa02_01.pdf was a multi-step process that required careful coordination between authors and the proceedings editor. Authors were asked to submit their final manuscripts as PostScript files, which were then distilled into PDF using Adobe Acrobat Distiller. The Information for Authors Kits in Mid-2000s Pattern Recognition describes the exact specifications: 300 dpi for images, Times Roman font family, and no embedded hyperlinks (to avoid broken references in offline viewing). The proceedings editor would then merge individual PDFs into a single volume using tools like pdftk or manual page insertion in Acrobat. The resulting file, despite its simplicity, was robust enough to be printed and distributed on CD-ROM to attendees.

Preservation and Access Today
Two decades later, prasa02_01.pdf remains accessible in online repositories such as the PRASA digital archive. However, the file lacks OCR and metadata tagging, making full-text search impossible without external tools. For researchers studying the history of pattern recognition in Africa, this PDF is a primary source that must be read page by page. Its limited accessibility is a reminder of how far academic publishing has come—modern proceedings are typically indexed, searchable, and linked to supplementary materials.
Yet the very simplicity of the 2002 PDF has an advantage: the file size is small (under 2 MB), and it renders faithfully on any device without dependency on proprietary plug-ins. Today, the file sits on servers alongside thousands of others. Without OCR, reading it requires manual scrolling, but that same limitation means you see the paper exactly as the authors intended in 2002—no auto-formatting, no hyperlinks, just the research as it was.
