Where to Find Old Ship Logbooks: Archives, Digitised Images, and the Condition They Arrive In
Where old ship logbooks are held by fleet and institution, how digitised images vary in condition, and why catalogues, images, climate datasets and transcriptions are not the same thing.
Leo Team
September 17, 2026
Contents
This is a working guide to where to find old ship logbooks — which institution holds which fleet's records, what is digitised, and what the images are actually like when they reach you. If your research depends on the Remarks column as much as the weather columns, the difference between a catalogue record, an image, a climate dataset, and a transcription will decide how your project runs.
Old ship logbooks are not held in one place. They sit with the institution that generated or received them — the Admiralty series at The National Archives in Kew, US Navy deck logs at NARA, East India Company journals at the British Library, VOC journals at the Nationaal Archief, whaling logs in New England museums, meteorological logs in national weather archives — so finding one means identifying the fleet, the flag, and the operating company before you search a catalogue. A significant share is digitised, but "digitised" covers everything from a preservation-grade colour capture to a bitonal microfilm scan. A catalogue record is not an image, an image is not a transcription, and a climate dataset built from logbooks is not the logbook's text. Plan for a distributed search and for images of very uneven condition.
That last point is where most projects lose time. What follows maps custody, then works through the three separate things that determine whether the pages you get can actually be read.
Why custody is dispersed, and how to search it
A logbook is a chronological operational record. NARA's own definition is a useful anchor: Navy logbooks, also called captains' logs or deck logs, are chronological entries documenting the daily activities of a ship or unit, arranged by date and, within each day, by time.
Because they were working records, custody follows the organisation that ran the ship. There is no universal maritime catalogue, and there will not be one. The practical consequence: your first research question is not "where are the logbooks?" but "whose ship was this, and which body received its paperwork?" A Royal Navy frigate, an East Indiaman, a Dutch VOC retourschip, a New Bedford whaler and a merchant vessel reporting weather to a national meteorological office all end up in different buildings, described under different rules.
This is also why aggregators are a discovery layer and nothing more. Internet Archive, Europeana, DPLA and HathiTrust are worth searching, but treat any hit as a pointer: go to the item record, identify the holding institution, and work from that institution's own catalogue for reference, extent, and rights.
The major custodial bodies
Royal Navy: The National Archives (UK)
TNA's holdings are organised by who kept the log, which matters because the same voyage may survive in several parallel versions. The research guide to Royal Navy ships' log books sets out the main series: ADM 51, captains' logs, 1669–1852; ADM 52, masters' logs, 1672–1871 (the guide also directs searches for 1672–1840); and ADM 53, ships' logs, from 1799 onwards, after the officers' logs were superseded by a single ship's log inspected weekly by the captain and forwarded to the Admiralty.
The masters' logs are usually the fuller source for anyone after position and weather: they record course, position and weather alongside punishments and the ship's working routine. Captains' logs were compiled from them, with whatever the captain thought worth adding.
Note also that lieutenants' logs went elsewhere — to the National Maritime Museum. Search the Caird Library at Royal Museums Greenwich as a parallel custodian, not a duplicate. And TNA states plainly that most records described in that guide are not available online; expect to order copies or visit Kew for a good deal of pre-1900 material.
US Navy: NARA
NARA's Navy logbook holdings span the American Revolutionary War to the late twentieth century, with the largest concentration in Record Group 24, Entry 118, and Special List #44 providing an item-level list of logbooks of US Navy ships, stations and units, 1801–1947. Deck logs continue to arrive from the Naval History and Heritage Command thirty years after each calendar year closes.
Digitisation is in progress rather than complete: NARA describes a multi-year project to digitise deck logs, with completed series appearing in the National Archives Catalog and the Vietnam-era tranche producing millions of images. Check the online lists (through 1940, and 1941 and later) before assuming a reproduction order is necessary.
East India Company: British Library and the Qatar Digital Library
EIC officers kept detailed journals, and the surviving series are IOR/L/MAR/A (1605–1705) and IOR/L/MAR/B (1702–1856), with IOR/L/MAR/C holding marine miscellaneous records. The Qatar Digital Library finding aid to IOR/L/MAR describes what the daily entries contain: arrivals and departures, wind and weather, crew activity, disease and deaths, punishments, encounters with other ships, cargo and treasure movements, and — at sea — latitude, longitude, magnetic variation, courses, and bearings of land, with occasional coastal sketches.
A single catalogue record shows what a "file" actually is. The British Library's entry for IOR/L/MAR/B/1D, the journal of the London, 1763–65, describes one file of 119 folios, kept in six columns — H, Courses, K, F, Winds &c., and Remarks — with digitised images linked via the QDL. That record also carries an access caveat worth noting: digitised items can be viewed online, but all manuscripts and archives must be consulted at the Library in London.
The QDL subset is geographically scoped — journals touching the Gulf and the Arabian Peninsula. It is a substantial digitised body, not a digitised series.
Dutch, Spanish and Portuguese material
The Nationaal Archief's overview of VOC archives, 1594–1814 is the entry point for ships' journals, alongside the inventory of archive 1.04.02. Some scans are unavailable in the viewer at times but can often still be downloaded — worth checking rather than concluding the item is undigitised.
For Spain, PARES is the state portal disseminating descriptive records and digitised heritage; for Portugal, Digitarq searches more than 8 million archival descriptions associated with more than 63 million images. Neither portal, in itself, tells you how much ship-log material is digitised. Both require series-level work in the catalogue.
Whaling and merchant logs
American whaling logs sit in museum collections. A CLIR-funded project at the New Bedford Whaling Museum covers 76 bound volumes containing 88 discrete logs from Nantucket voyages, 1769–1870 — a useful sense of scale for museum digitisation grants. Separately, WhalingHistory.org aggregates data extracted from logs rather than the logs themselves: the American Offshore Whaling Log database draws on 1,381 logbooks from voyages between 1784 and 1920, yielding 466,134 data records, most complete for the nineteenth century.
For merchant weather logs, the UK National Meteorological Archive holds many thousands of worldwide records from Merchant and Royal Navy ships, typically from the mid-nineteenth century onward. The digitised proportion and licensing are not established by the catalogue description; ask.
Four layers that get confused: catalogue, image, dataset, text
A great deal of wasted effort comes from treating these as one thing.
The catalogue record identifies a holding and its extent. It tells you a file exists, its dates, and often its columnar structure. It does not tell you the image exists.
The image layer is what a repository chooses to expose. IIIF is the common delivery standard: the IIIF Image API lets a client request a specified region, size and rotation of an image over HTTP. That is useful for pulling a single column of a wide logbook opening at full resolution. But IIIF governs delivery, not deposit: a IIIF endpoint does not guarantee that the repository exposes the preservation master or permits unrestricted download.
The dataset layer contains extracted observations, not pages. CLIWOC covers 1750–1854, released in original IMMA format with a GeoPackage and OpenDocument bundle also available. ICOADS spans surface marine data from 1662 to the present, with gridded monthly products at 2°×2° back to 1800. RECLAIM's stated objective is to locate and image logbooks and digitise the meteorological and oceanographic observations for merger into ICOADS — a targeted extraction, not a full-text edition. Likewise oldWeather recovers weather observations from Royal Navy, US Navy, Coastguard and Coast Survey logbooks through volunteer transcription, coordinated within the broader ACRE initiative for recovering surface weather observations over the last 250 years.
If your question is climatological, these datasets may answer it directly. If your question is about the Remarks column — a mutiny, a landfall, a sick list, a punishment, a note on a passing sail — the dataset almost certainly does not contain it, and you are back to the images. This is the most common mismatch between what a historian needs and what the data-rescue community has published, and it is worth naming early in a project rather than discovering at month three. The wider historical climate and scientific data rescue effort has grown out of exactly this division of labour: numbers first, narrative later or not at all.
The text layer — a searchable transcription of the page as written — is the rarest of the four. For most series it does not exist, and producing it is your problem.
What condition the images arrive in
Assume nothing from the word "digitised." Three separate causes degrade legibility, and they call for different responses.
Original deterioration
Sea-going volumes have had a hard life. Moisture damage, iron-gall ink burn, foxing, and show-through from the verso are all common. Bleed-through in particular is described in the conservation and imaging literature as one of the most common and invalidating effects affecting documents — and it matters more than aesthetics, because verso strokes visible through the leaf give a recognition system character-shaped evidence that is not part of the recto text. Treat specific damage claims as collection-specific unless a conservator's report or the catalogue documents them.
Capture geometry and lighting
Logbooks are bound volumes, often thick and tightly sewn, and the gutter is where the columns run out. Flattening a bound volume against a platen produces geometric distortion: text near the gutter appears stretched, compressed, or angled, sometimes to the point of being unreadable. Cradle-based capture at a controlled angle reduces the problem without stressing the binding, but you are inheriting whatever decision the digitising institution made, possibly decades ago.
Where capture followed a published standard, you can at least know what you have. The FADGI Technical Guidelines for Digitizing Cultural Heritage Materials, Third Edition and the Metamorfoze Preservation Imaging Guidelines v2.0 both specify measurable capture targets; Metamorfoze defines Full, Light and Extra Light profiles with sampling-rate rules (300 ppi for originals A5 and above), bit depth, colour space, tonal and geometric tolerances, and its governing principle is that all visible information in the original must be visible in the preservation master. No published evidence supports a claim about what percentage of any maritime holding meets these levels — ask the repository which profile a given batch was shot to.
Reformatting history
A large volume of naval and merchant logbook imagery is not a scan of the page. It is a scan of a microfilm of the page, and microfilm was made for legal substitution and long-term storage rather than for machine reading. Bitonal thresholding, tonal collapse, scratches and generational loss all reduce the visual evidence available. When you request material, ask directly: is this direct capture or a microfilm derivative, at what resolution and bit depth, in colour or greyscale or bitonal? The answer changes what you can expect from any subsequent reading, human or machine. Our guide to scanning resolution and bit depth for old documents sets out the specifications that determine whether text is recoverable, and the triage guide to faded ink, foxing and show-through helps separate a capture problem from a recognition problem from a conservation problem.
From image set to readable text
Once you have images, the logbook's own conventions govern what you can do with them. Sea days run from noon, positions may be observed or dead-reckoned, and longitudes may be measured from meridians other than Greenwich — all of which is best sorted out before transcription, not after. Our guide to reading an old ship's log column by column covers those conventions in detail, and the ship logbook transcription guide covers the transcription routes for naval and merchant logs specifically.
On the reading itself: these are Latin-script records in vernacular languages — English, Dutch, French, Spanish, Portuguese — not Latin-language texts. That distinction matters when you evaluate any transcription tool, since what a recognition model can read depends on script, language and hand together, not on a marketing language count. Period spelling and dense nautical abbreviation remain real obstacles regardless of alphabet.
The prevailing route for large logbook backlogs is volunteer transcription, and it works: oldWeather has demonstrated it at scale. Its constraint is that it is selective and slow, and it is generally pointed at the observation columns because that is where the funded value lies. That leaves the narrative untranscribed. If your project needs the full page — every column plus the Remarks — crowdsourced transcription runs out well before your series does.
Machine transcription is the alternative, with one condition attached: on tabular, damaged, mixed-hand material, the failure mode to guard against is not garbled output but fluent output. A general chatbot handed a logbook opening will produce a plausible column of temperatures. Some of them will not be the numbers on the page, and nothing in the output will indicate which. This is the fluent-but-wrong problem in its most damaging form, because a corrupted observation propagates silently into a dataset.
Leo's transcription model, ATR-1, is built for this specific gap: a zero-shot handwritten text recognition model for Latin-script material — whatever the language on the page — that transcribes what is written rather than normalising it, preserving strikethroughs, marginal additions, table structure and archaic spelling instead of smoothing them into modern prose. There is no per-collection model training step, which matters when a series shifts hands every few months of voyage. Around it sits the rest of the workflow: upload a batch of QDL or NARA page images, keep the image beside the transcription for verification, add archival metadata (archive, collection, box, folder, identifier) so a reading is traceable back to its shelfmark, search across the whole corpus, and export to Word, PDF, HTML or TEI. Translation, correction and modernisation are separate operations that write to their own tabs, so the base transcription stays untouched. Two honest limits: complex tabular layouts vary in quality, and pages where dominant printed structure carries dense handwriting — pre-printed log forms — are the model's known weak spot. Test on your own images before committing a series, and keep a human on the numbers regardless.
Before you commit to a series
The verification list is short and saves months. For any candidate collection: confirm the current catalogue record and reference; establish whether images exist and whether you can get derivatives or the preservation master; check the IIIF manifest or download endpoint; read the copyright and reuse terms; determine whether the scan is direct capture or a microfilm generation; get the resolution, bit depth and colour profile; and ask whether the narrative columns were ever transcribed or only the observations.
Do that on a sample of ten openings before ordering ten thousand. The logbooks that survive are a substantial record — a quarter-millennium of daily observation taken at sea, by people with no idea anyone would read it — and the work of getting it back is mostly patient, unglamorous verification of what you actually hold.
Frequently Asked Questions
Where can I find old ship logbooks from the Royal Navy, US Navy, or the East India Company?
Old ship logbooks sit with the institution that generated or received them, so start with the flag and operating company rather than a single catalogue. Royal Navy logs are at The National Archives in Kew, with lieutenants' logs at the National Maritime Museum. US Navy deck logs are at NARA, mostly in Record Group 24, Entry 118. East India Company journals are in the British Library's IOR/L/MAR series, part of it digitised through the Qatar Digital Library. Dutch VOC journals are at the Nationaal Archief, American whaling logs in New England museums, and merchant weather logs in national meteorological archives.
Are old ship logbooks available online for free?
A significant share is digitised, but coverage is uneven and access varies by institution. The National Archives states that most records in its Royal Navy log books guide are not available online, so pre-1900 material often means ordering copies or visiting Kew. NARA's deck log digitisation is a multi-year project in progress, with completed series appearing in the National Archives Catalog. The Qatar Digital Library hosts East India Company journals, but only those touching the Gulf and Arabian Peninsula. Aggregators such as Internet Archive, Europeana, DPLA and HathiTrust are discovery layers — follow the hit back to the holding institution.
What is the difference between ADM 51, ADM 52 and ADM 53?
These are the three main Royal Navy log series at The National Archives, and the same voyage may survive in more than one. ADM 51 holds captains' logs, 1669–1852. ADM 52 holds masters' logs, 1672–1871. ADM 53 holds ships' logs from 1799 onwards, after the separate officers' logs were superseded by a single ship's log inspected weekly by the captain and forwarded to the Admiralty. For position and weather, the masters' logs are usually the fuller source: they record course, position and weather alongside punishments and the ship's working routine, and captains' logs were compiled from them.
Do climate datasets like CLIWOC and ICOADS contain the logbook's Remarks column?
No. These are extraction projects: they contain observations, not pages. CLIWOC covers 1750–1854, ICOADS spans surface marine data from 1662 to the present, and RECLAIM's objective is to image logbooks and digitise the meteorological and oceanographic observations for merger into ICOADS. oldWeather's volunteer transcription is likewise pointed at the observation columns. If your question is climatological, they may answer it directly. If you need a mutiny, a landfall, a sick list or a note on a passing sail, the narrative is not there and you are back to the images.
Why are digitised logbook images often hard to read?
Three separate causes degrade legibility, and each calls for a different response. Original deterioration — moisture damage, iron-gall ink burn, foxing and show-through from the verso — is common in sea-going volumes. Capture geometry adds distortion: flattening a thick, tightly sewn book against a platen stretches or angles text near the gutter, exactly where the columns run out. And much naval and merchant logbook imagery is a scan of microfilm rather than the page, bringing bitonal thresholding, tonal collapse and generational loss. When ordering, ask whether it is direct capture or a microfilm derivative, and at what resolution, bit depth and colour profile.