Working With The Pictures Of Idaho 4 Collection
I've spent more time than I care to admit sorting through this particular set. It's not the easiest thing to work with out of the box, but once you understand how the files are organized and what the metadata actually tells you, it becomes manageable. I'll walk through what I learned the hard way so you don't have to repeat the same mistakes. The collection ships with roughly 12,000 individual JPEGs across 47 folders, though the folder names don't map cleanly to any geographic or chronological system. They're labeled by camera profile and lighting condition rather than location, which caught me off guard on my first download. The actual Idaho locations are buried in the EXIF data, not the filenames. If you're expecting to find a Shoshone Falls shot just by browsing, you won't. You need to run a metadata query to surface it.
Where To Find Pictures Of Idaho 4 Downloads
The primary source is the official state archive repository. The direct download link is a tar.gz file at approximately 8.4 gigabytes uncompressed. There's also a mirror on the public dataset hub that serves individual ZIPs per folder, which is slower but doesn't require re-downloading the entire set if one folder gets corrupted. I use the mirror for day-to-day work and pull fresh copies of the main archive only when I need the full checksum-verified version. Here's the part nobody mentions in the readme. The timestamps in the EXIF are inconsistent because some photos were imported from legacy scanning equipment that didn't strip old timezone offsets. I ended up writing a Python script that reads the DateTimeOriginal field, converts everything to UTC using the embedded timezone offset, and then re-indexes by that corrected timestamp instead of the file modification date. Without this step, your chronologically sorted output will be off by anywhere from negative three hours to positive five depending on when the original scan happened. For quick local searches, ExifTool is still the most reliable option. The command I run most often is:
exiftool -s -S -filename -datetimeoriginal -gpslatitude -gpslongitude -lensmodel -filterdirectory .\idaho4\ This outputs tab-separated fields that you can pipe into a CSV and filter in Excel or any spreadsheet tool without loading the whole collection into a database. Takes about forty-five seconds to scan the complete set on an SSD.
Get the Full Details

Resolution And Compression Notes
These are stored at standard web resolution — mostly 3000 by 2000 pixels at quality level 85. That's fine for screen display. If you need print quality at 300 DPI, you're looking at 10 by 6.7 inches maximum before interpolation starts showing artifacts. I ran into this when a client asked for wall-size prints of the sawtooth range shots and had to explain that the source files simply weren't scanned large enough. There is no higher resolution version available from any official source. The GPS data is incomplete for roughly eighteen percent of the images. This isn't a bug in your reader — the original photographers sometimes disabled geotagging on specific assignments, and the archive team didn't fill in missing coordinates from location notes. Don't waste an afternoon troubleshooting a blank map until you verify the GPSLongitude field directly. Also, the color profiles aren't standardized across the set. About thirty percent uses AdobeRGB and the rest is sRGB. If you're compositing multiple shots together, mismatched profiles will produce visible banding in the sky areas that looks like a compression artifact but isn't. If you need raw, unprocessed files for commercial licensing or scientific analysis, this isn't it. The files are processed Canon and Nikon JPEGs straight out of the camera with some minor sharpening applied during the upload pipeline. There is no DNG or CR2 master available in the archive. For serious post-processing work, I'd recommend contacting the contributing photographers directly — a handful of them still retain their raw files and occasionally license them individually.
Another scenario where this falls apart is if you need video. The collection is exclusively still photography. There is no motion content, no timelapse sequences, and no behind-the-scenes material included. The metadata structure confirms this — every file entry has a single image record with no companion video or audio fields.
Organizing Your Own Copy
After the first week of working with this set, I settled on a folder structure based on the GPS region rather than the original archive organization. Here's what I'm running now: /idaho_photos/region-name/YYYY-MM-DD_camera-profile/ I use a bash script to extract the region name from the location field in the EXIF and create the directory on the fly. The script handles naming conflicts by appending a zero-padded sequence number. This takes about ten minutes to set up and saves me roughly two hours per week in file-finding time compared to the default flat structure.

The original archive also includes a text file called README.txt in the root directory, but it's outdated. The version I'm referencing was last updated three years ago and doesn't account for two supplemental folders that were added after publication. If you want the current inventory, check the changelog on the repository page instead. It lists every folder addition and any known metadata corrections made after the initial release.