Free Press Houston

In Retrospect · Corpus analysis

Fifteen Years of Houston in Print:What the Free Press Houston Archive Contains

A description of the recovered corpus — articles, issues, bylines, subjects, places and objects — with every figure computed from the archive's own records.

Every figure on this page is computed from the archive's records when the page is served; the corpus they describe was last reconciled on 2026-09-16. Where a count depends on a matching rule, the rule is printed beside it. The same figures are available as datasets.

Author
Omar Afra
Published

Production note. This piece was written collaboratively by Omar Afra and AI Mind Control. On method.

Free Press Houston cover, May 2004
Free Press Houston cover, May 2005
Free Press Houston cover, January 2006
Free Press Houston cover, January 2007
Free Press Houston cover, March 2008
Free Press Houston cover, January 2009
Free Press Houston cover, January 2010
Free Press Houston cover, February 2011
Free Press Houston cover, June 2012
Free Press Houston cover, January 2013
Free Press Houston cover, January 2014
Free Press Houston cover, January 2015
One held cover for each year the archive holds a digitised issue, in order. Every cover shown is a scanned or supplied original catalogued as a visual object.
  1. 01 Free Press Houston, May 2004 — issue cover
  2. 02 Free Press Houston, May 2005 — issue cover
  3. 03 Free Press Houston, January 2006 — issue cover
  4. 04 Free Press Houston, January 2007 — issue cover
  5. 05 Free Press Houston, March 2008 — issue cover
  6. 06 Free Press Houston, January 2009 — issue cover
  7. 07 Free Press Houston, January 2010 — issue cover
  8. 08 Free Press Houston, February 2011 — issue cover
  9. 09 Free Press Houston, June 2012 — issue cover
  10. 10 Free Press Houston, January 2013 — issue cover
  11. 11 Free Press Houston, January 2014 — issue cover
  12. 12 Free Press Houston, January 2015 — issue cover

1. What “recovered” means

The Free Press Houston Archive is a recovered record, not a complete one. The paper published monthly from 2003 to 2018; the archive's issue index reserves 203 monthly slots for that run and holds a digitised object — a scan, a press proof or a supplied photograph — for 62 of them. Everything else the archive knows about the paper's contents comes from 4,142 article records, of which 4,122 were recovered from the paper's former website and 20 were transcribed from printed issues. The web records themselves come from two generations of site: 3,578 from the wordpress-era site and 544 from the earlier blogger-era site. “Recovered” therefore means retrieved from a surviving copy and reconciled — for title, byline, date and text — against whatever other evidence the archive holds, not reproduced from a publisher's master file. There is no such file.

Reconciliation is most visible in the dates. The migrated content system behind the former website printed the year of retrieval on most pages instead of the year of publication: 3,876 of the recovered records carried such a date, and 1,664 carry a rewritten year inside their own text. The archive corroborates a publication date from other evidence where it can — original archive paths, dated newsletters, printed issues, references in the text — and otherwise leaves the record undated. The result is 3,072 dated records (74%) and 1,070 undated ones (26%); the date-reconciliation dataset records the basis for each. Only 66 records are placed on the pages of a held printed issue; the rest are dated but not paginated.

2. The run in issues

Free Press Houston cover, June 2010
Free Press Houston, June 2010 — issue coverFree Press Houston cover — June 2010 A held scan from the middle of the run.Object record · Issue 2010.06Cover of the scanned original held for the June 2010 issue.
Free Press Houston cover, July 2014
Free Press Houston, July 2014 — issue coverFree Press Houston cover — July 2014 One of the twelve months of 2014, the densest year of digitised issues.Object record · Issue 2014.07Cover of the scanned original held for the July 2014 issue.

The digitised issues run from 2005.05b to 2015.10 and fall in 11 calendar years (2005, 2006, 2007, 2008, 2009, 2010, 2011, 2012, 2013, 2014, 2015); 60 of the 62 are complete. Covers are held for 97 issues, more than twice the number of full issues, because the paper reproduced its own covers online:

Cover images held, by source
SourceIssuesShare
From a held scan or proof6264%
From the paper's own Flickr reproduction3435%
Supplied photograph of the printed cover11%

One cover per issue, classified by the source the issue record names.

Digitised issues held, by year
  1. 20058
  2. 200617
  3. 20075
  4. 20083
  5. 20092
  6. 20106
  7. 20113
  8. 20122
  9. 20134
  10. 20147
  11. 20155

Years without a digitised issue are omitted. The paper published in every year of the run; the gap is in the archive's holdings, not the paper's.

3. Articles by year

The recovered corpus is not spread evenly across the run. 580 records fall in 2016, the largest single year, and the three years 2014, 2015, 2016 together hold 1,672 of the 3,072 dated records. The early years are represented by a handful of records each. That shape reflects the paper's web publishing — which grew through the run and was what could be recovered — far more than it reflects the size of the printed paper in any given year, and it should be read as a fact about the archive.

Recovered article records by year of publication
  1. 20043
  2. 20056
  3. 20061
  4. 2007168
  5. 2008288
  6. 200986
  7. 201054
  8. 201180
  9. 2012122
  10. 2013175
  11. 2014550
  12. 2015542
  13. 2016580
  14. 2017402
  15. 201818
  16. Undated1,067

A record is placed in a year by its corroborated publication date or, failing that, by the printed issue it sits in. Nothing is inferred from the text.

By year — dated, bylined, titled as interviews, digitised issues
YearRecordsDatedBylinedInterviewsIssues held
20043320
200566118
2006110017
20071681686005
200828828814923
200986863062
201054542326
201180806003
20121221226272
2013175175122114
201455054744577
201554254246985
201658058054650
201740240239645
20181818172
Undated1,06746528

“Interviews” here means records whose own title says so (Title begins with “Interview:”, ends in “Interview” or “Q&A”, or contains “an interview with” or “the … interview”.). It is a title census, narrower than the evidence-based interview index kept for individual contributors.

4. Who wrote it

2,847 of the 4,142 records (69%) carry a byline the archive could resolve to a contributor record; 1,295 do not, either because the recovered page carried no byline, because it carried a house credit, or because the name could not be matched with confidence. The bylines resolve to 195 distinct writers. The contributor directory holds 243 records in all: 195 with recovered bylines and 48 evidenced only by a staff box, an artwork credit or a festival roster. 10 records carry a byline naming more than one person.

The distribution is steep. The three most frequent bylines account for 1,380 records, 48% of everything bylined, and the 15 listed below for 75%. Most of the remaining writers are represented by a handful of pieces — which is, again, a description of what the website preserved rather than of who filled the printed paper. The transcribed staff boxes, by contrast, name 72 distinct people under 28 role headings across 9 issues (2010.08, 2012.11, 2013.08, 2014.07, 2014.08, 2014.12, 2015.01, 2015.03, 2015.10), and many of those names have no recovered byline at all.

Most frequent recovered bylines
ContributorRecordsShare of bylined
David Garrick55620%
Michael Bergeron50818%
Ramon Medina31611%
Elizabeth Rhodes1134%
Alex Wukman893%
Harbeer Sandhu883%
Jack Daniel Betz763%
Russel Gardin682%
Jef Rouner632%
Meghan Hendley Lopez522%
Nick Cooper492%
Mariam Afshar452%
Marini Van Smirren401%
Amanda Hart341%
Kyle Nazario341%

Bylines are counted after the archive's contributor curation, which folds variant byline strings onto one record where the evidence supports it. See the contributor and byline census and the masthead census.

5. Sections and subjects

The former website filed most records under a section. 37 section labels survive; 581 records carry none. The largest sections are the arts desks — Music, Film, Art — with the civic sections (Editorial, Local and State, News, Politics) forming a second tier. One label, Featured, is a placement category rather than a subject and is reported as found. A section label named for the paper's festival, FPSF, appears on its own, as does FPTV, the label the site used for Free Press TV, the video strand whose creative director appears in the later staff boxes.

Section labels, as recovered
SectionRecordsShare of sectioned
Music80923%
Film60817%
Featured43612%
Entertainment41112%
Art37411%
Editorial1534%
Local And State1534%
News1133%
Politics892%
Debris772%
Uncategorized481%
FPSF451%
Food And Drink391%
FPTV301%

The 14 largest of 37 labels. Labels are reproduced as the source system spelled them.

Subject tags are looser. 3,648 records carry at least one tag, and between them they use 7,354 distinct strings, of which 6,077 occur on a single record. The most frequent tags repeat the section structure, which is a property of how the site was run rather than a finding; the archive does not normalise them, so “Local And State” and “Local and State” are counted separately, as they were filed.

Most frequent subject tags
TagRecordsShare of tagged
Music1,07429%
Film92725%
Entertainment43012%
Featured40211%
Art3459%
News2868%
Local And State2748%
Featured Event2136%
Editorial1644%
Houston1334%
Uncategorized1183%
Local Events1093%
Politics832%
Debris 2772%
Free Press Houston632%
FPH592%
Video562%
Food And Drink531%
Technology481%
Local and State461%
Featured Coverage451%
Sports Desk391%
Podcast381%
Comedy371%

6. Entities and places

The archive registers a small number of entities — people, organisations, places and events — for which it holds evidence beyond a passing mention; 25 are currently evidenced, and each has a record in the entity directory. A record is joined to an entity only where the entity's registered name or a registered alias appears in the record's own subject or entity fields. The counts are therefore conservative and describe tagging, not coverage.

Registered entities most often named in record metadata
EntityRecords
Houston, Texas153
Free Press Houston107
Free Press Summer Fest77
Day for Night29
Westheimer Block Party13
Kendrick Lamar7
Omar Afra7
Pegstar6
Run The Jewels5
Sylvester Turner3
Oneohtrix Point Never2
Alex Czetwertynski1

Neighbourhoods are harder to count honestly, because the paper rarely tagged them. The table below matches a fixed list of Houston place names against record titles and subject tags only. A record counts once per place when its title or any subject tag contains the place name. Body text is not searched. A place absent from the list is not counted, and a match is a mention, not a classification of what the piece is about.

Houston places named in record titles or tags
PlaceRecords
Montrose40
Westheimer21
Downtown19
The Heights9
Discovery Green8
Galveston7
East End / EaDo6
Third Ward3
Fifth Ward2
Galleria2
Eleanor Tinsley Park1
Midtown1
Rice Village1
Washington Avenue1

Body text was not searched, so the count understates every place on the list; the ranking, not the totals, is the finding.

7. Interviews across the corpus

By the title rule stated above, 169 records are interviews, 141 of them dated, conducted by 33 different bylined writers. The form belongs overwhelmingly to the last years of the run: the year table in section 3 shows the count rising sharply in 2016 and 2017, when the paper's website was carrying question-and-answer pieces at a rate the earlier years do not approach.

Writers with the most titled interviews
ContributorInterviewsShare
Russel Gardin5331%
Elizabeth Rhodes2012%
Michael Bergeron127%
Kwame Anderson106%
Omar Afra95%
Will Guess42%
Amanda Hart32%
Josh Bosarge32%
Mills-mc Coin32%
Rob McCarthy32%

Title census only. For one contributor the archive keeps a stricter, evidence-based interview index that also admits question-and-answer transcripts and editor's notes; the companion essay In Conversation is built from it.

8. Festivals and events

Stage guide printed for the October 11, 2008 edition
Westheimer Block Party, 2008 — scheduleStage guide printed for the October 11, 2008 edition. Event ephemera catalogued as an object and gathered under its collection.720 × 972 px · Object record · Westheimer Block Party, fall 2008 · Westheimer Block Party · SourceRecovered from FPH archival file, 2008.
Free Press Summer Fest 2013 final site map · May 24, 2013
Free Press Summer Fest, 2013 — mapFree Press Summer Fest 2013 final site map · May 24, 2013. Houston — Eleanor Tinsley Park, event date June 1–2, 2013; drawn at 1" = 100' by Brungardt Enterprises, L.L.C.; areas 1–6 with the Saturn, Neptune, Mars, Jupiter, Venus and Mercury stages and festival headquarters. Production documents make up most of the festival collections' object counts.3031 × 2167 px · Object record · Free Press Summer Fest 2013 · Free Press Summer Fest · SourceRecovered from /fpsf/2013/Site-Map-FPSF-5-24-13-FINAL.pdf.
Final site map for Day for Night 2015 at Silver Street Studio, Sawyer Yards, Houston
Day for Night, 2015 — mapFinal site map for Day for Night 2015 at Silver Street Studio, Sawyer Yards, Houston. Four sheets, drawn by Brungardt Enterprises, L.L.C., revision dated December 14, 2015 and circulated December 16, 2015 as the map to be used from that point. Supersedes the earlier R9 map of December 4, 2015. First sheet shown; the original is held in full. The Day for Night collection is held mainly as production and promotional files.Brungardt Enterprises, L.L.C.2491 × 1591 px · Object record · Day for Night 2015 · Day for Night · SourceRecovered from /day-for-night/2015/Site-Map-FINAL-Day-for-Night-12-15-15.pdf.

Five subjects are treated as collections because the archive holds more than articles about them — edition records, posters, production files, artwork. Their article counts are small against the corpus as a whole, which is worth stating plainly: the paper's festivals are heavily documented in the archive's objects and edition records, and comparatively lightly in its recovered articles.

Collections — recovered articles and established editions
CollectionArticlesEditionsEdition years
Free Press Summer Fest9592009—2017
Westheimer Block Party1982005—2009
Day for Night3232015—2017
Fitzgerald’s: The Pegstar / Free Press Houston Years, 2010–201636
HOU Decide222015
Performer Registry0
Worst of Houston7

An edition is counted only where a held document establishes it. The Westheimer Block Party's collection record dates its launch to 2005; the established editions currently run 2005—2009. Worst of Houston is an annual feature, not an event, and has no edition rows.

9. The visual record

Artwork by Blake Jones for the Free Press Houston article “False Positive: An Introduction to Houston’s Prosecution Problem”
Blake Jones — artwork for “False Positive: An Introduction to Houston’s Prosecution Problem”A commissioned illustration with a credited artist, held with its article record.Art by Blake JonesObject record · False Positive: An Introduction to Houston’s Prosecution Problem · Blake JonesArtwork held with the recovered article “False Positive: An Introduction to Houston’s Prosecution Problem”, which prints the credit “Art by Blake Jones”.
Artwork by Austin Smith for the Free Press Houston article “Wasting the San Jacinto River Waste Pits”
Austin Smith — artwork for “Wasting the San Jacinto River Waste Pits”Illustration credits are recovered from the printed credit line and resolved to contributor records.Illustration by Austin SmithObject record · Wasting the San Jacinto River Waste Pits · Austin SmithArtwork held with the recovered article “Wasting the San Jacinto River Waste Pits”, which prints the credit “Illustration by Austin Smith”.

Free Press Houston commissioned its covers and much of its illustration from working artists, and the archive holds that work in two forms. First, as credits: 138 artwork and illustration credits have been resolved to 15 contributor records, with many more credit lines still unresolved, as the artwork-credits dataset states. Second, as objects: 159 registered archive objects — covers, posters, maps, schedules, billboards and printed ephemera — each with its own provenance and fixity record.

Registered objects, by type
TypeObjectsShare
Cover6340%
Illustration3522%
Poster2314%
Photograph128%
Map117%
Schedule64%
Printed ephemera64%
Billboard21%
Object11%

The article records themselves carry 1,690 held images across 996 records, and reference 1,523 further image addresses that no longer resolve and that the archive lists as lost rather than substituting. Recovered text survives for 4,111 records — 73,426 paragraphs in all. Beyond the paper's pages the archive holds 598 items in its photograph and video staging collection (76 of them video), 37 catalogued videos and 4 published audio recordings.

10. The apparatus

Every figure above is exported. The data catalog publishes 10 datasets in CSV and JSON, generated from the same records as the pages; 11 corrections and revisions are logged at /about/corrections; and 982 held files carry SHA-256 checksums.

Published datasets
DatasetRows
Article metadata census4,142
Contributor and byline census243
Issue census203
Masthead and staff-box census307
Artwork and illustration credits141
Event edition records22
Publication-date reconciliation4,101
Former-domain observation ledger5
File fixity (SHA-256)982
Corrections and revisions log11

11. What is not here

A researcher or a language model reading this page should take the following as the archive's own statement of its limits. There are no digitised issues before 2005.05b or after 2015.10, and the paper's first five years are represented by 178 article records in total. A quarter of the recovered articles are undated and will stay so unless new evidence appears. Nearly a third carry no resolvable byline. The entity register is small and its counts measure tagging. There are no circulation figures, no attendance figures, no financial or corporate records and no editorial correspondence in the public archive, and no figure on this page should be read as implying one. The paper's former website remains online outside the archive's control and displays the corrupted dates described in section 1; the archive's statement on the former domain explains the difference.

Cite this page with its URL and the reconciliation date shown at the top. The archive's methodology describes how records are recovered, dated and attributed.