The Open and Async audiobook is out. It’s read by a synthetic voice, which means that at some point a machine had to decide how to say GIF, with no help from me. It picked “jif.” Correctly. I have never been prouder of a robot.
In a bit, you’ll hear the audiobook both ways, the same line read by the same robot, and you can settle the argument yourself. Headphones on.
The fight only exists out loud
GIF-versus-JIF is the one internet argument that text is structurally incapable of settling. You can go your whole career typing “GIF” and never once reveal which side you’re on, because the letters look identical in both camps. Type “actually, it’s pronounced jif” into a thread until your keyboard wears out, and nobody hears a thing. The fight only exists in sound. Record audio and it comes for you, wanting an answer before the next sentence. That’s how the audiobook became the first version of this book that could even have the argument. It took my side without being asked.
For the record: soft G. It’s “jif,” like the peanut butter.1 Reasonable, thoughtful, wrong people say it with a hard G. We work together anyway.
There’s a chapter in the book about using chat well, and a section in it about emoji and animated GIFs being a kind of body language for text. A well-placed “that escalated quickly” carries something a paragraph can’t. The book already argues, in print, that GIFs matter. It just never had to say the word out loud. Then I made an audiobook, and it did, eleven times, in one chapter.
I recorded it wrong on purpose
I wrote a whole book about working in the open, though, and the open-and-async answer to a disagreement is not to win the thread. It’s to ship the artifact and let people decide. I built the most literal version of that I could think of. I re-recorded the entire chapter with the narrator forced to say hard-G “gif,” beginning to end, and posted it, free, right next to the correct one. An entire edition in the wrong pronunciation, rendered for the heretics to enjoy their own wrongness. Same words, same narrator, one phoneme of difference. Pick your fighter. Mine already won.
Ten seconds each, the same line both ways: “an urgent message from your manager gets the same screen space as a GIF of a cat playing the piano.”
The pronunciation was never the point. The GIF is. (share this quote)
That’s the point hiding inside the bit. The whole reason the chapter defends GIFs is that the shared reference does the work. Nobody who gets the joke has ever cared how you say the three letters. Arguing about the phoneme is bikeshedding the one part of the thing that carries no meaning at all. It’s impact over input, applied to a debate about peanut butter.
Both cuts live on the book’s feedback repository, which is where readers tell me what’s broken, what’s missing, and what actually landed.
About the robot narrator
People kept asking for an audiobook, to listen at the gym or on a commute, and I’m a big audiobook fan myself. I did some of my best listening on parental leave, one hand on a stroller. I wanted the book to reach people that way too. What I didn’t have was a way to pay for it. A human narration doesn’t pencil out for me right now, not a professional’s four-figure, months-long project, and not the fifty-plus hours it would take me to read a hundred thousand words myself, badly. That left a synthetic narrator, and I’d rather say so straight than have you notice it later. No human voice actor, no studio.
Synthetic is the shortcut, and I’m not going to pretend it’s anything else. (share this quote)
Something real is lost with it. A good narrator doesn’t just say the words, they perform them: the timing, the dry aside, the breath before a hard sentence. My robot reads cleanly. It isn’t an actor, and if you’ve heard a great human narration, you’ll feel the difference. The bigger cost lands past my one book. Synthetic narration takes work from the people who do this for a living, and “it was cheaper for me” is exactly the logic that adds up to their livelihoods shrinking. I don’t have a clean answer to that.
A human narrator for this book is the goal, not a synthetic stand-in forever. If it earns its way there, it’ll get one. Until then, imperfect and shipped beats perfect and imaginary. Ship early, ship often.
Here’s where that leaves me: the voice is disclosed, it isn’t cloned from a real person, and a human (me) listened to all eleven and a half hours before it shipped.2 Machine-narrated, human-checked. If a synthetic voice is a hard pass for you, I get it, and the ebook and paperback let you hear “jif” in your own head, at your own pace.
Get the audiobook
That free chapter is a free sample, and I mean that literally. If you want the same synthetic narrator to read you the whole book, saying “jif” correctly the entire time, the full audiobook is out now. Eleven and a half hours of it, and not one hard G.
I am not going to win this argument. Nobody wins this argument. That’s exactly why you hand people the choice instead of typing one more reply, then go build the next thing. Even if one of the choices is objectively incorrect.
The right edition is the whole audiobook. The other one is just me being generous to the wrong ones.
So, which are you? Choose carefully. My narrator already did.
Footnotes
-
Steve Wilhite created the format at CompuServe in 1987, and when he accepted a lifetime achievement award he stood on a stage and said “it’s pronounced JIF.” The man who invented the thing gets to name it, and the hard-G crowd is out here overruling him. “But graphics has a hard G” is not the counterargument they think it is. It’s an acronym, not a word, and the inventor already ruled. Yes, there’s the peanut butter. Yes, I’m on that side too. I have made my peace with being right. ↩
-
The audiobook has its own QA suite. It transcribes the generated narration back to text, diffs that against the script, and flags mispronunciations, so a stray hard-G “gif” can’t slip into the correct edition. The machine reads. I still get the last ear. ↩