
A gallery opening doesn't fail on audio because nobody cared about sound. It fails because the plan for capturing it lived in someone's head instead of on a checklist: which mic goes where, what's running as backup, and when the level check actually happens. None of that requires expensive gear. It requires knowing the room, testing it before the first guest walks in, and treating audio as a setup with steps, not an afterthought bolted onto the camera plan.
A usable on-site audio setup for a gallery opening or artist talk needs the right mic for each source (lavalier for speakers, shotgun or room mic for ambience), a second independent recorder as backup, a pre-event read of the room's acoustics, and a level check finished before doors open, not during the event.
Which microphone actually works for a gallery opening or artist talk?
The right microphone depends on the moment, not a single default: a clip-on lavalier for anyone giving remarks, a directional shotgun mic for ambient room sound and crowd texture, and a fixed room mic as a wide safety layer. Most gallery events use two of the three at once, not just one.
A lavalier clips to a collar or lapel and stays close to the mouth no matter how someone turns their head, which is why it's the right call for an artist giving a short introduction or a curator doing a formal walkthrough. It isolates the speaker's voice from room noise, so a lav is the first mic to grab whenever there's a single, identifiable voice that needs to come through clean.
A shotgun mic, mounted on the camera or a short boom, picks up a narrower cone of sound in front of it. That's useful for an artist answering questions from a seated audience, where a lav on one person would miss the room's questions, or for ambient texture, clinking glasses, conversation hum, footsteps on a hard floor, that a lav filters out on purpose.
A fixed room mic, left running from a stable position for the event, exists as a wide net. It won't sound as clean as a lav or as directional as a shotgun, but it captures the room's soundscape continuously, useful as reference audio if either other mic fails or misses a moment. Choosing one mic over the others usually isn't the real decision; deciding which two to run together, and where, is.
Why is a second, independent recording device non-negotiable?
Backup recording exists because a single point of audio failure at a live event can't be reshot. A second recorder running independently, on a different device, different battery, and ideally a different mic, means a dead battery, a bumped cable, or a corrupted card on the primary system doesn't erase the only usable audio from the night.
Every part of an on-site audio chain can fail quietly. A lavalier's battery can die thirty minutes into an hour-long talk. A cable can work loose from a camera's mic input without anyone noticing until playback. A memory card can corrupt mid-record. None of these failures make noise on site; they just show up later as silence where there should be a voice.
A gallery opening or artist talk only happens once. There's no reshoot scheduled for the following week if the primary audio drops out twenty minutes in. That's the real argument for a backup system: not that it improves quality, but that it protects against the one failure mode that can't be fixed after the fact.
In practice, backup recording is usually a second lav or a small handheld recorder running on its own battery and storage, positioned close enough to the same source to serve as a usable substitute if the primary fails. Treat it as part of the setup, not an afterthought: fresh batteries before doors open, a level check on both channels, and a spot check partway through if the event runs long.

How do you read a gallery room's acoustics before setting up?
Reading a room means walking it before guests arrive and noticing what will fight the audio: hard floors and bare walls that bounce sound, glass surfaces that add a hard reflection, high ceilings that thin voices out, and HVAC or refrigeration units humming steadily underneath everything. Each of those changes where a mic goes, not whether the room can be worked.
Hard floors, tile, polished concrete, bare wood, and bare walls reflect sound instead of absorbing it, which is why a voice recorded in an empty gallery can sound harsher and more echoey than the same voice recorded in a living room. Large glass windows or glass partitions add a specific problem: a hard, flat reflection that can create a faint slap-back or ring on close mics if a speaker is positioned near the glass. High ceilings, common in loft-style and converted gallery spaces, thin out a voice by giving the sound more room to disperse before it reaches a mic, so the same mic placement that works in a low-ceilinged room can sound distant in a taller one.
HVAC systems and refrigeration units near a bar setup add a low, constant hum that a room mic picks up continuously and that can bury quieter moments if it isn't accounted for. None of these problems mean the room can't be used. They mean the walkthrough before doors open should include listening, not just looking, and adjusting mic placement and type based on what the room does to sound, not how it looks.
Why does the level check happen before doors open, not during the event?
A level check before doors open means testing input levels against the loudest and quietest moments the event will produce, a raised voice during remarks, a quiet conversation, room noise once the space fills, while there's still time to adjust gear without interrupting anyone. Doing it live means guessing at settings while guests are already watching.
An empty room and a full room don't sound the same. A gallery that reads as acoustically dead fills up with conversation, footsteps, and clinking glassware within the first twenty minutes of an opening, and levels set for the empty room usually clip once that happens. The fix is simulating the loudest realistic moment in advance, someone speaking at normal volume from where remarks will happen, background noise approximating a full room, and adjusting gain before a single guest walks in.
This also means there's room to fix a bad setup. A lav picking up clothing rustle, a shotgun mic aimed wrong, a mic placed too close to a speaker, these are fixable with fifteen minutes and no audience watching. Discovered twenty minutes into a live talk, the same problems mean an audible fix or letting bad audio run for the rest of the event.
The habit that makes this reliable is treating the pre-doors window as production time, not downtime. A quick line read, a walk-through at normal event volume, and a check of both channels takes minutes and removes most of the guesswork that would otherwise happen live, in front of the people the recording is meant to serve.
What changes between covering a reception and covering an artist talk?
A reception's audio setup spreads across the whole room to catch ambient conversation and short exchanges, usually a room mic plus a roaming lav on the host. An artist talk narrows everything to one or two speaking positions, with a lav on each speaker and a shotgun or room mic aimed at the audience for laughs, applause, and questions.
A standard opening reception doesn't have a single audio focal point. Guests circulate, conversations start and end in clusters, and the goal is usable ambient sound plus the occasional clean exchange, an artist explaining a piece to a visitor, a toast, a short thank-you. That calls for coverage over precision: a fixed room mic capturing the general soundscape, and a lav on whoever's most likely to speak, that can be handed off if the center of activity shifts.
An artist talk or panel flips that priority. There's a defined speaking position, sometimes two with a moderator, and the setup should treat that position as the one thing that has to be captured cleanly. A lav on each speaker becomes the primary source, not a supplement, because the recording's value depends on that voice being intelligible. A second mic, shotgun or room, covers the audience for reaction and floor questions, useful for context but not the priority.
The practical result is that a reception's setup tolerates more compromise if something goes wrong, while a talk's setup has almost no room for it. Knowing which one is happening before arrival changes how much time gets spent on a single speaking position versus general room coverage.

How do multiple sound sources get handled during a single event?
Multiple sound sources get handled by running separate channels for each one, a lav per speaker, a shotgun for ambient texture, a backup recorder as a safety net, rather than trying to blend everything into one mic. Each channel gets checked and adjusted on its own; combining them into a clean final mix happens later, in the edit, not live on site.
A single event can generate several sound sources at once: brief remarks, a question from an interviewer, ambient crowd noise reacting to it, and a backup channel running as insurance. Capturing all of that on one mic forces compromises, a mic close enough to isolate the speaker misses the crowd reaction; a mic catching the room misses the clarity the speaker's voice needs.
The more reliable approach records each source on its own channel and blends later. A lav on the speaker stays dedicated to that voice regardless of what else is happening. A shotgun or room mic runs independently, capturing ambient sound without carrying dialogue. A backup recorder does the same job as the primary, without depending on other channels working correctly.
This separation matters most in moments that produce good footage later, an artist pausing while the room laughs, a question shouted during a Q&A, someone off-mic adding a comment the room channel picks up. None of those need to be captured perfectly live. They need to exist as separate recordings an editor can balance later, an easier problem to solve after the fact than live, with one mic doing every job at once.
How does the audio actually sync to the camera footage in post?
Audio syncs to camera footage using either a shared timecode between devices, a physical clap at the start of each recording, or waveform matching software that lines up the audio track from a separate recorder with the camera's built-in audio. Most event shoots use a clap or slate moment as the simplest, most reliable method.
Every recorder running separately from the camera produces its own audio file, which means someone has to line that file up with the picture before the footage is usable. Shared timecode, where every device runs on the same clock, is the most precise, but it requires gear that supports timecode jamming and setup before the event starts, not always practical for a small gallery shoot with borrowed or mixed equipment.
The simpler, far more common method on a small event is a clap, an actual hand clap in frame near the start of each recording, or a slate if one's used. That single sharp sound creates a visible spike on the waveform and a visible motion on camera, giving an editor a precise point to line up in post using nothing more than editing software's waveform view.
Some editing software can also auto-sync using waveform matching, comparing the camera's scratch audio against the recorder's cleaner track and aligning them automatically without a clap. That works when the camera's internal mic picked up something usable, but it's not something to rely on alone; a manual clap at the start of every take or recorder restart costs seconds and removes any dependency on software guessing correctly.
What does the full on-site audio setup look like from start to finish?
A full setup runs in order: walk the room for acoustic problems, place the primary mic based on what's covered, start a backup recorder on separate power and storage, run a level check before doors open, and confirm a sync point on camera before the first guest arrives.
Arriving early enough to walk the room matters more than any piece of gear. That walkthrough is where hard floors, glass, high ceilings, and HVAC hum get identified, and where the event's format decides how mics get positioned. This happens with the room empty and quiet, the only time it's possible to hear these problems before guests add their own noise.
Mic placement follows from that walkthrough: a lav on whoever's speaking, a shotgun or room mic covering the wider space, positioned to avoid the issues just identified. The backup recorder gets set up alongside the primary, on its own battery and storage, close enough to serve as a real substitute, not a token second device in a bag.
The level check comes next, testing against a realistic loud moment, a normal speaking voice from the remarks position, background noise approximating a full room, adjusted before a guest walks in. Last is a sync point: a clap or slate in frame, confirmed on the camera and every recorder, giving post-production a clean reference.
None of this is complicated on its own. What makes it reliable is doing every step before doors open, in order, rather than compressing them into a live event's first few minutes with no room left to fix what gets missed.
Frequently asked questions
Do you always need three microphones at a gallery event?
Not always. Many small gallery events run well on two, a lavalier on whoever's speaking and a fixed room mic for ambient sound, plus a backup recorder. A third dedicated shotgun mic becomes worth adding when there's a Q&A, a moderated panel, or a room large enough that one wide mic can't cover both the speaker and the audience clearly. The backup recorder, though, is never optional regardless of how many primary mics are running.
What's the single biggest mistake in event audio setup?
Testing levels live instead of before doors open. A gallery sounds completely different empty versus full of guests, and levels set for a quiet, empty room usually distort or clip once conversation, footsteps, and glassware fill the space. The fix is simple: simulate a realistically loud moment before anyone arrives, a normal speaking voice at the actual remarks position, and adjust gain then, when there's still time to fix a bad setup without an audience watching.
Can a phone or a basic recorder work as backup audio?
Yes, and it often does. A backup doesn't need to match the primary system in quality, it needs to run independently on its own power and storage and stay close enough to the same source to be a usable substitute if the main recording fails. A phone's voice memo app, or a small handheld recorder, both work fine for this role as long as someone actually checks that it's running and has enough battery and storage for the full event.
How much extra time does proper audio setup add to an event?
Usually 15 to 30 minutes before doors open, on top of any camera setup already happening. That covers walking the room for acoustic problems, placing mics, starting the backup recorder, and running a level check against a realistic loud moment. It's not extra time added for its own sake; it's the window where mistakes get caught and fixed before an audience is in the room, which is the only point in the event where that's still possible.
What happens if a Q&A or applause blows out a lav mic?
A lav clipping on a sudden laugh or applause is a limiter problem, not a placement problem, and it's exactly why the level check happens against the loudest realistic moment beforehand, not just a normal speaking volume. Setting gain with some headroom above a calm speaking voice, and letting the wider room mic pick up crowd reaction instead of relying on the lav for it, keeps a single burst of noise from distorting the whole recording.
One night, two openings, no dropped coverage.
We plan clustered opening nights before they turn into a scramble, whether that means a split schedule, a second shooter, or an early enough booking that the calendar never becomes the problem.
Explore event coverage →