- How does Eddie identify who is speaking?
- Eddie uses two signals. It separates the voices in the audio and labels each line by speaker. On a multicam shoot, it also watches the mouths on each camera while each speaker talks, and it matches each voice to the camera that shows that person.
- Why use faces as well as voices?
- Voice labels say who is talking, but not which camera shows them. Faces answer that. In our tests on a two-camera interview, the current version of Eddie labelled 99.0% of speaking time with the right speaker, against 93.3% for the earlier version. On a three-camera podcast it was 99.8% against 83.1%.
- Does it work on one camera?
- Yes, for voices. Eddie separates the speakers and labels every line. The camera matching applies to multicam shoots, where Eddie needs to know which camera shows which person.
- What if Eddie gets a label wrong?
- Rename the speaker, move the lines to the right person, or merge two labels. Renaming and merging speakers is free. Eddie never overwrites a label you set by hand, so your correction stays through later automatic passes.
- What about a wide shot that shows two people?
- When both people are clearly in frame, Eddie labels the camera with both names, from left to right. If it cannot tell, it leaves the camera unlabelled and asks you. Your cuts still follow each speaker's voice, and for a vertical clip Eddie crops each segment onto whoever is speaking.
- Does speaker identification change my exports?
- It changes the cut. Where you export a multicam edit to the Premiere Pro project or the DaVinci Resolve project, the multicam clips and the angle Eddie chose for each cut come with it. FCPXML carries the same cuts as flat clips.