The clip count floor now follows the length of the source: at least five under twenty minutes, nine from twenty to forty, thirteen beyond, instead of five everywhere. Voice activity detection also moved to run alongside transcription and scene detection, which it does not depend on. Measured the same day on real jobs: five clips in eighteen minutes forty eight before, then thirteen clips in ten minutes eighteen and twenty two clips in eight minutes thirty six.