A Study on the Efficiency of Film Music Composition Education Using Generative AI - Focusing on a Comparative Case Study of Composition-Major and Non-Major Learners -
본 연구는 생성형 AI 도구를 활용한 영상음악 작곡 교육이 장르 분석 기반의 AI 비활용 교육과 비교하여 어떠한 가능성과 한계를 갖는지 확인하는 데 목적이 있다. 연구 방법으로 탐색적 성격의 질적 비교 사례연구를 채택하고, 사례 ①로는 작곡 전공 대학생 12명을 대상으로 15시간 동안 진 행된 대학 영상음악 작곡 수업, 사례 ②로는 비전공 고등학생 12명을 대상으로 수노(Suno)를 활용 하여 3시간 동안 진행된 고교생 전공체험 프로그램을 분석하였다. 두 사례는 영상 소재와 산출물 형식, 교수자가 동일하다. 분석 결과, 비전공 고등학생들은 장면 분석 정형문과 프롬프트 공식, DAW 최소 조작 세트 등의 교수적 지원을 바탕으로 전원이 3시간 안에 영상에 결합한 음악을 완성 하였으며, 그 산출물은 교수자의 질적 판단 기준에서 영상 분위기와의 부합, 음악 배치의 적절성, 결과물 완성도의 세 관점 가운데 일부 측면에서 전공 대학생의 산출물과 비교 가능한 양상을 보였 다. 두 사례의 교육 시간(15시간과 3시간) 차이는 교육 내용과 운영 방식이 서로 다른 조건에서 관찰된 것일 뿐 일반화된 효과는 아니지만, 생성형 AI와 구조화된 교수적 지원이 결합될 때 교육 의 중심이 구현 훈련에서 장면 분석과 음악적 판단·선택 훈련으로 이동할 수 있음을 시사한다. 이를 토대로 본 연구는 5단계의 잠정적 교육 절차를 제시하였다. 다만 생성 음원의 정밀한 동기화 한계와 교수자 1인의 질적 판단이라는 제약이 있으므로, 이 결과는 전문 교육의 대체가 아닌 입문 단계의 교육 방안으로서 의의를 갖는다.
This study examined the possibilities and limitations of film music composition education using generative AI. An exploratory qualitative comparative case study compared two classes that used the same videos and format: a 15-hour genre-analysis university course for 12 composition majors without generative AI, and a 3-hour major exploration session using Suno for 12 non-major high school students. With scaffolding—a scene-analysis template, a prompt formula, and a minimal DAW operation set—all 12 high school students completed music combined with video in 3 hours. In the instructor's judgment, their outputs were comparable to those of the major students in some of the three dimensions: mood congruence, music placement, and completeness. The instruction-time difference (15 vs. 3 hours) arose under differing content and delivery and is not a generalized effect, but it suggests that generative AI with scaffolding may shift instruction from implementation training to scene analysis, musical judgment, and selection. The study presents a provisional five-stage sequence from these cases. Given limited synchronization precision and single-rater judgment, the findings support an introductory-level approach, not a replacement for professional training.