NOW BUZZING 헤어컬러가 브라운과 오렌지로 쪼개지는 이… · 코치, 85주년에 '낡음'을 미래로 바꾸… · 프레피룩, 왜 더 플라자에서 다시 태어났…◆
NOW BUZZING 헤어컬러가 브라운과 오렌지로 쪼개지는 이… · 코치, 85주년에 '낡음'을 미래로 바꾸… · 프레피룩, 왜 더 플라자에서 다시 태어났…◆
JELIBI
⌕ SUBSCRIBE
SEOUL · DAILY · ISSUE No.327 WE TRIED IT · THE PICK · BEHIND IT 2026.09.30 · WED
JELIBI
⌕ SUB
Beauty Body Gadget Bites Buzz World
HOME/GADGET/BEHIND IT
GADGET BEHIND IT · 5 MIN READ

Meta, 음성 분리 AI 모델 SAM Audio 공개

JD 젤리비 편집국 · 2025.12.16 SHARE · COPY LINK

무슨 발표인가

  • 텍스트·시각·시간 범위 프롬프트로 복합 오디오에서 음성 격리
  • 음악·팟캐스트·영상 편집·접근성·과학 연구 등 다양 분야 적용
  • Segment Anything 컬렉션의 최신 모델

원문 (영어)

Today, we’re introducing SAM Audio, a state-of-the-art AI model that enables you to segment sound. Imagine recording a video of your favorite band and isolating the guitar or vocals with a single click, using text prompts to filter traffic noise from a video filmed outside, or removing the sound of a dog barking from your entire podcast recording.

SAM Audio, the latest addition to our Segment Anything collection , transforms audio processing by making it easy to isolate any sound from complex audio mixtures using text, visual, and time span prompts. This intuitive approach mirrors how people naturally engage with sound, making professional-grade audio separation more accessible and easier than ever before.

SAM Audio has the potential to transform audio and video editing and drive innovation in areas like music, podcasting, television, film, scientific research, accessibility, and more. Until now, audio segmentation and editing has been a fragmented space, with a variety of tools designed for single-purpose use cases.

As a unified model, SAM Audio is the first to support use cases that match how people naturally think about audio, and achieves cutting-edge performance across diverse, real-world scenarios. SAM Audio supports three kinds of prompts: Text prompting : Type “dog barking” or “singing voice” to extract specific sounds.

https://about.fb.com/wp-content/uploads/2025/12/01_Text-Prompts.mp4 Visual prompting : Click on the person or object in the video that’s making a sound to isolate their audio. https://about.fb.com/wp-content/uploads/2025/12/02_Visual-Prompts.mp4 Span prompting : An industry first, this method lets you mark time segments where target audio occurs.

https://about.fb.com/wp-content/uploads/2025/12/03_Span-Prompts.

원문: Meta Newsroom — "Our New SAM Audio Model Transforms Audio Editing" (2025-12-16) 공식 원문: https://about.fb.com/news/2025/12/our-new-sam-audio-model-transforms-audio-editing/

#Meta Newsroom
THE JELIBI BRIEF

A small team that reads too much internet so you don't have to.

SUBSCRIBE →
GADGET

인스로픽, 안전장치 없이 사이버공격 자동화 모델 공개

GADGET

삼성 6개 계열사, AI 인프라 기업 헬릭스에 10억 달러 투자

GADGET

TBC, AWS와 협력해 신경세포 유래 AI 비디오 모델 출시

← GADGET BEHIND IT

Meta, 음성 분리 AI 모델 SAM Audio 공개

젤리비 편집국·2025.12.16·5 MIN
IN THIS PIECE
무슨 발표인가 원문 (영어)

무슨 발표인가

  • 텍스트·시각·시간 범위 프롬프트로 복합 오디오에서 음성 격리
  • 음악·팟캐스트·영상 편집·접근성·과학 연구 등 다양 분야 적용
  • Segment Anything 컬렉션의 최신 모델

원문 (영어)

Today, we’re introducing SAM Audio, a state-of-the-art AI model that enables you to segment sound. Imagine recording a video of your favorite band and isolating the guitar or vocals with a single click, using text prompts to filter traffic noise from a video filmed outside, or removing the sound of a dog barking from your entire podcast recording.

SAM Audio, the latest addition to our Segment Anything collection , transforms audio processing by making it easy to isolate any sound from complex audio mixtures using text, visual, and time span prompts. This intuitive approach mirrors how people naturally engage with sound, making professional-grade audio separation more accessible and easier than ever before.

SAM Audio has the potential to transform audio and video editing and drive innovation in areas like music, podcasting, television, film, scientific research, accessibility, and more. Until now, audio segmentation and editing has been a fragmented space, with a variety of tools designed for single-purpose use cases.

As a unified model, SAM Audio is the first to support use cases that match how people naturally think about audio, and achieves cutting-edge performance across diverse, real-world scenarios. SAM Audio supports three kinds of prompts: Text prompting : Type “dog barking” or “singing voice” to extract specific sounds.

https://about.fb.com/wp-content/uploads/2025/12/01_Text-Prompts.mp4 Visual prompting : Click on the person or object in the video that’s making a sound to isolate their audio. https://about.fb.com/wp-content/uploads/2025/12/02_Visual-Prompts.mp4 Span prompting : An industry first, this method lets you mark time segments where target audio occurs.

https://about.fb.com/wp-content/uploads/2025/12/03_Span-Prompts.

원문: Meta Newsroom — "Our New SAM Audio Model Transforms Audio Editing" (2025-12-16) 공식 원문: https://about.fb.com/news/2025/12/our-new-sam-audio-model-transforms-audio-editing/

#Meta Newsroom
THE JELIBI BRIEF

A small team that reads too much internet so you don't have to.

SUBSCRIBE →
MORE IN BEHIND IT
질병관리청, WHO 의료대응수단 네트워크 포럼 참석해 감염병 대응 협력 강화인스로픽, 안전장치 없이 사이버공격 자동화 모델 공개삼성 6개 계열사, AI 인프라 기업 헬릭스에 10억 달러 투자Samsung, 세계 심장의 날 캠페인으로 심장 건강 관리 강조CJ도너스캠프, 아동복지시설 교사 260명 교사교육 실시
THE JELIBI BRIEF

인터넷을 너무 많이 보는 팀이 대신 골라 옵니다.

SUBSCRIBE

바로가기 등록하시면, 더 쉽게 찾아보실 수 있습니다.