Ross ROSS = Recommend OSS · open-source software intelligence for agents

facebookresearch/sam-audio

The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how to use the model. observed · 2026-08-28

github.com/facebookresearch/sam-audio · Python · NOASSERTION (other) observed · 2026-08-28

Health v2 · maintenance only

55/100

  • Activity 84
  • Release rhythm 35
  • Longevity 25

Flags: no_releases no_license

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 363
  • days_rel: n/a
  • days_push: 99
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

3612 stars · 330 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

SAM-Audio is Meta's foundation model for isolating any sound in audio using text, visual, or temporal prompts. This repository provides inference code, model checkpoints via Hugging Face, and example notebooks for separating specific sounds from complex audio mixtures.

Use cases

  • separate a specific sound from an audio mixture using a text description
  • isolate a sound in audio based on visual cues from video
  • extract audio segments within a given time span
  • run inference with the SAM Audio foundation model
  • split an audio file into a target sound and residual audio
  • download and experiment with Meta's audio segmentation checkpoints

When to choose

  • you need prompt-driven audio source separation with natural language descriptions
  • you want state-of-the-art sound isolation from a research-grade foundation model
  • you have a CUDA GPU and want to integrate audio separation into a Python pipeline

When to avoid

  • you need a lightweight CPU-only audio separation tool
  • you want a ready-made end-user application rather than a Python library
  • you cannot obtain access to the gated Hugging Face checkpoints

Facets

library · maturity active

audio-processing machine-learning deep-learning llm-inference machine-learning artificial-intelligence speech-processing python windows audio-separation source-separation foundation-model segment-anything meta-ai text-prompting audio-visual inference audio gpu linux macos

1 source

Member repositories

RepositoryRoleHealth v2
facebookresearch/sam-audiomain55

For agents

markdown · JSON · MCP: product_card(name="facebookresearch/sam-audio")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem