# Moving between speech and text across languages

SeamlessM4T joins speech recognition, text translation, speech translation, and speech generation in one multilingual research model family.

Canonical: https://brightaifuture.com/discoveries/seamlessm4t-speech-translation
Format: discovery
Source publication: 2023-08-22
Bright publication: 2026-09-19
Substantive update: None recorded
Evidence and review: Emerging; confidence: unassessed; source-checked; ai-assisted. AI-assisted comparison with the cited sources. Source-checked means the record was checked against those sources; it does not claim independent reproduction, expert review, or validation of the publisher’s results.

## The human problem

Language differences can block conversation and access to spoken information, especially where specialist translation tools are scarce.

## The prior constraint

Speech translation pipelines commonly stitched together separate recognition, translation, and synthesis systems.

## AI’s actual role

The model accepts speech or text and generates translated speech or text across its documented language set.

## The documented result

Meta publishes inference tooling, checkpoints, and supported-language documentation that allow researchers to run and evaluate the model family.

## Why it may matter

One inspectable model family can simplify research across the pipeline while making language-by-language failure analysis more urgent.

## Limitations

SeamlessM4T v1 and v2 weights are CC-BY-NC-4.0, so commercial use cannot be assumed. Translation quality, toxicity, and speech identity require evaluation for each language and setting.

Capabilities and language coverage are developer-described. MIT code and noncommercial model weights have materially different reuse permissions.

## Unresolved questions



## Provenance and history

{
  "dates": {
    "eventDate": null,
    "publicationDate": "2023-08-22",
    "captureDate": "2026-09-19",
    "lastReviewedDate": "2026-09-19"
  },
  "provenance": {
    "origin": "editorial",
    "externalId": "https://github.com/facebookresearch/seamless_communication"
  },
  "revisions": [
    {
      "id": "revision:open-models-added:seamlessm4t-speech-translation",
      "recordedAt": "2026-09-19",
      "summary": "Bright added this source-checked open-model application record. The cited source publication date is 2023-08-22; 2026-09-19 is when Bright added this record.",
      "sourceIds": [
        "seamless-communication-repository",
        "seamlessm4t-paper"
      ]
    }
  ],
  "corrections": []
}

## Original sources

- [Seamless Communication](https://github.com/facebookresearch/seamless_communication)
- [SeamlessM4T: Massively Multilingual & Multimodal Machine Translation](https://arxiv.org/abs/2308.11596)

## Continue exploring

- [Open Models](https://brightaifuture.com/open-models)
- [Open intelligence](https://brightaifuture.com/worlds/open)
- [What changes when powerful models become open-weight?](https://brightaifuture.com/threads/open)
- [Someone builds on it](https://brightaifuture.com/open-intelligence)
