Ross ROSS = Recommend OSS · open-source software intelligence for agents

LargeWorldModel/LWM

Large World Model -- Modeling Text and Video with Millions Context observed · 2026-08-28

github.com/LargeWorldModel/LWM · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

25/100

  • Activity 0
  • Release rhythm 35
  • Longevity 66

Flags: no_releases

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-02. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 937
  • days_rel: n/a
  • days_push: 683
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

7425 stars · 563 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

Large World Model (LWM) is a family of open-source 7B-parameter multimodal autoregressive transformer models trained on long videos and books with RingAttention, supporting context up to 1M tokens. The repository provides training and inference code for language, image, and video understanding and generation.

Use cases

  • answer questions about hour-long YouTube videos
  • retrieve facts across a 1M-token context
  • chat with images and videos
  • generate videos and images from text
  • train transformers on million-length multimodal sequences
  • process entire books with a large-context language model

When to choose

  • you need a model that understands very long videos or documents
  • you want open weights for million-token-context multimodal research
  • you need a reference implementation of RingAttention and masked sequence packing

When to avoid

  • you need a production-ready, well-supported inference stack
  • you only need short-context text chat
  • you lack multi-GPU resources, since training targets large clusters

Facets

library · maturity active

machine-learning deep-learning llm-training llm-inference video-processing nlp large-language-models deep-learning machine-learning python multimodal ring-attention long-context vision-language-model video-understanding research video natural-language-processing linux gpu

2 sources

Member repositories

RepositoryRoleHealth v2
LargeWorldModel/LWMmain25

For agents

markdown · JSON · MCP: product_card(name="LargeWorldModel/LWM")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem