Ross ROSS = Recommend OSS · open-source software intelligence for agents

microsoft/Oscar

Oscar and VinVL observed · 2026-08-28

github.com/microsoft/Oscar · Python · MIT (permissive) · archived observed · 2026-08-28

Health v2 · maintenance only

10/100

  • Activity 0
  • Release rhythm 35
  • Longevity 100

Flags: no_releases archived

How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: n/a
  • age_days: 2302
  • days_rel: n/a
  • days_push: 1102
  • n_releases_24m: 0

Full methodology

Adoption not part of the score

1053 stars · 249 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-30, confidence not recorded

Oscar is Microsoft's research code for object-semantics aligned cross-modal pre-training of vision-language models, with VinVL providing improved visual representations. It includes pretrained checkpoints and fine-tuning code for vision-language understanding and generation tasks.

Use cases

  • pretrain a vision-language model on image-text pairs
  • fine-tune a model for visual question answering
  • generate image captions with a pretrained model
  • search images using text queries
  • reproduce Oscar or VinVL paper results
  • extract object-attribute features for V+L tasks

When to choose

  • you need proven state-of-the-art vision-language models like Oscar/VinVL
  • you are doing research on VQA, image captioning, or image-text retrieval
  • you want pretrained checkpoints for cross-modal pre-training experiments

When to avoid

  • you need a maintained production library - the repo is in maintenance mode with no active development
  • you want modern multimodal LLMs - the authors point to LLaVA for instruction-tuned models
  • you need a simple inference API rather than research training code

Facets

library · maturity maintenance

machine-learning deep-learning nlp image-processing artificial-intelligence computer-vision deep-learning python vision-and-language pre-training image-captioning vqa image-text-retrieval multimodal research-code vinvl natural-language-processing

1 source

Member repositories

RepositoryRoleHealth v2
microsoft/Oscarmain10

For agents

markdown · JSON · MCP: product_card(name="microsoft/Oscar")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem