# DirtyHarryLYL/Transformer-in-Vision

Recent Transformer-based CV and related works.

Repository: https://github.com/DirtyHarryLYL/Transformer-in-Vision
Canonical: https://ross.abutalabs.com/products/transformer-in-vision
License Family: other
Topics: transformer, vision-transformers, computer-vision, self-attention, multi-modal, visual-language, deep-learning, paper
Last push: 2023-08-22T15:15:34+00:00

## Health v2 (maintenance only)
Score: 32/100 (v2, computed 2026-09-02T17:46:02.011165+00:00)
- activity 0, release rhythm 35, longevity 100
- inputs: {"age_days": 2029, "days_push": 1107, "days_rel": null, "gap_med": null, "n_releases_24m": 0}
- flags: no_releases, no_license
- formula: round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10)

## Adoption (not part of the score)
Stars 1343, forks 140 (observed 2026-08-28T04:04:26.515000+00:00)

## What it is
A curated list of recent Transformer-based computer vision and related research works, with links to papers and code. It serves as a reading/reference resource covering vision transformers, self-attention, and multi-modal visual-language models.

## Use cases
- find papers on vision transformers
- survey transformer-based computer vision research
- learn about self-attention in vision models
- discover multimodal visual-language model papers
- keep up with transformer research in CV
- find code implementations of transformer CV papers

## When to choose
- you want a curated, categorized reading list of transformer CV papers
- you need quick links to papers and code for vision transformers and multimodal models

## When to avoid
- you need runnable software or a library
- you need a regularly updated list (updates are irregular)
- you need a license for redistribution

## Facets
- artifact type: learning-resource
- maturity: maintenance
- function: deep-learning, computer-vision, nlp
- domain: computer-vision, deep-learning, artificial-intelligence, tutorials
- platform: python
- tags: awesome-list, vision-transformer, paper-collection, self-attention, multimodal, visual-language

## Member repositories
- DirtyHarryLYL/Transformer-in-Vision (main) score 32

## Provenance
- Observed fields: from GitHub, fetched 2026-08-28T04:04:26.515000+00:00.
- Health v2: computed from the inputs above; adoption is never an input.
- Inferred fields (summary, facets, guidance): AI-extracted, prompt v1, taxonomy v1, on 2026-08-30T04:42:51.948740+00:00, confidence not recorded.
  - readme: https://github.com/DirtyHarryLYL/Transformer-in-Vision (fetched 2026-08-28T04:04:26.515000+00:00, sha 0a3e24f08c37)
- Data as of 2026-08-30T08:39:29.467469+00:00.
