Ross ROSS = Recommend OSS · open-source software intelligence for agents

aws/aws-sdk-pandas

pandas on AWS - Easy integration with Athena, Glue, Redshift, Timestream, Neptune, OpenSearch, QuickSight, Chime, CloudWatchLogs, DynamoDB, EMR, SecretManager, PostgreSQL, MySQL, SQLServer and S3 (Parquet, CSV, JSON and EXCEL). observed · 2026-08-28

github.com/aws/aws-sdk-pandas · homepage · Python · Apache-2.0 (permissive) observed · 2026-08-28

Health v2 · maintenance only

94/100

  • Activity 99
  • Release rhythm 84
  • Longevity 100
How is this computed?

round(0.45*activity + 0.35*rhythm + 0.20*longevity); archived -> min(score, 10) — computed 2026-09-03. Adoption (stars, forks) is never an input.

  • gap_med: 43.0
  • age_days: 2746
  • days_rel: 30
  • days_push: 7
  • n_releases_24m: 13

Full methodology

Adoption not part of the score

4118 stars · 742 forks observed · 2026-08-28

What it is AI-extracted, prompt v1, taxonomy v1, 2026-08-29, confidence not recorded

AWS SDK for pandas (awswrangler) is a Python library that extends pandas with high-level APIs for reading and writing data across AWS services like S3, Athena, Glue, Redshift, Timestream, DynamoDB, and OpenSearch. It simplifies building data lakes and ETL pipelines by bridging pandas DataFrames with AWS analytics and storage services.

Use cases

  • read and write parquet datasets on S3 from pandas
  • query Amazon Athena and get results as a pandas DataFrame
  • load pandas DataFrames into Redshift
  • build a data lake on S3 with Glue catalog integration
  • run ETL jobs in AWS Lambda with pandas
  • export data to DynamoDB or Timestream from pandas
  • read CSV, JSON, and Excel files from S3 into pandas

When to choose

  • you work with pandas and need seamless AWS data lake integration
  • you want to move data between S3, Athena, Redshift, and other AWS analytics services
  • you build ETL pipelines in Python on AWS
  • you need Glue Catalog-backed datasets with partitioning

When to avoid

  • you work outside AWS or with non-AWS data stores
  • you need a general-purpose ORM or database driver rather than analytics integration
  • your project does not use pandas or Arrow-based workflows

Facets

library · maturity active

etl database data-science serialization data-science cloud-computing databases big-data python cloud serverless pandas aws athena redshift glue s3 parquet data-lake awswrangler apache-arrow data-engineering

1 source

Member repositories

RepositoryRoleHealth v2
aws/aws-sdk-pandasmain94

For agents

markdown · JSON · MCP: product_card(name="aws/aws-sdk-pandas")

Data as of 2026-08-30T08:39:29.467469+00:00 · Report a problem