NewYour coding agent can read the release notes before it upgrades.Set up the MCP server →
PyPI · #297 most downloaded on PyPI
Pandas on AWS.
Last release 2 months ago
03 Aug 2026
Ships fairly regularly
a new release about every 2 months
Nearly every release is documented
notes for 60 of the last 60 stable releases
1 version withdrawn
withdrawn after publishing
8 years old
167 releases · first in 2019
fix: map Athena varbinary correctly in athena2pyarrow by @hsusul in https://github.com/aws/aws-sdk-pandas/pull/3413
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.17.0...3.17.1
One column per quarter.
Full Changelog: 3.17.0...3.17.1
chore(lambda-layer): bump pyarrow to 24.0.0 by @PatrickBunker in https://github.com/aws/aws-sdk-pandas/pull/3375
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.16.1...3.17.0
Full Changelog: 3.16.1...3.17.0
fix(deps): upgrade lxml to 6.1.0 and redshift-connector to 2.1.13 (CVE-2026-41066) by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3309
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.16.0...3.16.1
Full Changelog: 3.16.0...3.16.1
feat: Support Pandas 3.x by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3272
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.15.1...3.16.0
Full Changelog: 3.15.1...3.16.0
fix: upgrade setuptools due to CVE-2026-23949 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3261
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.15.0...3.15.1
Full Changelog: 3.15.0...3.15.1
fix: upgrade aiohttp due to CVE-2025-69223 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3250
athena.to_iceberg documentation by @villoro in https://github.com/aws/aws-sdk-pandas/pull/3246Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.14.0...3.15.0
athena.to_iceberg documentation by @villoro in #3246Full Changelog: 3.14.0...3.15.0
chore: upgrade pg8000 due to CVE-2025-61385 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3225
CLEANPATH by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3211s3_output parameter to _start_query_execution call in "overwrite" mode by @sergeymazin in https://github.com/aws/aws-sdk-pandas/pull/3205Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.13.0...3.14.0
CLEANPATH by @kukushking in #3211s3_output parameter to _start_query_execution call in "overwrite" mode by @sergeymazin in #3205Full Changelog: 3.13.0...3.14.0
updated aiohhtp==3.12.15to fix CVE-2025-53643 (LOW) by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3197
aiohhtp==3.12.15to fix CVE-2025-53643 (LOW) by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3197Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.12.1...3.13.0
aiohhtp==3.12.15to fix CVE-2025-53643 (LOW) by @kukushking in #3197Full Changelog: 3.12.1...3.13.0
Moved to uv package manager 🔥 🔥 🔥
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.12.0...3.12.1
Full Changelog: 3.12.0...3.12.1
AWS Lambda Layers: pyarrow was upgraded to 20.0.0
large_list by @ashrielbrian in https://github.com/aws/aws-sdk-pandas/pull/3086large_list and large_string unit test for read_parquet_metadata by @ashrielbrian in https://github.com/aws/aws-sdk-pandas/pull/3089COPY with SERIALIZETOJSON test case by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3104Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.11.0...3.12.0
large_list by @ashrielbrian in #3086large_list and large_string unit test for read_parquet_metadata by @ashrielbrian in #3089COPY with SERIALIZETOJSON test case by @kukushking in #3104Full Changelog: 3.11.0...3.12.0
add support for Python 3.13 & deprecate Python 3.8 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/3045
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.10.1...3.11.0
fix: update references in introduction notebook by @emmanuel-ferdman in https://github.com/aws/aws-sdk-pandas/pull/3009
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.10.0...3.10.1
feat: Support numpy 2.0 by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2944
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.9.1...3.10.0
address Ray deprecation warnings by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2929
athena.read_sql_query failing for time columns by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2895s3.select_query by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2928Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.9.0...3.9.1
Replace deprecated ray parallelism arg with override_num_blocks by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2876
redshift.copy_from_files function by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2849athena.to_iceberg function by @aldder in https://github.com/aws/aws-sdk-pandas/pull/2861NULL values in athena.to_iceberg merge statement by @aldder in https://github.com/aws/aws-sdk-pandas/pull/2872tz attribute check, it was checking dtype instead of dt by @sanrodari in https://github.com/aws/aws-sdk-pandas/pull/2855Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.8.0...3.9.0
support client-side parameter resolution in athena.create_ctas_table by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2797
postgresql.to_sql by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2820s3_output parameter to athena.delete_from_iceberg_table by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2829Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.7.3...3.8.0
Iceberg schema evolution fails for map, array and struct types by @LeonLuttenberger in #2755
s3_output in athena.to_iceberg by @jaidisido in #2767to_iceberg by @jaidisido in #2768fixed_size_binary dtype support by @jaidisido in #2775_id by @kukushking in #2784list_to_arrow_table by @kukushking in #2778athena.to_iceberg overwrite to delete table in order to preserve Iceberg transactions history by @erwan-simon in #2776idna from 3.6 to 3.7 by @dependabot in #2772aiohttp from 3.9.3 to 3.9.4 by @dependabot in #2777Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.7.2...3.7.3
replace deprecated np.split_array by @jaidisido in #2735
wr.athena.to_iceberg - Insert query has mismatched column types #2678 by @GalvFionic in #2715s3_output in athena.to_iceberg by @jaidisido in #2727np.split_array by @jaidisido in #2735to_iceberg fails with non-lowercase column names by @LeonLuttenberger in #2736Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.7.1...3.7.2
fix breaking change in _create_table by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/2711
_create_table by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/2711redshift.to_sql doc indentation error by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2706Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.7.0...3.7.1
Lake Formation Governed tables are being phased out and we are dropping support (#2692).
Lake Formation Governed tables are being phased out and we are dropping support (#2692).
site-packages folder by @AlJohri in https://github.com/aws/aws-sdk-pandas/pull/2698Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.6.0...3.7.0
Enable Iceberg row deletion & add mode parameter to to_iceberg by @LeonLuttenberger in #2632
mode parameter to to_iceberg by @LeonLuttenberger in #2632large_string by @joakibo in #2663max_results to athena.list_query_executions by @LeonLuttenberger in #2665Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.5.2...3.6.0
Add vulnerability label to dependabot PRs with alert state by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/2629
to_iceberg support for filling missing columns in the DataFrame with None by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2616ignore_nulls for container types by @raaidarshad in #2636s3_additional_kwargs to docstrings by @malachi-constant in https://github.com/aws/aws-sdk-pandas/pull/2627Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.5.1...3.5.2
Deserialization error when reading from DynamoDB using KeyConditionExpression by @LeonLuttenberger in #2607
KeyConditionExpression by @LeonLuttenberger in #2607show_create_table to Athena API page by @MikeSchriefer in #2610bump2version with bump-my-version by @LeonLuttenberger in #2608Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.5.0...3.5.1
Due to CVEs, Ray is capped to patched version 2.9.x. As a result, the latest version of the library cannot be used on the Glue for Ray runtime. We hav
Due to CVEs, Ray is capped to patched version 2.9.x. As a result, the latest version of the library cannot be used on the Glue for Ray runtime. We have raised the CVEs issue to the Glue team
spark_properties to athena spark by @rajagurunath in https://github.com/aws/aws-sdk-pandas/pull/2508MERGE INTO support for Iceberg by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2527analysis_template_arn to cleanrooms.read_sql_query by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/2584table and schema params for Redshift by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2551oracledb to 2.0 by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2574Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.4.2...3.5.0
Update pyarrow to 14.0.1 to fix arbitrary code execution security vulnerability
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.4.1...3.4.2
feat: Add schema evolution to athena.to_iceberg by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2465
athena.to_iceberg by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2465client_request_token by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2474cn-north-1 & cn-northwest-1 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2514sanitize_column_name in create_*_table by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2464requests_aws4auth not being treated as an optional dependency by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2471Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.4.0...3.4.1
Geospatial - parse Athena geospatial types via geopandas by @kukushking in #2346
wr.cloudwatch queries by @LeonLuttenberger in #2430athena.to_iceberg wait_query by @jaidisido in #2428athena.to_iceberg by @jaidisido in #2446wr.s3.to_parquet by @kukushking in #2455Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.3.0...3.4.0
Support Athena query prepared statements & Athena parameterized queries by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2344
cleanrooms.wait_query by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2381Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.2.1...3.3.0
Fix error where library could not be imported on Windows due to No module named 'pyarrow._orc' by @LeonLuttenberger in #2341 #2337
No module named 'pyarrow._orc' by @LeonLuttenberger in #2341 #2337packaging version requirement by @LeonLuttenberger in #2340Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.2.0...3.2.1
Adapt benchmark tests to Glue for Ray GA breaking changes by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/2316
s3.read_orc and s3.to_orc by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2312 🔥wr.athena.create_spark_session & wr.athena.run_spark_calculation by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2314 🚀to_sql for RDS Data API by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2287UNLOAD by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2284allowed_to_use and allowed_to_manage when creating QuickSight resources by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2278PARTITIONED BY and additional table properties support by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2322s3.read_parquet by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2328test_spectrum_decimal_cast by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2283dtype_backend use in read_parquet_table by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2307register_func to handle type checking by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2309Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.1.1...3.2.0
fix: Add missing packaging dependency by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2281
packaging dependency by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2281Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.1.0...3.1.1
Add neptune.bulk_load for bulk loading data into Neptune by @LeonLuttenberger in #2238 #2267
neptune.bulk_load for bulk loading data into Neptune by @LeonLuttenberger in #2238 #2267s3.to_deltalake function by @LeonLuttenberger in #2228chunked parameter to DynamoDB read functions by @LeonLuttenberger in #2227ignore_metadata to False by default by @jaidisido in #2206path_ignore_suffix by @LeonLuttenberger in #2240test_spectrum_decimal_cast test by @LeonLuttenberger in #2244emr.create_cluster was not passing security configuration to internal method by @malachi-constant in #2246timestream.list_tables by @SukruHan #2275layers.rst with Python 3.10 layers by @LeonLuttenberger in #2219pyi files by @LeonLuttenberger in #2229 #2255 #2256Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0...3.1.0
Deprecate wr.s3.merge_upsert_table by @kukushking in #2076 ⚠️
pip install awswrangler[<MODULE_NAME>], for example pip install awswrangler[redshift]dt.datetime is parsed into DATETIME xxxx-xx-xx xx:xx:xx, while a parameter of type str is formatted into "x"TypeDict by @LeonLuttenberger and @kukushking in #1855 #1996 #2016 #2055 #2081 💼
to_parquet, to_csv and to_jsonwr.s3.merge_upsert_table by @kukushking in #2076 ⚠️updated_name parameter in update_ruleset by @jaidisido in #2122 ⚠️AWS SDK for pandas can now run at scale 🚀💻🚀
use_theads parameter to dynamodb.read_items by @LeonLuttenberger in #2113 📈wr.dynamodb.put_df with executor task by @LeonLuttenberger in #2118 📈DatabaseInput by @malachi-constant in #2067 🔧timestream.create_table by @cnfait in #1819_read_parquet_metadata_file function based on the PyArrow file system by @LeonLuttenberger in #2050@Experimental and @Deprecated annotations by @kukushking in #2062describe_objects by @jaidisido in #2069bulk_read option for reading large amounts of Parquet files quickly by @LeonLuttenberger in #2033s3.to_json and s3.to_csv by @LeonLuttenberger in #1631s3.read_csv, s3.read_json and s3.read_fwf by @LeonLuttenberger in #1567 #1607s3.wait_objects by @LeonLuttenberger in #1539s3.to_parquet by @kukushking in #1526s3.delete objects by @malachi-constant in #1474s3.read_parquet by @jaidisido in #1513s3.select_query by @kukushking in #1446Literal typing for mode and projection_types by @LeonLuttenberger in #2191read_parquet_metadata_distributed by @jaidisido in #2196utcnow argument in start_query by @LeonLuttenberger in #2193awswrangler.distributed from coverage report by @LeonLuttenberger in #1884Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/2.20.1...3.0.0
breaking change: Move dependencies to optional by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1992
names parameter support to PyArrow reading by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2008_read_parquet_metadata_file function based on the PyArrow file system by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2050bulk_read option for reading large amounts of Parquet files quickly by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2033parallelism and bulk_read into ray_modin_args by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2081awswrangler.distributed from coverage report by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/1884 test_modin_s3_read_parquet_many_files by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2096Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0rc2...3.0.0rc3
(enhancement): Enable missing unit tests and Redshift, Athena, LF load tests by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1736
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0rc1...3.0.0rc2
(enhancement): Move RayLogger out of non-distributed modules by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1686
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0b3...3.0.0rc1
(feat): Add partitioning on block level by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/1653
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0b2...3.0.0b3
(feat) Update to Ray 2.0 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/1635
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0b1...3.0.0b2
(test) Consolidate unit and load tests by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1525
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/3.0.0a2...3.0.0b1
This is a pre-release for the Wrangler@Scale project
This is a pre-release for the Wrangler@Scale project
Full Changelog: https://github.com/awslabs/aws-data-wrangler/compare/3.0.0a1...3.0.0a2
This is a pre-release for the Wrangler@Scale project
This is a pre-release for the Wrangler@Scale project
Full Changelog: https://github.com/awslabs/aws-data-wrangler/compare/2.16.1...3.0.0a1
deprecate: boto3 resources by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/2097
chunksize=True by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2087to_csv and to_json by @LeonLuttenberger in https://github.com/aws/aws-sdk-pandas/pull/2104Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/2.20.0...2.20.1
(breaking change): Use ExecuteStatement instead of Scan for DynamoDB read_partiql by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1964
dynamodb.read_partiql no longer performs a Scan operation under the hood. Instead the ExecuteStatement API is used. It means that the PartiQL* IAM permission is required instead of ScanExecuteStatement instead of Scan for DynamoDB read_partiql by @jaidisido in https://github.com/aws/aws-sdk-pandas/pull/1964xfail's in tests by @malachi-constant in https://github.com/aws/aws-sdk-pandas/pull/1930Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/2.19.0...2.20
Glue Data Quality now supported, checkout the tutorial 🔥
read_items method by @a-slice-of-pydelete_all methods by @malachi-constant in https://github.com/aws/aws-sdk-pandas/pull/1913measure_name in wr.timestream.write() by @malachi-constant in https://github.com/aws/aws-sdk-pandas/pull/1925We thank the following contributors/users for their work on this release: @jaidisido, @kukushking, @LeonLuttenberger, @cnfait, @malachi-constant, @mdavis-xyz, @dydc, @enricomarchesin
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/2.18.0...2.19.0
Pyarrow 10 support 🔥 by @kukushking in https://github.com/aws/aws-sdk-pandas/pull/1731
af-south-1 (Cape Town) 🌍 by @malachi-constants3.read.parquet() validate schema by @malachi-constant in https://github.com/aws/aws-sdk-pandas/pull/1735We thank the following contributors/users for their work on this release: @lucasasmith, @vikramsg, @mycaule, @pal0064, @LeonLuttenberger, @cnfait, @malachi-constant, @kukushking, @jaidisido
Full Changelog: https://github.com/aws/aws-sdk-pandas/compare/2.17.0...2.18.0
RedshiftDataAPI serverless support 🔥 #1530
get_query_results to the Athena module #1496
generate_create_query to the Athena module #1514
INSERT IGNORE for mysql.to_sql #1429use_column_names to redshift.copy akin to redshift.to_sql #1437redshift.connect #1467timestream_endpoint_url property to the config #1483validate_schema=True for wr.s3.read_parquet breaks with partition columns and dataset=True #1426wr.neptune.to_property_graph failing for Neptune version 1.1.1.0 #1407catalog_id in wr.catalog.create_database #1480TagColumnOperation in quicksight.create_athena_dataset #1570s3.to_json compression parameters is passed twice when dataset=True #1585Since the last release, the library has been accepted as an official SDK for AWS, and rebranded as AWS SDK for pandas 🚀. The module names in Python will remain the same. One noteworthy change, however, is that the AWS Lambda Manager layer name has been renamed from AWSDataWrangler to AWSSDKPandas.
You can view the ARN value for the layers here.
⚠️ For platforms without PyArrow 7 support (e.g. MWAA, EMR, Glue PySpark Job):
pip install pyarrow==2 awswrangler
We thank the following contributors/users for their work on this release:
@bechbd, @maxispeicher, @timgates42, @aeeladawy, @KhueNgocDang, @szemek, @malachi-constant, @cnfait, @jaidisido, @LeonLuttenberger, @kukushking
> 🐛 Fixed issue introduced by 2.16.0 to method s3.read_parquet()
🐛 Fixed issue introduced by
2.16.0to methods3.read_parquet()
s3.read_parquet() #1412P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
Full Changelog: https://github.com/awslabs/aws-data-wrangler/compare/2.16.0...2.16.1
> ⚠️ For platforms without PyArrow 7 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 7 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
We thank the following contributors/users for their work on this release:
@bnimam, @IldarAlmakaev, @syokoysn, @IldarAlmakaev, @thomasniebler, @maxdavidson91, @takeknock, @Sleekbobby1011, @snikolakis, @willsmith28, @malachi-constant, @cnfait, @jaidisido, @kukushking
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ Dropped Python 3.6 support
⚠️ Dropped Python 3.6 support
⚠️ For platforms without PyArrow 7 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
sparql extra & make SPARQLWrapper dependency optional #1252P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ Dropped Python 3.6 support
⚠️ Dropped Python 3.6 support
⚠️ For platforms without PyArrow 7 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
We thank the following contributors/users for their work on this release:
@bechbd, @sakti-mishra, @mateogianolio, @jasadams, @malachi-constant, @cnfait, @jaidisido, @kukushking
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 6 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 6 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
ExcludeColumnSchema=True argument to the glue.get_partitions call to reduce response size #1094write_parquet via pyarrow_additional_kwargs #1057rename_duplicate_columns and handle_duplicate_columns flag to sanitize_dataframe_columns_names method #1124timestamp_as_object argument to all database read_sql_table methods #1130ignore_null to read_parquet_metadata method #1125to_parquet method #1058We thank the following contributors/users for their work on this release:
@dennyau, @kailukowiak, @lucasmo, @moykeen, @RigoIce, @vlieven, @kepler, @mdavis-xyz, @ConstantinoSchillebeeckx, @kukushking, @jaidisido
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 6 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 6 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
We thank the following contributors/users for their work on this release:
@csabz09, @Falydoor, @moritzkoerber, @maxispeicher, @kukushking, @jaidisido
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 5 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 5 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 5 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 5 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
We thank the following contributors/users for their work on this release:
@AssafMentzer, @mureddy19, @isichei, @DonnaArt, @kukushking, @jaidisido
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 5 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 5 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
connect methods #871We thank the following contributors/users for their work on this release:
@pwithams, @maxispeicher, @kukushking, @jaidisido
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 4 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 4 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
wr.s3.to_csv #787wr.s3.to_csv when dataset = True #765CSV as unload format to wr.redshift.unload_files #761We thank the following contributors/users for their work on this release:
@maxispeicher, @kukushking, @jaidisido, @mohdaliiqbal
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 4 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 4 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
S3 Select 🚀 #678VersionId parameter for S3 read operations #721wr.redshift.unload_to_files #729wr.redshift.to_sql #705We thank the following contributors/users for their work on this release:
@maxispeicher, @kukushking, @jaidisido
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 4 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 4 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
to_parquet #669LOCK before concurrent COPY calls in Redshift #665iter_batches (>= 3.0.0 only) #660drop, truncate, cascade) #671dtypes for empty ctas athena queries #659We thank the following contributors/users for their work on this release:
@maxispeicher, @kukushking, @igorborgest, @gballardin, @eferm, @jaklan, @Falydoor, @chariottrider, @chriscugliotta, @konradsemsch, @gvermillion, @russellbrooks, @mshober.
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run or use them from our S3 public bucket!
> ⚠️ For platforms without PyArrow 3 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 3 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
wr.athena.read_sql_query params argument use #609wr.s3.write.to_parquet #617exist_ok flag to safely create a Glue database #642chunked mode in wr.s3.read_parquet_table #627\ character from wr.s3.read_parquet_table method #638postgres as an engine value #630merge_upsert_table fails or data_quality is insufficient #601athena2pyarrow method #612We thank the following contributors/users for their work on this release:
@maxispeicher, @igorborgest, @mattboyd-aws, @vlieven, @bentkibler, @adarsh-chauhan, @impredicative, @nmduarteus, @JoshCrosby, @TakumiHaruta, @zdk123, @tuannguyen0901, @jiteshsoni, @luminita.
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run!
> ⚠️ For platforms without PyArrow 3 support (e.g. MWAA, [EMR](https://aws-data-wrangler.readthedocs.io/en/stable/install.html#emr-cluster), [Glue PyS
⚠️ For platforms without PyArrow 3 support (e.g. MWAA, EMR, Glue PySpark Job):<br> ➡️
pip install pyarrow==2 awswrangler
chunksize parameter to the to_sql function. Default set to 200. Decreased insertion time from 120 to 1 second #599path argument is now optional in s3.to_parquet and s3.to_csv functions #586map_types boolean (set to True by default) to convert pyarrow DataTypes to pandas ExtensionDtypes #580ctas_database_name argument to store ctas_temporary_table in an alternative database #576We thank the following contributors/users for their work on this release:
@maxispeicher, @igorborgest, @ilyanoskov, @VashMKS, @jmahlik, @dimapod, @Reeska
P.S. The AWS Lambda Layer file (.zip) and the AWS Glue file (.whl) are available below. Just upload it and run!
Your coding agent can read these notes before it upgrades. Set up the MCP server →