AWS Storage Blog

Category: Learning Levels

How Precisely transforms user experience with AI agents using Amazon S3 Vectors

At Precisely, the team is reimagining the user experience for its Data Integrity Suite by adding a conversational interface powered by AI agents to the traditional UI. With this enhancement, users can interact with the platform more naturally and intuitively (asking questions, making requests, and exploring data assets through dialogue) while still benefiting from the […]

Connect workloads to Amazon S3 Files across VPCs and accounts

Organizations store vast amounts of data in Amazon S3 for machine learning, data analytics, media processing, and generative AI workloads. Many of the applications, agents, and teams that work with that data are file-based: they read and write on a mounted path using the file and directory operations and POSIX tools they already depend on. […]

AWS Transfer Family Featured Image

Build AI-powered file classification with AWS Transfer Family

Organizations that receive files from external partners through SFTP face a persistent operational challenge: routing each file to the correct downstream system. Invoices, contracts, images, CSVs, and reports all arrive in a single landing zone, and each requires a different destination. The traditional approach—pattern-matching on file names with regular expressions—is inherently fragile. It relies on […]

Enable zero-copy access to AWS services on Amazon FSx for NetApp ONTAP with Amazon S3 Access Points

Semiconductor verification teams run thousands of simulation jobs every night using Electronic Design Automation (EDA) tools. A large verification environment can generate logs from 100,000 or more test executions per night. A single regression cycle produces simulation logs, compilation logs, and scheduler logs. For a regression with dozens of failures, manual triage typically takes 45–60 […]

Secure team file sharing on Amazon WorkSpaces with Amazon EFS

As more organizations move to remote and distributed work, builders and development teams face growing challenges in sharing files securely. Without a central file system, these teams deal with version control problems, unclear access permissions, and security risks from scattered storage locations. Central IT becomes a bottleneck when handling every permission change, yet granting full […]

Amazon S3 Object Lock

Flexibly control Amazon S3 Object Lock retention based on real business events

Immutability is a foundational data protection control, protecting data in place against unintended changes and deletions by authorized users, and changes by unauthorized users. In many cases, however, it isn’t clear at the time data is written how long it needs to stay immutable, or when that period should begin. A signed contract might need […]

Amazon S3 Tables

How Tubular Labs reclaimed 50% of engineering capacity by rebuilding their 70TB pipeline on Apache Iceberg and Amazon S3 Tables

Customer Story | Amazon S3 Tables – Learn how Tubular Labs (part of Chartbeat Inc.) reclaimed 50% of engineering capacity by replacing fragile, file-based data pipelines with a Common Pipeline Runtime (CPR) built on Apache Iceberg and Amazon S3 Tables. By Tubular Labs Engineering (part of Chartbeat Inc.), in collaboration with the AWS Solution Architecture team […]

Migrate VMware Storage to Amazon FSx for NetApp ONTAP using AWS Transform

Enterprise storage underpins every critical workload in a VMware environment. It’s not just capacity, it’s the operational backbone that delivers automatic failover, instant snapshots, writable clones, inline deduplication, and multi-protocol access that production applications depend on every day. When the time comes to migrate these workloads to the cloud, teams expect those same capabilities on […]

Amazon S3 Express One Zone thumbnail

Run Spark 31% faster and optimize compute costs with Amazon S3 Express One Zone on Amazon EMR

As your Spark datasets grow, storage latency often becomes the constraint, impeding application performance. Query runtimes stretch, and the bottleneck shifts from compute to how fast each node can read from Amazon Simple Storage Service (Amazon S3). We benchmarked this directly on Amazon EMR with TPC-DS at 3 TB scale. On an 8-node Graviton4 cluster […]

Detect stalled Amazon S3 live replication to prevent unexpected storage costs

Organizations that replicate data across storage locations for disaster recovery, compliance, or analytics need a way to detect when replication stalls, because without early warning, unreplicated objects accumulate silently and drive unexpected storage costs. Most teams rely on automated retention policies to clean up the source copy after a defined period, but when replication breaks […]