Learn AI

by Amazon Research Publication

How Optimization and Forecasting Power Amazon Fulfillment

Deploying programmatic tool calling with pre-execution validation for production agentic systems
...
...

How Optimization and AI Power Amazon’s Fulfillment Network

On optical add/drop multiplexing architecture for hyperscale backbone networks
...
...

How Optimization and Forecasting Power Amazon Fulfillment

Beyond the harness: End-to-end optimization of context artifacts for enterprise Text-to-SQL
...
...

How to Evaluate LLM Agents Beyond Task Success

Beyond Task Success: An eight-metric tiered evaluation protocol for LLM agents over fragmented operational data
...
...

How Optimization and AI Power Amazon’s Delivery Network

MapScout: An agentic harness for map editing and geospatial data labeling
...
...

Off-Policy Ranking Evaluation: Rethinking Position Weights

Generalized position-based model: Rethinking position weights in ranking off-policy evaluation
...
...

The blind curator: How a biased judge silently disables skill retirement in self-evolving agents

The blind curator: How a biased judge silently disables skill retirement in self-evolving agents
...
...

How Optimization and Forecasting Power Amazon Fulfillment

Query-aware index pruning for retrieval under budget constraints
...
...

How Optimization and AI Power Amazon’s Fulfillment Network

Scalable conflation for maps data replay between heterogeneous geospatial data sources
...
...

How WARP Aligns RAG with Population Opinions

WARP: Wasserstein-Aligned RAG for population opinions
...
...

How to Build Realistic AI Twins of Mobile Users

AgenTwin: An end-to-end framework for building agentic twins of mobile users
...
...

How Sparse Attention Speeds Up Language Model Inference

SAS: Sparse attention synthesizer for efficient language model inference
...
...

How AI Turns BIM Requirements into Machine-Checkable IDS

Ishigaki-IDS-Bench: A benchmark for generating information delivery specification from BIM information requirements
...
...

How AI Learns to Check Evidence in Text and Images

Eliciting self-verification in multimodal reasoning agents with reinforcement learning
...
...

How Connectivity Shapes Attacks on Multi-Agent Recommenders

Attacking and defending multi-agent collaborative filtering systems through connectivity
...
...

How AI Finds Errors in Seismic Data Workflows

Analyzing seismic workflows with large language models: A case study on error detection
...
...

What Static Pruning Transfers Across Sparse Search Engines

Static pruning across sparse retrieval regimes: What transfers, what breaks, and what still helps
...
...

Predicting Trajectories from Historical Movement Patterns

Non-parametric spatiotemporal trajectory prediction via state-conditioned transition sampling
...
...

How Static Analysis Generates Least-Privilege IAM Policies

IAM policy autopilot: Static analysis for policy generation from application code
...
...

Why Pairwise Ranking Beats RL for Offline Explanation Choice

Pairwise ranking outperforms single-action RL for offline explanation selection: A practical lesson
...
...