The latest news and research from Amazon's science community. #AmazonScience

Global
Using the Neuron Kernel Interface, @reactorworld and Amazon's Neuron Science team built a kernel-centric path to real-time autoregressive diffusion video generation on Trainium. They tackled the dynamic shapes, memory access patterns, and cache management that make these workloads hard for generic compilers, and built techniques that generalize across models. amazon.science/blog/a-kernel…
4
3
18
4,157
"One workload is power-bound. The next workload is memory-bandwidth-bound. The next is memory-bound. It's one of the most interesting hardware design problems that we've seen in ages." Amazon SVP Peter DeSantis sat down with @dylan522p of @SemiAnalysis_ at #AIInfraSummit: aboutamazon.com/news/innovat…
1
4
2,290
Amazon Science retweeted
The most durable skill in the age of AI isn’t any single technology. It’s learning. I was back at @Stanford last week celebrating the Stanford-Amazon Research Initiative, bringing our researchers together to work on hard problems across AGI, robotics, healthcare and more. Excited to see what we learn and solve together.
4
3
24
2,154
Amazon Bio Discovery developed three AI approaches to accelerate antibody drug design: MochiBind (sequence-based affinity ranking), CA-MAP (developability prediction with batch effect correction), and an agent-guided design system with 46 lab-validated hits against a novel cancer target. amazon.science/blog/advancin…
1
3
4
2,685
Amazon Science retweeted
We spent tens of billions of tokens building a benchmark for AI security — so you don't have to. Deception Benchmark is now open on GitHub: 14,822 samples, 16 languages, 70+ CWEs. Download the dataset, run your tools, hold them accountable. go.aws/4xUbAjP #AWS #OpenSource
1
11
28
2,430
"Machine learning, at its core, is about generalization, not memorization," write @awscloud Applied Scientist Martin Bertran Lopez and @WarrenCntrPenn faculty affiliate @Aaroth for @AmazonScience. amazon.science/blog/why-dont…
1
3
319
Years of iterating against the same benchmarks should, by textbook logic, produce overfitting. It largely doesn't. New research explains why: strategies that generalize can be expressed in too compact a form to allow memorization, while the ones that overfit don't survive a compression bottleneck. amazon.science/blog/why-dont…
1
2
7
2,670
Amazon Science retweeted
Trainium has by far the best profiler of any accelerator. As you can see we have nanosecond-accurate traces of our programs, allowing you to write a program that generates images reliably in the profiler.
bad apple, but actually running on a trainium chip
10
23
371
32,961
.@Amazon and @DARPA brought together 150+ researchers in Seattle last week, with speakers including @awscloud CEO @mattsgarman, Fields Medalist @TaoistTerence, @EPrinceton Associate Professor @BorisHanin, and @NSAGov's Michael O'Hara to explore how AI is transforming mathematical discovery and reasoning.
2
4
31
17,564
Verus is an open-source, automated program verifier for Rust that mechanically checks code against a formal mathematical specification for all possible inputs. Amazon used it to prove correctness of Nitro Isolation Engine primitives. amazon.science/blog/developi…
8
70
540
74,367
When LLM judges agree, the right question is why. Shared prompts, model families, or training lineage can make a majority look stronger than it is. Dependence-aware aggregation via Ising models accounts for this, improving accuracy 9–14% over weighted majority vote. amazon.science/blog/when-llm…
5
14
2,926
Amazon Science retweeted
Honored to be named to @TIME's 2026 TIME100 AI list. This one belongs to our customers and the teams at @awscloud. Still early.
AWS CEO @MattSGarman has been named to @TIME’s 2026 TIME100 AI list. Congratulations, Matt. go.aws/4gFhle3
16
16
162
21,575
How did a model upgrade make agents worse? By pairing real enterprise SOPs with functioning tools and ground-truth grading across 12 industries and 2,000+ tasks, SOP-Bench helps find such anomalies. amazon.science/blog/sop-benc…
2
1,784
Amazon's Automated Reasoning Group started by demoing tools to prove AWS systems secure and correct. A decade later, they have proved the Nitro Isolation Engine, cryptographic code, and S3 correct. Now they are applying the same techniques to AI. amazon.science/blog/a-decade…
1
1
2,418
📣 AWS Trainium Frontier is open for registration. Train language models from scratch on purpose-built AI chips for @NeurIPSConf. Prizes include $25K for first place, co-publication with Annapurna Labs researchers, and a presentation in Sydney. Deadline is September 30. #NeurIPS2026 amazon.science/news/aws-trai…
6
8
3,228
Training a graph neural network on multiple objectives usually means blending conflicting gradients at every step. Instead of compromising among parameter updates from different training objectives, ControlG allocates capacity to objectives sequentially and dynamically via PID control. #ICML2026 amazon.science/blog/how-cont…
1
2
1,977
Most healthcare AI benchmarks test static medical knowledge or evaluate tool-using agents on provider-facing tasks. PatientAgentBench generates synthetic patient records and clinical vignettes, then runs multiturn dual-agent conversations scored by an LLM-as-a-jury panel across over 100 clinician-vetted criteria. amazon.science/blog/a-new-be…
3
5
1,625