Blog
Notes, deep dives, and experimental results from training agents with reinforcement learning on Bedrock AgentCore Runtime.
- Training multi-step agents on AgentCore Runtime with reinforcement learning — an end-to-end system for online RL on multi-step agents, with results from training on two long-horizon benchmarks: MigrationBench (Java code-migration) and OfficeBench (office automation).