Reactive Machines

Group-Specific Correlative Policy Development for Heterogeneous Preference Alignment

Group-Specific Correlative Policy Development for Heterogeneous Preference Alignment

Despite their complex general purpose capabilities, Large Language Models (LLMs) often fail to align with diverse preferences because standard post-training…
Automating competitive price intelligence with Amazon Nova Act

Automating competitive price intelligence with Amazon Nova Act

Monitoring competitor prices is essential for ecommerce teams to maintain a market edge. However, many teams remain trapped in manual…
Build reliable AI agents with Amazon Bedrock AgentCore Evaluations

Build reliable AI agents with Amazon Bedrock AgentCore Evaluations

Your AI agent worked in the demo, impressed stakeholders, handled test scenarios, and seemed ready for production. Then you deployed…
Build a FinOps agent using Amazon Bedrock AgentCore

Build a FinOps agent using Amazon Bedrock AgentCore

Managing costs across multiple AWS accounts often requires finance teams to query data from several sources to get a complete…
Building an AI powered system for compliance evidence collection

Building an AI powered system for compliance evidence collection

Compliance audits require comprehensive evidence trails, often involving hundreds of screenshots across multiple systems. Your compliance teams likely spend hours…
AWS introduces border agents to assess cloud security and performance

AWS introduces border agents to assess cloud security and performance

I'm excited to announce that AWS on-demand penetration testing and the AWS DevOps Agent are now generally available, representing a…
Can your governance keep pace with your AI ambitions? AI risk intelligence in the agentic era

Can your governance keep pace with your AI ambitions? AI risk intelligence in the agentic era

DevOps used to be predictable: same input, same output, binary success, static dependencies, concrete metrics. You could control what you…
ProText: A Benchmark Dataset for Measuring (Mis)gendering in Long Form Texts

ProText: A Benchmark Dataset for Measuring (Mis)gendering in Long Form Texts

Introducing ProText, a dataset for measuring gender and gender bias in English texts of various styles. ProText includes three dimensions:…
How Ring scales global customer support with Amazon Bedrock Knowledge Bases

How Ring scales global customer support with Amazon Bedrock Knowledge Bases

This post is cowritten with David Kim, and Premjit Singh from Ring. Scaling self-service support globally presents challenges beyond translation.…
Back to top button