Reactive Machines

Build an explainable next-best-product recommendation system for banking on AWS

Build an explainable next-best-product recommendation system for banking on AWS

Building a deep learning-based explainable next-best-product recommendation system helps banking institutions predict which product a customer needs next. Banks hold…
HOLAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

HOLAD: Breaking the No-Recovery Bottleneck in Long-Horizon Reasoning

Long-term modeling of Large-scale Language Models (LLMs) remains unstable even when advanced techniques are provided. Examining controlled algorithmic puzzles, we…
Best practices for applying Amazon Bedrock Guardrails to code generation workflows

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

This post continues our series on best practices with Amazon Bedrock Guardrails. For the previous post, see Build safe generative…
Evaluating AI Agents: A production blueprint with Strands and AgentCore

Evaluating AI Agents: A production blueprint with Strands and AgentCore

This post was co-written with Motorway and the AWS Prototyping and AI Customer Engineering (PACE) team. Motorway, a UK-based online…
AI Teammates: how monday.com runs production AI agents on Amazon Bedrock

AI Teammates: how monday.com runs production AI agents on Amazon Bedrock

AI Teammates are agentic AI on Amazon Bedrock, and few engineering organizations run them in production at the scale that…
Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova

Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova

When you fine-tune a model using Supervised Fine-Tuning (SFT), creating high-quality chain-of-thought (CoT) reasoning traces for your training data is…
Environment-free Synthetic Data Generation for API-Calling Agents

Environment-free Synthetic Data Generation for API-Calling Agents

API training called large language modeling (LLM) requires large amounts of high quality. However, collecting such data at scale often…
Accelerating Text-to-Video Production with Limited Attention

Accelerating Text-to-Video Production with Limited Attention

The latest distribution models enable high-quality video production, but suffer from slow runtimes. The large transformer-based backbones used in these…
Custom OS installations are now available for AWS DeepRacer devices

Custom OS installations are now available for AWS DeepRacer devices

With stock firmware and software, developers could not modify their AWS DeepRacer devices to run the latest operating systems. Now,…
RayRoPE: Projective Ray Positional Encoding for Multi-View Attention

RayRoPE: Projective Ray Positional Encoding for Multi-View Attention

We study the spatial encoding of multi-view converters that process tokens from a set of input images, and seek a…
Back to top button