Reactive Machines

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

Best practices for applying Amazon Bedrock Guardrails to code generation workflows

This post continues our series on best practices with Amazon Bedrock Guardrails. For the previous post, see Build safe generative…
Evaluating AI Agents: A production blueprint with Strands and AgentCore

Evaluating AI Agents: A production blueprint with Strands and AgentCore

This post was co-written with Motorway and the AWS Prototyping and AI Customer Engineering (PACE) team. Motorway, a UK-based online…
AI Teammates: how monday.com runs production AI agents on Amazon Bedrock

AI Teammates: how monday.com runs production AI agents on Amazon Bedrock

AI Teammates are agentic AI on Amazon Bedrock, and few engineering organizations run them in production at the scale that…
Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova

Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova

When you fine-tune a model using Supervised Fine-Tuning (SFT), creating high-quality chain-of-thought (CoT) reasoning traces for your training data is…
Environment-free Synthetic Data Generation for API-Calling Agents

Environment-free Synthetic Data Generation for API-Calling Agents

API training called large language modeling (LLM) requires large amounts of high quality. However, collecting such data at scale often…
Accelerating Text-to-Video Production with Limited Attention

Accelerating Text-to-Video Production with Limited Attention

The latest distribution models enable high-quality video production, but suffer from slow runtimes. The large transformer-based backbones used in these…
Custom OS installations are now available for AWS DeepRacer devices

Custom OS installations are now available for AWS DeepRacer devices

With stock firmware and software, developers could not modify their AWS DeepRacer devices to run the latest operating systems. Now,…
RayRoPE: Projective Ray Positional Encoding for Multi-View Attention

RayRoPE: Projective Ray Positional Encoding for Multi-View Attention

We study the spatial encoding of multi-view converters that process tokens from a set of input images, and seek a…
LVSum: A Timestamp Benchmark for Long Video Summarization

LVSum: A Timestamp Benchmark for Long Video Summarization

Long video summarization presents significant challenges for large-scale linguistic models (MLLMs), especially in maintaining temporal fidelity over extended time periods…
Longitudinal Value Modeling: Training of Quantitative Value for Token Level Modeling

Longitudinal Value Modeling: Training of Quantitative Value for Token Level Modeling

A token serves as the basic unit of calculation in modern automation models, and the length of production directly influences…
Back to top button