Artificial Intelligence

Category: Artificial Intelligence

Grok 4.7 is now available on Amazon Bedrock

Grok 4.7 is now available on Amazon Bedrock

xAI’s Grok 4.7 is now available on Amazon Bedrock: a frontier model for coding, long-running agents, and knowledge work. It offers a 500K token context window and four configurable reasoning effort levels, reachable through the Responses, Chat Completions, and Converse APIs.

Generate images and video with vLLM-Omni on SageMaker AI - Part 2

Generate images and video with vLLM-Omni on SageMaker AI – Part 2

Deploy two generative media models from one AWS vLLM-Omni Deep Learning Container on Amazon SageMaker AI. Generate an image with FLUX.2-klein through real-time inference, then animate it into video with Wan2.1-VACE through asynchronous inference, and retrieve the MP4 from Amazon S3.

Automating Amazon Textract adapter lifecycle management across accounts

Automating Amazon Textract adapter lifecycle management across accounts

Learn how to operationalize Amazon Textract Custom Queries adapters for production: infrastructure as code with AWS CloudFormation and Terraform, a cross-account adapter promotion process, a pre-classification routing pattern for multiple form versions, and production security controls such as VPC endpoints, encryption, and least-privilege IAM.

Accelerate multimodal RL training with SkyRL on Amazon SageMaker HyperPod

Accelerate multimodal RL training with SkyRL on Amazon SageMaker HyperPod

Learn how to run SkyRL, an open-source reinforcement learning framework, on Amazon SageMaker HyperPod to post-train a Qwen3-VL-8B vision-language model with GRPO. This walkthrough covers building the container image, launching a Ray cluster from SageMaker Studio, submitting and monitoring the job, and hosting the trained LoRA adapter for inference.

NarrateAI: production-ready LLM quality assurance on Amazon Bedrock

NarrateAI: production-ready LLM quality assurance on Amazon Bedrock

NarrateAI delivers production-ready LLM quality assurance on Amazon Bedrock. This post details five techniques—adaptive pipeline orchestration, cross-account multi-model failover, real-time streaming evaluation, composite evaluation, and data accuracy verification—that reach about 99% numerical accuracy while streaming responses in real time.

Multi-Region training with Amazon SageMaker HyperPod and Qumulo

Multi-Region training with Amazon SageMaker HyperPod and Qumulo

Amazon SageMaker HyperPod and Cloud Native Qumulo let you place training compute in one AWS Region while keeping your dataset in another. This post shares the architecture and validation results from a cross-Region training run, where a remote cluster matched a co-located cluster’s throughput after a brief NeuralCache warmup.