Skip to content Skip to footer

AI Infrastructure, GPUs & Deployment — September 28, 2026

Findings

  1. [1] 2026-09-28 Manufacturing data and AI: Connecting the product value chain

    A manufacturing defect rarely belongs to one system. A scrap spike may relate to…

  2. [2] 2026-09-28 Introducing Claude Sonnet 5.5 on AWS

    Today, we’re excited to announce the availability of Claude Sonnet 5.5 on Amazon Bedrock and Claude Platform on AWS. Claude Sonnet 5.5 is a smarter, more efficient Sonnet model suited for focused coding and knowledge work with lower cost per… Dani Mitchell Dani is a Senior Specialist Solutions Architect for Generative AI at AWS, working on go-to-market for Anthropic on Amazon Bedrock. He helps enterprises across the world design and deploy generative AI solutions using Anthropic’s models and capabilities on…

  3. [3] 2026-09-28 How to roll out Genie One: A step-by-step enterprise playbook

    A regional sales director wants to know why the Northeast pipeline looks soft this quarter…

  4. [4] 2026-09-28 Build real-time voice applications with vLLM-Omni on SageMaker AI – Part 1

    Voice agents, interactive learning applications, accessibility tools, and customer service assistants need to respond without long silent pauses. In this tutorial, you deploy a text-to-speech (TTS) model on Amazon SageMaker AI that can start playing speech before it finishes generating… Clone the hosting examples repository. Clone the repository and enter the vLLM-Omni bidirectional streaming sample directory. git clone https://github.com/aws-samples/sagemaker-genai-hosting-examples.git cd sagemaker-genai-hosting-examples/03-features/bidirectional-streaming-vLLM-Omni Install the required Python packages.Create a virtual environment and install the versions defined by the sample. python3.12 -m venv…

  5. [5] 2026-09-28 Generate images and video with vLLM-Omni on SageMaker AI – Part 2

    In this post, you turn a text prompt into an image, then animate that image into a short video on Amazon SageMaker AI. You deploy two endpoints from the same AWS vLLM-Omni Deep Learning Container (DLC): a real-time endpoint for… Clone the hosting examples repository.Clone the repository and enter the vLLM-Omni image and video sample directory. git clone https://github.com/aws-samples/sagemaker-genai-hosting-examples.git cd sagemaker-genai-hosting-examples/03-features/vllm-omni-image-video Install the Python dependencies.Create a virtual environment and install the sample requirements. python -m venv .venv source .venv/bin/activate python… Yadan Wei Yadan is a Software Development Engineer on the AWS Deep Learning Containers team. She builds containers that package tested framework versions, dependencies, and AWS deployment configuration for Amazon SageMaker AI, Amazon Elastic Compute Cloud (Amazon EC2), Amazon Elastic…

  6. [6] 2026-09-28 Implementing synthetic monitoring using Amazon Nova Act

    Synthetic monitoring emulates real user journeys through automated transactions. Rather than waiting for customers to encounter problems, teams continuously validate critical workflows (logins, purchases, form submissions) on a scheduled basis. With this approach, you detect problems faster when performance degrades… An AWS account with access to Amazon Nova Act, Amazon Bedrock AgentCore (Runtime and Browser tool), Amazon Elastic Container Registry (Amazon ECR), IAM, Amazon EventBridge Scheduler, and Amazon SNS. Python 3.11 or later, Docker, and AWS Command Line Interface (AWS… The act workflow show command returns the AgentCore Runtime ARN that you reference as the Scheduler target. The sample repository’s deploy.py script wraps these commands with prerequisite checks (Docker, AWS credentials), creates the SNS alert topic, and wires the Amazon… Security and compliance Synthetic monitoring must produce reliable results without introducing false positives or masking real failures. AgentCore Runtime provides session isolation for each test execution using a strict one-session-one-microVM model built on Firecracker microVMs. By default, each session terminates…

  7. [7] 2026-09-28 Automating Amazon Textract adapter lifecycle management across accounts

    Amazon Textract is a fully managed machine learning (ML) service that automatically extracts text, handwriting, layout elements, and structured data from scanned documents. Organizations use Amazon Textract to automate document processing workflows such as invoice processing, mortgage application intake, insurance… For production deployments, route API calls through AWS PrivateLink for network isolation. IAM enforces least-privilege access. AWS CloudTrail provides API audit logging and Amazon CloudWatch handles operational monitoring and alerting. This architecture decouples adapter management from application logic. When you… AWS CLI v2 installed and configured. Sample documents (minimum 5 training and 5 test documents) for adapter training. For multi-account promotion: access to both source and destination AWS accounts in the same Region. (Optional) AWS CloudFormation or Terraform (version 1.4… This approach trades cross-account networking complexity for operational simplicity in adapter management. It works particularly well when you have a dedicated ML platform team that owns adapter training and quality. Cross-account IAM role trust policy In the hub account, create… Service quotas consideration: Before adopting this centralized hub approach, evaluate your aggregate transactions per second (TPS) requirements across all workload accounts, centralizing API calls means all environments compete for a single account’s Amazon Textract quotas. Request quota increases proactively through… Important: When you copy an adapter between accounts, only the trained model weights transfer. Query definitions and training data do not transfer. Maintain a separate configuration store (such as Parameter Store or a version-controlled config file) that maps each adapter…

  8. [8] 2026-09-28 Next.js applications, powered by Vite: introducing Vinext 1.0

    When we launched Vinext in February, it was the result of an audacious week-long AI-driven experiment to see how far one engineer, and a stack of tokens, could get to replicating the NextJS framework backed by Vite.In the seven months… Applications use generateStaticParams() and getStaticPaths() to identify pages that should be rendered when building, and they expect page-level ISR to connect those initial responses to background and on-demand revalidation.Vinext 1.0 supports that lifecycle for both routers. It can prerender App…

  9. [9] 2026-09-28 The road to the agentic browser: A Kitesurf update

    In August, we introduced Kitesurf, a browser for the agentic age that runs entirely on Cloudflare Workers.  We built it around what agents need from the web, rather than carrying all the features and bloat of a browser designed for… Now you can also use them from inside a Worker script using the env.BROWSER.quickAction() binding:Kitesurf runs in the terminal nowAs we detailed in the How we built it section of our announcement blog post, Kitesurf separates PageScript, the isolate that handles the page…

  10. [10] 2026-09-28 NVIDIA Announces a $150 Billion Share Repurchase Authorization Increase

    NVIDIA today announced that its Board of Directors has authorized an additional $150 billion under the company’s existing share repurchase program, increasing the total remaining amount authorized to $235 billion.

  11. [11] 2026-09-28 NVIDIA Launches Open Agent Safety Platform to Secure Agents From Testing to Deployment

    NVIDIA today announced NVIDIA Open Agent Safety Platform, an open software platform and reference system design to strengthen AI security from agent testing to deployment, with full-stack governance and control across software and the hardware, compute and robotics systems that run agents.

  12. [12] 2026-09-28 NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring

    To understand where agentic AI stands today, consider the last seismic shift in technology: the rise of the internet in the 90s. It was new and full of…To understand where agentic AI stands today, consider the last seismic shift in technology: the rise of the internet in the 90s. It was new and full of possibilities. You could build a…

  13. [13] 2026-09-28 Add Runtime Controls to AI Agents with NVIDIA OpenShell

    AI agents can be given a goal, write code, use tools, and keep working as new information becomes available. This opens the door to applications that…AI agents can be given a goal, write code, use tools, and keep working as new information becomes available. This opens the door to applications that investigate software failures, run experiments, and carry out business-critical…

  14. [14] 2026-09-28 The 24 most commonly misunderstood marketing data terms

    Imagine you’re a marketer planning a win-back campaign and you ask your data team for a list of “inactive customers.” Y…

  15. [15] 2026-09-28 How NVIDIA DSX MaxLPS Maximizes AI Factory Throughput and Efficiency

    Every unused watt is capacity left on the table. AI factories are typically provisioned for the unlikely moment when every GPU reaches peak power, creating a…Every unused watt is capacity left on the table. AI factories are typically provisioned for the unlikely moment when every GPU reaches peak power, creating a protective buffer that can leave valuable infrastructure underused during…

Where Kimbodo Comes In

Kimbodo builds and operates this in production for businesses — see our AI Infrastructure & MLOps practice, or Estimate My Infrastructure.

Sources

  1. [1] Manufacturing data and AI: Connecting the product value chain
  2. [2] Introducing Claude Sonnet 5.5 on AWS
  3. [3] How to roll out Genie One: A step-by-step enterprise playbook
  4. [4] Build real-time voice applications with vLLM-Omni on SageMaker AI – Part 1
  5. [5] Generate images and video with vLLM-Omni on SageMaker AI – Part 2
  6. [6] Implementing synthetic monitoring using Amazon Nova Act
  7. [7] Automating Amazon Textract adapter lifecycle management across accounts
  8. [8] Next.js applications, powered by Vite: introducing Vinext 1.0
  9. [9] The road to the agentic browser: A Kitesurf update
  10. [10] NVIDIA Announces a $150 Billion Share Repurchase Authorization Increase
  11. [11] NVIDIA Launches Open Agent Safety Platform to Secure Agents From Testing to Deployment
  12. [12] NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring
  13. [13] Add Runtime Controls to AI Agents with NVIDIA OpenShell
  14. [14] The 24 most commonly misunderstood marketing data terms
  15. [15] How NVIDIA DSX MaxLPS Maximizes AI Factory Throughput and Efficiency

Leave a comment

0.0/5