A company is designing an API for a generative AI (GenAI) application that uses a foundation model (FM) that is hosted on a managed model service. The API must stream responses to reduce latency, enforce token limits to manage compute resource usage, and implement retry logic to handle model timeouts and partial responses.
Which solution will meet these requirements with the LEAST operational overhead?
A logistics company is using Amazon Bedrock to build an autonomous routing agent that coordinates with APIs that support warehouse, shipping, and international customs operations. The agent must meet the following requirements:
• Break requests into reasoning steps.
• Retry failed tool calls with backoff.
• Stop retrying after three consecutive failures.
• Require human approval for shipments that are valued over $100,000.
• Use MCP to provide access to tools and new integrations without requiring code changes.
Which combination of solutions will meet these requirements? (Select THREE.)
A book publishing company wants to build a book recommendation system that uses an AI assistant. The AI assistant will use ML to generate a list of recommended books from the company ' s book catalog. The system must suggest books based on conversations with customers.
The company stores the text of the books, customers ' and editors ' reviews of the books, and extracted book metadata in Amazon S3. The system must support low-latency responses and scale efficiently to handle more than 10,000 concurrent users.
Which solution will meet these requirements?
A financial services company is deploying a generative AI (GenAI) application that uses Amazon Bedrock to assist customer service representatives to provide personalized investment advice to customers. The company must implement a comprehensive governance solution that follows responsible AI practices and meets regulatory requirements.
The solution must detect and prevent hallucinations in recommendations. The solution must have safety controls for customer interactions. The solution must also monitor model behavior drift in real time and maintain audit trails of all prompt-response pairs for regulatory review. The company must deploy the solution within 60 days. The solution must integrate with the company ' s existing compliance dashboard and respond to customers within 200 ms.
Which solution will meet these requirements with the LEAST operational overhead?
A company uses Amazon Bedrock to build a Retrieval Augmented Generation (RAG) system. The RAG system uses an Amazon Bedrock Knowledge Bases that is based on an Amazon S3 bucket as the data source for emergency news video content. The system retrieves transcripts, archived reports, and related documents from the S3 bucket.
The RAG system uses state-of-the-art embedding models and a high-performing retrieval setup. However, users report slow responses and irrelevant results, which cause decreased user satisfaction. The company notices that vector searches are evaluating too many documents across too many content types and over long periods of time.
The company determines that the underlying models will not benefit from additional fine-tuning. The company must improve retrieval accuracy by applying smarter constraints and wants a solution that requires minimal changes to the existing architecture.
Which solution will meet these requirements?
A company is building a meeting analysis solution for its executive team. The solution uses AWS generative AI services. The solution must extract speaker-attributed content from recorded meetings, analyze visual elements from presentation slides, and create searchable summaries that link speaker comments to relevant visual context.
The solution must process 200 hours of meeting recordings each week. The solution must maintain data privacy by processing all meeting data within the AWS Cloud. The solution must store the source data for future retrieval and must be able to perform full-text searches.
Which solution will meet these requirements with the LEAST operational overhead?
A company is designing a solution that uses foundation models (FMs) to support multiple AI workloads. Some FMs must be invoked on demand and in real time. Other FMs require consistent high-throughput access for batch processing.
The solution must support hybrid deployment patterns and run workloads across cloud infrastructure and on-premises infrastructure to comply with data residency and compliance requirements.
Which combination of steps will meet these requirements? (Select TWO.)
A company is using Amazon Bedrock to build an AI assistant to help internal teams analyze unstructured customer feedback data. The company stores the customer feedback in an Amazon S3 bucket. The S3 bucket contains more than 25 TB of historical data from mobile app reviews, chat conversations, and call center transcripts. The company expects the data source to grow by 3 GB every day. The data entries often contain multiple unrelated topics within the same input.
The company needs a solution that reliably delivers accurate answers to questions based on the data source. The solution must not export any personally identifiable information (PII) to the Amazon Bedrock model during processing or response generation.
Which solution will meet these requirements with the LEAST operational overhead?
A media company is building an AI-powered content moderation system by using Amazon Bedrock. The system first classifies text by using a small, low-latency model. Then the system escalates requests that have a confidence score below 0.65 to a larger, more expensive model.
The system must respond in near real time for high-confidence results. The system must process low-confidence requests asynchronously. The system must scale to meet sudden spikes in demand. The company wants to optimize costs for the system by invoking the larger model only when required. The company wants to use decoupled components to achieve high resiliency for the system.
Which solution will meet these requirements?
A healthcare company is using Amazon Bedrock to build a GenAI application to analyze patient feedback data from CSV files, JSON documents, and text files. The company needs to make the data available for a RAG solution that requires high data quality to prevent hallucinations. The GenAI application will use the data to make accurate clinical recommendations. The application must be highly scalable to handle data in near real time. Data attrition is also high.
Before the company feeds data to a foundation model (FM), the company needs to validate data completeness, detect anomalies, remove personally identifiable information (PII), and monitor quality metrics. The application must be serverless, provide automated rule recommendations, and generate quality scores for regulatory compliance.
Which solution will meet these requirements?