Quotas for the bedrock-runtime endpoint
Added a dedicated page that describes the per-model token quotas that govern inference traffic to the <code class="code">bedrock-runtime</code> endpoint and points to related quotas for custom inference profiles, batch inference, and Provisioned Throughput.