You can now send one prompt to a panel of models in parallel, then have a judge model synthesize their answers into a single response.
Read more in the Merge Fusion docs.
You can now set indirect prompt injection protection per axis, picking a mode (off, alert, or block) and a sensitivity threshold for each one.
Test the configuration against injection patterns with a built-in detection tester in Security settings before it goes live. Read more in the prompt injection protection docs.
You can now set a service_tier request parameter to toggle between flex and priority processing, trading cost for latency or the reverse.
Per-tier pricing and billing now show up directly on /v1/models, so the tradeoff is visible before you pick a tier. Read more in the service tiers docs.