Showing 41–47 of 47 dossiers

Azure Document Intelligence v2.0 API retires August 31 — old integrations need a version migration

Azure Document Intelligence v2.0 reaches retirement on August 31, 2026. Microsoft recommends moving workloads to the current v4.0 API; the post-v2 REST surface was redesigned, so teams should verify the actual api-version their SDK or HTTP client sends rather than assuming a package upgrade is enough.

Gemini 3.8 Flash raises agent capability at the same token price — but may use more tokens per task

Gemini 3.8 Flash keeps 3.7 Flash’s promotional per-token rate and Flash-tier latency, but early independent analysis suggests harder reasoning can increase tokens consumed per task. A separate 3.8 Flash Cyber model is available only through Google’s Fairwind defensive-security program.

Inference is where an AI product meets its latency target, reliability budget and monthly bill. Model quality matters, but so do rate limits, caching, batching, regional availability, data terms, observability and the provider behaviour that only appears under production traffic.

BTN tracks important API launches, price changes and serving techniques across hosted and self-managed systems. Coverage connects provider documentation with benchmarks and operating experience so builders can compare more than headline token prices. The useful outcome is knowing when an infrastructure change makes a product newly viable, when migration is worth the work and where apparent savings hide another constraint.

The beat includes routing layers, gateways and compatibility standards when they reduce switching cost or improve control. It also watches changes to retention, abuse monitoring and service terms, because the fastest endpoint is not a safe default if its data handling conflicts with the product being built.