Azure AI Search

AzureAI & MLFree tier available

Enterprise search-as-a-service (formerly Azure Cognitive Search) with vector, hybrid, and semantic ranking, built-in AI skills for OCR and NLP enrichment, first-class integration with Azure OpenAI for RAG workloads, and 90+ data-source connectors including SharePoint, OneDrive, and Salesforce

Jurisdictional exposure

Provider HQ
USRedmond, USA

Subject to CLOUD Act, FISA-702, DPF

Region locations
APACCNEEAEUUKUSOther73 regions across 7 jurisdictions
Sovereign option
Yes — 13 sovereign-flagged regions available

Attributes

SLA Uptime
99.9%
Vector Search
Yes

Sub-services (4)

Vector Search

ANN vector indexing with HNSW and exhaustive KNN for RAG workloads

Semantic Ranking

Deep-learning re-ranker that surfaces the most relevant chunks

AI Enrichment Skillsets

Pipeline of OCR, entity, and key-phrase skills applied at ingestion

Integrated Vectorisation

Built-in chunking and embedding via Azure OpenAI at index-time

Compliance & Certifications

This service is attested for the following frameworks. Always verify with the provider before relying on a specific compliance posture.

Where this runs

73 regions
36 countries
13sovereign
Sovereign regions (13)
  • Australia Central · CanberraAzure Australia Government
  • Australia Central 2 · CanberraAzure Australia Government
  • US Gov Virginia · VirginiaAzure Government (US)
  • US Gov Arizona · ArizonaAzure Government (US)
  • US Gov Texas · TexasAzure Government (US)
  • US DoD East · VirginiaAzure Government Secret (US)
  • US DoD Central · IowaAzure Government Secret (US)
  • China North (Beijing) · BeijingMicrosoft Azure China (21Vianet)
  • China East (Shanghai) · ShanghaiMicrosoft Azure China (21Vianet)
  • China North 2 · BeijingMicrosoft Azure China (21Vianet)
  • China East 2 · ShanghaiMicrosoft Azure China (21Vianet)
  • China North 3 · HebeiMicrosoft Azure China (21Vianet)
  • China East 3 · ShanghaiMicrosoft Azure China (21Vianet)
Commercial regions (60)

Europe (21)

  • Austria East
  • Belgium Central
  • Denmark East
  • Finland Central
  • France South
  • France Central
  • Germany North
  • Germany West Central
  • Greece Central
  • North Europe
  • Italy North
  • West Europe
  • Norway East
  • Norway West
  • Poland Central
  • Spain Central
  • Sweden Central
  • Switzerland West
  • Switzerland North
  • UK West
  • UK South

North America (13)

  • Canada East
  • Canada Central
  • Mexico Central
  • West US
  • East US 3
  • North Central US
  • Central US
  • West US 3
  • South Central US
  • East US
  • East US 2
  • West US 2
  • West Central US

South America (3)

  • Brazil Southeast
  • Brazil South
  • Chile Central

Asia (13)

  • East Asia
  • South India
  • Jio India West
  • West India
  • Jio India Central
  • Central India
  • Indonesia Central
  • Japan West
  • Japan East
  • Malaysia West
  • Southeast Asia
  • Korea South
  • Korea Central

Oceania (3)

  • Australia East
  • Australia Southeast
  • New Zealand North

Middle East (5)

  • Israel Central
  • Qatar Central
  • Saudi Arabia Central
  • UAE Central
  • UAE North

Africa (2)

  • South Africa West
  • South Africa North

Tags

Equivalent services on other platforms

Alibaba Qwen (Tongyi Qianwen)Alibaba

Alibaba's flagship open-source foundation model family covering Qwen (text), Qwen-VL (vision-language), Qwen-Audio, and Qwen-Coder — accessible via the DashScope API with chat, completion, embeddings, and function-calling endpoints

Amazon OpenSearch ServiceAWS

Managed OpenSearch and Elasticsearch 7.10 compatible search, log analytics, and vector search service with Multi-AZ, cold / UltraWarm storage tiers, OpenSearch Ingestion (Data Prepper), and a serverless option for variable workloads

Amazon BedrockAWS

Build generative AI applications with foundation models from Anthropic (Claude Opus 4.7 from April 2026), Cohere, Meta, Mistral, Stability AI, TwelveLabs (video understanding), and Amazon's own Nova family — accessed via a single API with fine-tuning, knowledge bases, agents, and a model marketplace for discovery and easy onboarding

Amazon NovaAWS

AWS-built foundation model family covering text (Micro, Lite, Pro, Premier), image generation (Canvas), and video generation (Reel) — accessed through the Bedrock runtime with tight pricing and low-latency streaming, launched at re:Invent 2024

Amazon S3 VectorsAWS

Native vector storage in S3 — up to 2 billion vectors per index, sub-100 ms query latency, S3-native durability, and pricing claimed up to 90 percent lower than dedicated vector databases for retrieval-augmented generation and embedding-heavy workloads

Cloudflare AI SearchCloudflare

Managed retrieval-augmented-generation service — index your content (R2 buckets, websites, Workers KV) and query it with natural language from a Workers binding, REST API, or MCP server. Originally launched as AutoRAG and renamed AI Search in 2026.

Mosaic AIDatabricks

End-to-end AI platform (formerly MLflow + Mosaic ML) for training, fine-tuning, deploying, and monitoring foundation models and custom ML models on the Lakehouse

Gcore Inference at the EdgeGcore

AI inference runtime that deploys models to Gcore's edge POPs and routes requests to the nearest GPU-backed endpoint, with support for open-source LLMs and custom model containers

Vertex AIGCP

Unified platform to build, deploy, and scale ML models with AutoML, custom training on TPUs and GPUs, model registry, pipelines, feature store, and generative AI studio

IBM watsonx.aiIBM

Enterprise AI studio for training, validating, tuning, and deploying foundation models and traditional ML models, with IBM's Granite model family, Hugging Face integration, prompt lab, synthetic data generation, and governance via watsonx.governance

Kanana AIKakao

Kakao's Korean-first foundation-model family (Kanana Flash / Essence / Nano) for chat, code, and embeddings — multilingual but tuned for Korean conversational performance

CLOVA StudioNaver

Naver's HyperCLOVA X foundation-model platform for Korean-language LLM workloads — chat completion, embeddings, function calling, RAG over Korean text with strong native-language performance

Red Hat OpenShift AIOpenShift

Managed MLOps platform (formerly Open Data Hub) for training, serving, and monitoring ML models on OpenShift with JupyterHub, KServe, Kubeflow, and PyTorch operators

OCI Generative AIOracle

Managed inference service hosting Cohere Command and Embed plus Meta Llama large language models — pay-per-token chat / completion / embedding APIs, plus fine-tuning on customer datasets via dedicated AI clusters

OCI Enterprise AIOracle

End-to-end platform for building, deploying, and governing production AI workloads on OCI — unifies Gen AI models, agent orchestration, retrieval-augmented generation, and policy-based governance controls in a single managed service so enterprises don't have to assemble them from primitives.

Generative APIsScaleway

Managed inference for open-source LLMs (Llama, Mistral, DeepSeek) hosted in EU datacentres

Snowflake CortexSnowflake

Fully managed AI and ML service offering hosted LLMs, vector search, and ML functions inside Snowflake SQL

Tencent HunyuanTencent

Tencent's in-house family of large language models (Hunyuan-Pro, Standard, Lite, plus multimodal Hunyuan-Vision) accessible via the Hunyuan API, with enterprise-grade context windows up to 256K, function calling, embeddings, and tuning

Pricing

Pricing model:pay-as-you-go