Back to AI Roadmap
🛠️
LAYER 07

07. AI Platform / MLOps / LLMOps

Benchmark Companies:
MLflowMLflowW&BW&BVertex AIVertex AISageMakerSageMakerAzure Machine LearningAzure Machine LearningOpenRouterOpenRouterBentoMLBentoMLKubeflowKubeflowPortkeyPortkeyLiteLLMLiteLLMCloudflareCloudflareHeliconeHeliconeKongKongBraintrustBraintrustLangfuseLangfuseArize AIArize AIPromptfooPromptfooHumanloopHumanloopDatadogDatadogOpenTelemetryOpenTelemetryCloudZeroCloudZeroDatabricksDatabricks

Industrializing the model lifecycle: model registries, automated CI/CD evaluation gates, intelligent model gateways, distributed tracing, and FinOps cost governance.

📊Layer View Mode:
🌐3D CROSS-CORRELATION MATRIX

07. AI Platform / MLOps / LLMOps · Cross-Correlation Ecosystem

Click any company, tech category, or career role to illuminate all cross-associations and dim unrelated entities.

🔍
🏢

Benchmark Enterprises

22
MLflowMLflow
3
W&BW&B
3
Vertex AIVertex AI
1
SageMakerSageMaker
2
Azure Machine LearningAzure Machine Learning
1
OpenRouterOpenRouter
2
BentoMLBentoML
2
KubeflowKubeflow
1
PortkeyPortkey
3
LiteLLMLiteLLM
2
CloudflareCloudflare
2
HeliconeHelicone
4
KongKong
1
BraintrustBraintrust
2
LangfuseLangfuse
4
Arize AIArize AI
2
PromptfooPromptfoo
2
HumanloopHumanloop
1
DatadogDatadog
1
OpenTelemetryOpenTelemetry
1
CloudZeroCloudZero
1
DatabricksDatabricks
3
🛠️

Tech Stack Categories & Atomic Nodes

4 Tracks
7.1

7.1 Model Lifecycle, Registry & Deployment

MLflow & W&B experiment metadata, end-to-end data/code/model lineage, GitOps model deployment, canary traffic shifting, and automated rollback.
Sub-Page
Model Registry, Lineage & Experiment Tracking
💰 $195K - $400K / year (MLOps & Model Registry) | ¥550K - ¥1.25M / year
Unified model artifact registry with MLflow & W&B, experiment hyperparameter tracking, full data/code/model lineage, S3/GCS artifact integrity, and multi-stage promotion workflows.
Benchmark:MLflowMLflowW&BW&BVertex AIVertex AISageMakerSageMakerBentoMLBentoML
Model RegistryMLflowWeights & BiasesExperiment TrackingLineageVertex AISageMaker
Model CI/CD, Canary Rollouts & Automated Rollback
💰 $190K - $390K / year (Model CI/CD & Deployment) | ¥520K - ¥1.2M / year
GitOps-based declarative model deployment, Kubeflow/BentoML pipelines, progressive canary traffic shifting (1%->10%->100%), blue-green deployments, and instant automated rollback.
Benchmark:BentoMLBentoMLKubeflowKubeflowMLflowMLflowSageMakerSageMakerAzure Machine LearningAzure Machine Learning
Model CI/CDGitOpsCanary DeploymentBlue-GreenAutomated RollbackKubeflowBentoML
7.2

7.2 Model Gateway, Routing & Traffic Control

Unified OpenAI-compatible API gateway (OpenRouter/LiteLLM/Portkey), multi-cloud load balancing, automatic failover, semantic intent routing, and multi-tenant rate limiting.
Sub-Page
Unified Model Gateway, Provider Abstraction & Failover
💰 $195K - $410K / year (LLM Gateway Architecture) | ¥520K - ¥1.2M / year
Unified OpenAI-compatible API abstraction, intelligent multi-provider load balancing, automatic timeout retries and fallback failover, high-performance SSE streaming, and enterprise audit logs.
Benchmark:OpenRouterOpenRouterPortkeyPortkeyLiteLLMLiteLLMCloudflareCloudflareHeliconeHeliconeKongKong
Model GatewayOpenRouterLiteLLMPortkeyProvider AbstractionFallback RetriesCircuit BreakerSSE Streaming
Semantic Intent Routing, Multi-Tenancy & Rate Limiting
💰 $190K - $395K / year (Semantic Routing & Rate Limiting) | ¥500K - ¥1.15M / year
Semantic intent routing via lightweight embeddings/SLMs (distributing simple queries to SLMs, deep reasoning to o1/R1), multi-tenant quota isolation, and Redis distributed token bucket rate limiting.
Benchmark:OpenRouterOpenRouterPortkeyPortkeyLiteLLMLiteLLMCloudflareCloudflareHeliconeHelicone
Semantic RoutingOpenRouterModel RoutingRate LimitingToken BucketMulti-TenancyQuota ManagementSLM Offloading
7.3

7.3 Evaluation, PromptOps & Quality Gates

Prompt-as-code versioning, environment config injection, golden benchmark datasets, automated LLM-as-a-judge calibration, and regression gates.
Sub-Page
Prompt Registry & PromptOps Lifecycle Management
💰 $175K - $360K / year (PromptOps & Prompt Management) | ¥450K - ¥1.05M / year
Prompt-as-code management, dynamic parameter templating, version tagging, GitOps review workflows, A/B prompt experiment testing, and team collaboration workspaces.
Benchmark:W&BW&BBraintrustBraintrustLangfuseLangfuseHumanloopHumanloopPromptfooPromptfoo
PromptOpsPrompt RegistryPrompt-as-CodeA/B TestingGitOpsBraintrustHumanloop
Golden Datasets, LLM-as-a-Judge & Automated Eval CI/CD
💰 $185K - $380K / year (Automated AI Evaluation & CI/CD) | ¥480K - ¥1.15M / year
Golden evaluation dataset curation, multi-criteria LLM-as-a-judge scoring calibrated against human review, statistical significance testing (Pass@k), and automated pre-deployment regression gates.
Benchmark:BraintrustBraintrustLangfuseLangfuseArize AIArize AIW&BW&BPromptfooPromptfoo
LLM-as-a-JudgeEval CI/CDGolden DatasetRegression TestingStatistical SignificanceBraintrustArize AI
7.4

7.4 Observability, Cost & FinOps Governance

OpenTelemetry-compatible distributed tracing, end-to-end span debugging, real-time token cost attribution, multi-tenant chargeback, and FinOps budget controls.
Sub-Page
LLM Distributed Tracing, Span Debugging & APM
💰 $195K - $405K / year (AI Observability & Distributed Tracing) | ¥520K - ¥1.22M / year
OpenTelemetry-compatible distributed tracing, nested agent/tool/RAG span inspection, TTFT and TPOT anomaly diagnosis, payload auditing, and automated real-time alert triggers.
Benchmark:LangfuseLangfuseArize AIArize AIDatadogDatadogHeliconeHeliconeOpenTelemetryOpenTelemetry
Distributed TracingOpenTelemetrySpan InspectionLangfuseArize PhoenixDatadogAPMTTFT / TPOT
Token Cost Attribution, Multi-Tenant Chargeback & FinOps
💰 $185K - $385K / year (Token Cost & AI FinOps) | ¥480K - ¥1.15M / year
Real-time token cost calculation across multi-cloud providers (prompt/completion/reasoning token pricing), departmental chargeback models, budget overrun triggers, cache hit analysis, and ROI efficiency reporting.
Benchmark:LangfuseLangfusePortkeyPortkeyHeliconeHeliconeCloudZeroCloudZeroDatabricksDatabricks
Token CostFinOpsCost AttributionChargebackBudget AlertsPrompt CachingCloudZeroLangfuse
💼

Career Track Roles

13
💼MLOps Engineer
2
💼LLMOps Engineer
3
💼AI Platform Engineer
4
💼ML Infrastructure Engineer
2
💼Evaluation Engineer
1
💼QA Automation Engineer
2
💼FinOps Engineer
1
💼DevOps / SRE
3
💼API Platform Engineer
1
💼Backend Architect
2
💼Reliability Engineer
1
💼Applied ML Engineer
1
💼Observability Engineer
1
🔄Sub-Domain Sequential Path (4 stages):
7.1

7.1 Model Lifecycle, Registry & Deployment

Open Sub-Page

MLflow & W&B experiment metadata, end-to-end data/code/model lineage, GitOps model deployment, canary traffic shifting, and automated rollback.

Model Registry, Lineage & Experiment Tracking

Unified model artifact registry with MLflow & W&B, experiment hyperparameter tracking, full data/code/model lineage, S3/GCS artifact integrity, and multi-stage promotion workflows.

🏢 Companies
MLflowMLflowW&BW&BVertex AIVertex AISageMakerSageMakerBentoMLBentoML
🛠️ Tech Stack
Model RegistryMLflowWeights & BiasesExperiment TrackingLineageVertex AISageMaker
💼 Roles & Salary
MLOps Engineer、ML Infrastructure Engineer、AI Platform Engineer
💰 $195K - $400K / year (MLOps & Model Registry) | ¥550K - ¥1.25M / year

Model CI/CD, Canary Rollouts & Automated Rollback

GitOps-based declarative model deployment, Kubeflow/BentoML pipelines, progressive canary traffic shifting (1%->10%->100%), blue-green deployments, and instant automated rollback.

🏢 Companies
BentoMLBentoMLKubeflowKubeflowMLflowMLflowSageMakerSageMakerAzure Machine LearningAzure Machine Learning
🛠️ Tech Stack
Model CI/CDGitOpsCanary DeploymentBlue-GreenAutomated RollbackKubeflowBentoML
💼 Roles & Salary
MLOps Engineer、DevOps / SRE、ML Infrastructure Engineer
💰 $190K - $390K / year (Model CI/CD & Deployment) | ¥520K - ¥1.2M / year
7.2

7.2 Model Gateway, Routing & Traffic Control

Open Sub-Page

Unified OpenAI-compatible API gateway (OpenRouter/LiteLLM/Portkey), multi-cloud load balancing, automatic failover, semantic intent routing, and multi-tenant rate limiting.

Unified Model Gateway, Provider Abstraction & Failover

Unified OpenAI-compatible API abstraction, intelligent multi-provider load balancing, automatic timeout retries and fallback failover, high-performance SSE streaming, and enterprise audit logs.

🏢 Companies
OpenRouterOpenRouterPortkeyPortkeyLiteLLMLiteLLMCloudflareCloudflareHeliconeHeliconeKongKong
🛠️ Tech Stack
Model GatewayOpenRouterLiteLLMPortkeyProvider AbstractionFallback RetriesCircuit BreakerSSE Streaming
💼 Roles & Salary
LLMOps Engineer、API Platform Engineer、Backend Architect
💰 $195K - $410K / year (LLM Gateway Architecture) | ¥520K - ¥1.2M / year

Semantic Intent Routing, Multi-Tenancy & Rate Limiting

Semantic intent routing via lightweight embeddings/SLMs (distributing simple queries to SLMs, deep reasoning to o1/R1), multi-tenant quota isolation, and Redis distributed token bucket rate limiting.

🏢 Companies
OpenRouterOpenRouterPortkeyPortkeyLiteLLMLiteLLMCloudflareCloudflareHeliconeHelicone
🛠️ Tech Stack
Semantic RoutingOpenRouterModel RoutingRate LimitingToken BucketMulti-TenancyQuota ManagementSLM Offloading
💼 Roles & Salary
LLMOps Engineer、Backend Architect、Reliability Engineer
💰 $190K - $395K / year (Semantic Routing & Rate Limiting) | ¥500K - ¥1.15M / year
7.3

7.3 Evaluation, PromptOps & Quality Gates

Open Sub-Page

Prompt-as-code versioning, environment config injection, golden benchmark datasets, automated LLM-as-a-judge calibration, and regression gates.

Prompt Registry & PromptOps Lifecycle Management

Prompt-as-code management, dynamic parameter templating, version tagging, GitOps review workflows, A/B prompt experiment testing, and team collaboration workspaces.

🏢 Companies
W&BW&BBraintrustBraintrustLangfuseLangfuseHumanloopHumanloopPromptfooPromptfoo
🛠️ Tech Stack
PromptOpsPrompt RegistryPrompt-as-CodeA/B TestingGitOpsBraintrustHumanloop
💼 Roles & Salary
LLMOps Engineer、QA Automation Engineer、AI Platform Engineer
💰 $175K - $360K / year (PromptOps & Prompt Management) | ¥450K - ¥1.05M / year

Golden Datasets, LLM-as-a-Judge & Automated Eval CI/CD

Golden evaluation dataset curation, multi-criteria LLM-as-a-judge scoring calibrated against human review, statistical significance testing (Pass@k), and automated pre-deployment regression gates.

🏢 Companies
BraintrustBraintrustLangfuseLangfuseArize AIArize AIW&BW&BPromptfooPromptfoo
🛠️ Tech Stack
LLM-as-a-JudgeEval CI/CDGolden DatasetRegression TestingStatistical SignificanceBraintrustArize AI
💼 Roles & Salary
Evaluation Engineer、QA Automation Engineer、Applied ML Engineer
💰 $185K - $380K / year (Automated AI Evaluation & CI/CD) | ¥480K - ¥1.15M / year
7.4

7.4 Observability, Cost & FinOps Governance

Open Sub-Page

OpenTelemetry-compatible distributed tracing, end-to-end span debugging, real-time token cost attribution, multi-tenant chargeback, and FinOps budget controls.

LLM Distributed Tracing, Span Debugging & APM

OpenTelemetry-compatible distributed tracing, nested agent/tool/RAG span inspection, TTFT and TPOT anomaly diagnosis, payload auditing, and automated real-time alert triggers.

🏢 Companies
LangfuseLangfuseArize AIArize AIDatadogDatadogHeliconeHeliconeOpenTelemetryOpenTelemetry
🛠️ Tech Stack
Distributed TracingOpenTelemetrySpan InspectionLangfuseArize PhoenixDatadogAPMTTFT / TPOT
💼 Roles & Salary
AI Platform Engineer、Observability Engineer、DevOps / SRE
💰 $195K - $405K / year (AI Observability & Distributed Tracing) | ¥520K - ¥1.22M / year

Token Cost Attribution, Multi-Tenant Chargeback & FinOps

Real-time token cost calculation across multi-cloud providers (prompt/completion/reasoning token pricing), departmental chargeback models, budget overrun triggers, cache hit analysis, and ROI efficiency reporting.

🏢 Companies
LangfuseLangfusePortkeyPortkeyHeliconeHeliconeCloudZeroCloudZeroDatabricksDatabricks
🛠️ Tech Stack
Token CostFinOpsCost AttributionChargebackBudget AlertsPrompt CachingCloudZeroLangfuse
💼 Roles & Salary
FinOps Engineer、AI Platform Engineer、DevOps / SRE
💰 $185K - $385K / year (Token Cost & AI FinOps) | ¥480K - ¥1.15M / year