Enterprise Agent Operating System

17-plane kernel, 16 platform adapters, 1,584 tests, 48/48 protocol conformance. Build, deploy, and govern autonomous AI agents at production scale.

πŸ† GOLD Certified (83.4/100) βœ“ 48/48 Conformance βœ“ 1,584 Tests MIT Licensed

Platform Overview

17Control Planes

Identity β†’ Safety & Alignment bounded planes

16Adapters

OpenClaw, CrewAI, LangGraph, Claude Code, MCP, A2A + more

10Providers

Multi-model routing with cost/latency optimization

48/48Conformance

Full protocol compliance across all adapters

1,584Tests

Enterprise-grade test coverage across all modules

100%Security

OWASP-compliant, zero critical vulnerabilities

Enterprise Features

πŸ›οΈ

Executive Dashboard

Real-time agent orchestration with role-based views. Monitor fleet health, budgets, and compliance from a single pane.

πŸ€–

Agent Registry

Discover, install, and manage certified agents. 8 marketplace agents with digital signing and lifecycle management.

🧩

Skill Registry

Versioned, composable agent skills. Swappable behavioral units with dependency resolution and provenance tracking.

⚑

Live Agent Execution

Execute agents in real-time with streaming output. HITL approval gates, retry with backoff, and Saga compensation.

🧠

Memory Visualization

4 memory types visualized: semantic, procedural, episodic, organizational. Knowledge graph with 142 entities and 163 edges.

πŸ”€

Workflow Builder

DAG-based workflow composition. Visual task chaining with checkpointing, rollback, and conditional branching.

πŸ“Š

Event Timeline

Immutable event sourcing with full replay capability. Every execution produces a verifiable receipt.

πŸ“ˆ

Evaluation Dashboard

14-dimension agent scoring with historical trends. Regression detection and continuous improvement tracking.

🧭

LLM Router

Intelligent model selection across 10 providers. Cost/latency/quality optimization with fallback chains.

πŸ”§

MCP Tool Registry

Model Context Protocol integration. Auto-discover and register external tools with capability-based permissions.

πŸ“‘

Runtime Metrics

OpenTelemetry traces, cost tracking, SLO dashboards. Full observability across the agent mesh.

🎭

Enterprise Demo Mode

Pre-built demo scenarios: multi-agent workflow, executive planning, memory replay, event sourcing, skill execution.

Architecture

Dual-Runtime Engine

Python Enterprise Kernel β€” 17 bounded control planes, 16 adapters, governance, orchestration, certification, evaluation, digital twin, economics.

TypeScript Monorepo β€” Express API (33 endpoints), React dashboard, CLI, MCP server, 82 source modules across 28 domains.

Shared Layer β€” Supabase (PostgreSQL), event bus, memory persistence, skill registry, cron scheduler.

Governance Pipeline

GateDescription
1. IntentParse and validate agent objective
2. RiskClassify risk level (LOW→CRITICAL)
3. PolicyMatch against allowed capabilities
4. ApprovalRoute to governance board if required
5. ExecutionDispatch through capability router
6. ReceiptImmutable execution record

Model Suite

AIF-Kernel

Core enterprise agent runtime. 17 control planes, event sourcing, CQRS command bus. The central nervous system of AIF.

View Model β†’

AIF-Orchestrator

Multi-agent coordination engine. DAG workflow execution, retry/rollback, HITL gates, and Saga compensation patterns.

View Model β†’

AIF-Planner

Task decomposition and dependency resolution. Converts objectives into executable DAGs with resource estimation.

View Model β†’

AIF-Execution

Capability router and execution engine. Dispatches tasks to the optimal adapter with cost/latency awareness.

View Model β†’

AIF-Router

Intelligent model selection across 10+ providers. Quality-aware routing with automatic fallback chains.

View Model β†’

Dataset Ecosystem

aif-system-prompts

Versioned system prompts, policy definitions, and evaluation criteria for enterprise agents.

View β†’

aif-agent-specifications

Structured agent definitions, capability manifests, and skill contracts.

View β†’

aif-benchmarks

Evaluation datasets for agent performance: GAIA, SWE-bench, AgentBench results.

View β†’

aif-memory-schemas

Event sourcing schemas, memory type definitions, and persistence models.

View β†’

aif-agent-skills

Reusable agent skill definitions with versioning, dependencies, and provenance.

View β†’

aif-execution-events

Immutable execution event logs for audit replay and compliance verification.

View β†’

Performance & Benchmarks

0.840GAIA Score
0.782SWE-bench
0.820AgentBench
83.4GOLD Score

Get Started

Clone the repo, install dependencies, and run your first agent in under 5 minutes.

GitHub Repository Hugging Face Hub
v3-1784589218