Kaizen

Self-improving agents through iterations.

Kaizen is a system designed to help agents improve over time by learning from their trajectories. It uses a combination of an MCP server for tool integration, vector storage for memory, and LLM-based conflict resolution to refine its knowledge base.

Features

MCP Server: Exposes tools to get guidelines and save trajectories.
Conflict Resolution: Intelligently merges new insights with existing guidelines using LLMs.
Trajectory Analysis: Automatically analyzes agent trajectories to generate tips and best practices.
Milvus Integration: Uses Milvus (or Milvus Lite) for efficient vector storage and retrieval.

Architecture

Quick Start

Installation

Prerequisites:

Python 3.12 or higher
uv (recommended) or pip

git clone <repository_url>
cd kaizen
uv sync && source .venv/bin/activate

Configuration

Set your OpenAI API key:

export OPENAI_API_KEY=sk-...

For detailed configuration options (custom LLM providers, backends, etc.), see CONFIGURATION.md.

Running the MCP Server

uv run fastmcp run kaizen/frontend/mcp/mcp_server.py --transport sse --port 8201

Verify it's running:

npx @modelcontextprotocol/inspector@latest http://127.0.0.1:8201/sse --cli --method tools/list

Available tools:

get_guidelines(task: str): Get relevant guidelines for a specific task.
save_trajectory(trajectory_data: str, task_id: str | None): Save a conversation trajectory and generate new tips.
create_entity(content: str, entity_type: str, metadata: str | None, enable_conflict_resolution: bool): Create a single entity in the namespace.
delete_entity(entity_id: str): Delete a specific entity by its ID.

Documentation

CONFIGURATION.md - Detailed configuration options
CLI.md - Command-line interface documentation
CLAUDE_CODE_DEMO.md - Claude Code demo walkthrough

Development

Running Tests

uv run pytest

Phoenix Sync Tests

Tests for the Phoenix trajectory sync functionality are skipped by default since they require familiarity with the Phoenix integration. To include them:

# Run all tests including Phoenix tests
uv run pytest --run-phoenix

# Run only Phoenix tests
uv run pytest -m phoenix

End-to-End (E2E) Low-Code Verification

To run the full end-to-end verification pipeline (Agent -> Trace -> Tip):

KAIZEN_E2E=true uv run pytest tests/e2e/test_e2e_pipeline.py -s

See docs/LOW_CODE_TRACING.md for more details.

Name		Name	Last commit message	Last commit date
Latest commit History 60 Commits
.claude-plugin		.claude-plugin
.github		.github
demo		demo
docs		docs
examples/low_code		examples/low_code
explorations/claudecode		explorations/claudecode
kaizen		kaizen
plugins/kaizen		plugins/kaizen
tests		tests
.env.example		.env.example
.gitignore		.gitignore
.pre-commit-config.yaml		.pre-commit-config.yaml
.python-version		.python-version
.secrets.baseline		.secrets.baseline
AGENTS.md		AGENTS.md
CHANGELOG.md		CHANGELOG.md
CLAUDE_CODE_DEMO.md		CLAUDE_CODE_DEMO.md
CLI.md		CLI.md
CONFIGURATION.md		CONFIGURATION.md
LICENSE		LICENSE
README.md		README.md
README_extract_trajectories.md		README_extract_trajectories.md
README_phoenix_sync.md		README_phoenix_sync.md
SAVE_SKILL_DESIGN.md		SAVE_SKILL_DESIGN.md
extract_trajectories.py		extract_trajectories.py
pyproject.toml		pyproject.toml
uv.lock		uv.lock

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Kaizen

Features

Architecture

Quick Start

Installation

Configuration

Running the MCP Server

Documentation

Development

Running Tests

Phoenix Sync Tests

End-to-End (E2E) Low-Code Verification

About

Uh oh!

Releases 6

Packages

Contributors 8

Languages

License

AgentToolkit/kaizen

Folders and files

Latest commit

History

Repository files navigation

Kaizen

Features

Architecture

Quick Start

Installation

Configuration

Running the MCP Server

Documentation

Development

Running Tests

Phoenix Sync Tests

End-to-End (E2E) Low-Code Verification

About

Resources

License

Uh oh!

Stars

Watchers

Forks

Releases 6

Packages 0

Contributors 8

Languages

Packages