Metadata-Version: 2.4
Name: haliosai-cli
Version: 2.0.2
Summary: Halios CLI for coding-agent evaluations, guardrails, and OpenTelemetry evidence
Author-email: HaliosLabs <support@halios.ai>
License-Expression: Apache-2.0
Project-URL: Homepage, https://halios.ai
Project-URL: Documentation, https://docs.halios.ai
Project-URL: Repository, https://github.com/HaliosAI/halios
Project-URL: Changelog, https://github.com/HaliosAI/halios/releases
Keywords: ai,agents,evaluation,guardrails,opentelemetry,cli
Classifier: Development Status :: 4 - Beta
Classifier: Environment :: Console
Classifier: Intended Audience :: Developers
Classifier: Programming Language :: Python :: 3 :: Only
Classifier: Programming Language :: Python :: 3.10
Classifier: Programming Language :: Python :: 3.11
Classifier: Programming Language :: Python :: 3.12
Classifier: Programming Language :: Python :: 3.13
Classifier: Programming Language :: Python :: 3.14
Classifier: Topic :: Software Development :: Testing
Requires-Python: >=3.10
Description-Content-Type: text/markdown
License-File: LICENSE
Requires-Dist: httpx<1,>=0.27
Requires-Dist: jsonschema<5,>=4.18
Requires-Dist: PyYAML<7,>=6.0
Requires-Dist: rich<15,>=13
Requires-Dist: typer<1,>=0.12
Requires-Dist: tomli>=2.0; python_version < "3.11"
Provides-Extra: dev
Requires-Dist: build>=1.2; extra == "dev"
Requires-Dist: pytest>=8; extra == "dev"
Requires-Dist: ruff>=0.9; extra == "dev"
Requires-Dist: twine>=6; extra == "dev"
Dynamic: license-file

# Halios

[![PyPI version](https://img.shields.io/pypi/v/haliosai-cli.svg)](https://pypi.org/project/haliosai-cli/)
[![License](https://img.shields.io/badge/license-Apache--2.0-blue.svg)](LICENSE)

**Your pair programmer for AI agent evaluations.**

Halios helps you build reliable AI agents in minutes. Add the Halios skill to **Codex**, **Claude Code**, **Cursor**, or your favorite coding agent, and create eval suites, run multi-turn simulations, investigate failures, and gate pull requests directly from your repository. 

Halios provides the evaluation framework, developer tools, and hosted evaluation runtime underneath.

---

## Quickstart

### 1. Install the Agent Skill

Add the Halios skill to your coding agent environment:

```bash
npx skills add HaliosAI/halios --skill halios
```

*(Works with Codex, Claude Code, Cursor, GitHub Copilot, Gemini CLI, OpenCode, and any harness supporting the open Agent Skills format).*

### 2. Ask Your Coding Agent

Prompt your agent to set up evaluations for your project:

```text
"Set up evals for this agent: inspect the repository, create realistic test scenarios and checks, and run a baseline."
```

Your agent will inspect your application entrypoint, configure standard OpenTelemetry, draft scenarios in `.halios/`, and run baseline evaluations.

---

### Standalone CLI Installation

If you prefer to drive evaluations directly from the command line or CI:

```bash
# Recommended: Install with uv tool
uv tool install 'haliosai-cli>=2.0.0'

# Or install with pipx
pipx install 'haliosai-cli>=2.0.0'

# See available commands and usage
halios --help
```

---

## How It Works

1. **Inspect & Connect**: Your coding agent inspects your agent's tools, policies, and runtime, then configures standard OpenTelemetry export.
2. **Author Scenarios & Checks**: Test cases, rubrics, and failure criteria are stored directly in your repository (`.halios/scenarios.yml` and `.halios/eval.yml`). You own the evaluation suite.
3. **Simulate Multi-Turn Trajectories**: Halios runs fresh simulation passes against your agent across edge cases, tool dependencies, and user personas.
4. **Investigate & Fix Failures**: Pinpoint hallucinated parameters, broken tool handoffs, or policy violations from complete trace evidence.
5. **Gate Pull Requests**: Run protected checks in CI to block regressions before merging to production.
6. **Monitor Production**: Evaluate production traces using the same checks, and turn real-world failures into new regression test scenarios.

---

## Key Principles

- **You own the evaluation suite**: Scenarios and checks are clean YAML files stored in your Git repository.
- **Fresh simulations, not static replays**: Halios tests real agent execution across multiple turns, rather than replaying outdated completions.
- **Framework agnostic**: Works with any agent architecture—OpenAI Agents SDK, LangChain, LlamaIndex, PydanticAI, or custom workflows.
- **Stock OpenTelemetry**: Applications emit standard OpenTelemetry GenAI spans; no proprietary vendor lock-in in your production runtime.
- **Unified local, CI, and production loop**: The same checks run during local development, gate merge requests in CI, and monitor live production traces.

---

## Resources

- **Website**: [halios.ai](https://halios.ai)
- **Documentation**: [docs.halios.ai](https://docs.halios.ai)
- **Python SDK**: [github.com/HaliosAI/haliosai-python-sdk](https://github.com/HaliosAI/haliosai-python-sdk)
- **Public Skill Source**: [`skills/halios/`](skills/halios/)

---

## Development

```bash
# Clone and install locally in editable mode
git clone https://github.com/HaliosAI/halios.git
cd halios
python -m pip install -e '.[dev]'

# Run the test suite
python -m pytest -q
```

---

## License

[Apache 2.0](LICENSE) © Anomalytica Inc. 2026
