hermes-agent/tests/hermes_cli/test_reasoning_full_command.py
Teknium 6b81590c55
test: prune low-value tests suite-wide (wave 1) — 46,820 → 28,106 test functions
Systematic prune per AGENTS.md test policy, one pass over every major
test tree (gateway, hermes_cli, tools, agent, run_agent, plugins, cli,
cron, tui_gateway, honcho/openviking, root-level):

- DELETE: source-reading tests (read_text/getsource on prod files),
  change-detector tests (exact catalog counts, model-name snapshots,
  config version literals), mock-echo tests (assert a mock returns what
  it was told), assertion-free/trivial tests, near-duplicate
  parametrizations (boundaries + one representative kept), async/sync
  twin duplicates, cosmetic within-file variations.
- KEEP (mandatory): security/redaction/approval guards, message-role
  alternation invariants, prompt-caching/deterministic-call-id
  invariants, issue-number regression tests (deduped), E2E tests.
- 6 test files deleted outright (script-style/no-assert or fully
  redundant); conftest.py, fakes/, fixtures/ untouched.
- tests/acp/conftest.py added: autouse fixture stubs the live
  models.dev/GitHub/Copilot/Anthropic inventory fetches that ACP server
  tests performed on every session create — test_server.py 147s → 3.4s,
  and the tests are now genuinely hermetic.
- Sleep-based slowness shrunk where safe (codex_ttfb_watchdog,
  compression_concurrent_fork, etc.); no wall-clock assertion tightened.

Verification: full hermetic suite via scripts/run_tests.sh —
2439 files, 31,130 tests passed, 0 failed, 0 flaky retries, 315s wall
(baseline: 583s wall, 13,564s subprocess CPU).
2026-07-29 13:10:23 -07:00

63 lines
2.1 KiB
Python

"""Tests for the CLI `/reasoning full` / `/reasoning clamp` recap toggle.
The post-response "Reasoning" recap box clamps long thinking to the first
10 lines. `/reasoning full` opts into uncapped display (Taelin's "show all
thinking tokens" ask); `/reasoning clamp` restores the 10-line collapse.
These assert the toggle sets the instance flag, persists to config.yaml,
and that the clamp gate honours the flag.
"""
import os
import yaml
from hermes_cli.cli_commands_mixin import CLICommandsMixin
from hermes_cli.config import DEFAULT_CONFIG
class _Stub(CLICommandsMixin):
"""Minimal carrier for the attributes `_handle_reasoning_command` reads."""
def __init__(self):
self.reasoning_config = None
self.show_reasoning = True
self.reasoning_full = False
self.agent = None
def _current_reasoning_callback(self):
return None
def test_default_config_clamps_reasoning():
# Behaviour contract: the recap defaults to clamped, not full.
assert DEFAULT_CONFIG["display"]["reasoning_full"] is False
def _seed_config(tmp_path, monkeypatch):
hh = tmp_path / ".hermes"
hh.mkdir()
(hh / "config.yaml").write_text("display:\n show_reasoning: true\n")
monkeypatch.setenv("HERMES_HOME", str(hh))
# cli captures _hermes_home at import; force it to the temp home.
import cli
monkeypatch.setattr(cli, "_hermes_home", hh, raising=False)
return hh
def test_reasoning_full_sets_and_persists(tmp_path, monkeypatch):
hh = _seed_config(tmp_path, monkeypatch)
s = _Stub()
s._handle_reasoning_command("/reasoning full")
assert s.reasoning_full is True
saved = yaml.safe_load((hh / "config.yaml").read_text())
assert saved["display"]["reasoning_full"] is True
def test_clamp_gate_honours_flag():
# The display gate at cli.py: clamp only when long AND not reasoning_full.
reasoning = "\n".join(f"line{i}" for i in range(25))
lines = reasoning.strip().splitlines()
assert (len(lines) > 10 and not False) is True # full=False -> clamp
assert (len(lines) > 10 and not True) is False # full=True -> show all