Repository navigation
docs(plans): phased plan for RLM language-based workflows (tinyagents rhai REPL) - #4510
Conversation
|
Warning Review limit reachedYou’ve reached a temporary PR review limit under our Fair Usage Limits Policy. Next review available in: 58 minutes Enable usage-based reviews in Billing to review now. Otherwise, wait until the next included review is available. How can I continue?After more reviews become available, a review can be triggered using the To avoid repeated limits, reduce automatic review volume by pausing incremental auto-reviews earlier, using label-based review opt-in, excluding WIP or generated PR titles, or requesting reviews manually when the PR is ready. If your team needs uninterrupted high-volume reviews, an organization admin can enable usage-based reviews. How do review limits work?CodeRabbit enforces per-developer PR review limits for each organization. Most developers receive the normal plan review availability. For paid Pro and Pro+ PR reviews, CodeRabbit uses adaptive limits for sustained high-volume activity. When a developer's recent PR review activity reaches the 95th percentile or higher among CodeRabbit users, additional reviews become available more gradually as earlier reviews age out of the rolling window. Please refer docs for additional details. Review details⚙️ Run configurationConfiguration used: Organization UI Review profile: CHILL Plan: Pro Run ID: 📒 Files selected for processing (8)
Comment |
There was a problem hiding this comment.
💡 Codex Review
Here are some automated review suggestions for this pull request.
Reviewed commit: decbbc6fdb
ℹ️ About Codex in GitHub
Your team has set up Codex to review pull requests in this repo. Reviews are triggered when you
- Open a pull request for review
- Mark a draft as ready
- Comment "@codex review".
If Codex has suggestions, it will comment; otherwise it will react with 👍.
Codex can also answer questions or update the PR. Try commenting "@codex address that feedback".
| `ToolAdapter` carries each tool's own security/approval behavior, scripts | ||
| get exactly the same gates as direct tool calls. |
There was a problem hiding this comment.
Put the approval gate in the RLM bridge
When a script calls an external-effect tool via the planned REPL CapabilityRegistry, it will not automatically pass through the harness tool middleware: ApprovalSecurityMiddleware is installed on harness runs (src/openhuman/tinyagents/mod.rs:1713-1720), while execute_openhuman_tool documents that approval was removed from the adapter and now happens before the executor (src/openhuman/tinyagents/tools.rs:213-215). The referenced ToolAdapter is also #[cfg(test)], so relying on it for production approval/security would let tool_call execute effectful inner tools without HITL unless the RLM bridge explicitly invokes the approval/permission gate before direct tool execution.
Useful? React with 👍 / 👎.
| `registry.replace_tool(name, adapter)`. **Exclusions** (recursion + | ||
| duplication guards): `rlm` itself, `spawn_subagent`/`spawn_parallel_agents` | ||
| (use `agent_query` instead), `run_workflow`/`await_workflow`. Because |
There was a problem hiding this comment.
Strip all delegation tools from RLM tool_call
This exclusion list only removes spawn_subagent/spawn_parallel_agents, but the default registry also includes other delegation surfaces such as spawn_async_subagent, agent_prepare_context, delegate_graph, and delegate_* tools (src/openhuman/tools/ops.rs:168-190). In scripts that use tool_call, those would be counted as tool calls rather than ReplPolicy.max_agent_calls/depth-limited agent_query calls, bypassing the intended RLM agent limits; mirror the broader spawn/delegate stripping used by the tinyagents registration guard (src/openhuman/tinyagents/mod.rs:421-427) or route all such tools through the agent capability path.
Useful? React with 👍 / 👎.
Summary
Plan-only PR — no code changes. Adds
docs/plans/rlm-workflows/, the phased implementation plan for exposing TinyAgents' Rhai-backed REPL language (the.ragsh/ RLM / CodeAct surface, gated behind thereplcargo feature invendor/tinyagents) as a first-classrlmtool in the OpenHuman core, so the orchestrator can author and execute its own workflows (fan-out over subagents, batched tool/model calls, loops) — similar to Claude Code Workflows and Recursive Language Models.Plan contents
README.mdphase-1-research.mdreplsurface (ReplSession, built-ins, fail-closed ReplPolicy, blocking async bridge) + openhuman integration points (Tool trait, ToolAdapter/SharedToolAdapter, ProviderModel, subagent runner, approval gate) + gapsphase-2-tinyagents.mdtinyhumansai/tinyagentsphase-3-rlm-domain.mdsrc/openhuman/rlm/domain: policy mapping, capability bridge, session manager, eval opsphase-4-rlm-tool.mdrlmtool: schema, registration, permission/approval posture, prompt surfacingphase-5-hardening.mdphase-6-tests.mdphase-7-delivery.mdKey design decisions captured
rlmmaps one tool call → oneeval_cell, with a persistentsession_idfor namespace continuity.ApprovalGateitself for external-effect tools so scripts can't bypass HITL review.spawn_blockingtimeout + harnessToolTimeoutbackstop; session LRU/TTL; call-count/output/depth limits fromReplPolicy.Notes
--no-verify: the pre-push hook fails on pre-existing Rust warnings unrelated to this docs-only change.https://claude.ai/code/session_014BU5kzHXUDn8yP1fWCq2QN