angelX is a terminal coding agent. It works inside your repository with the models you choose, asks before it runs a command or edits a file, checks its work before it reports done, and keeps what it learns about your project.
Field proven in kernel and cryptography work: GPU MODE Cholesky · 2nd · 100+ records on Yukon

# Requires Linux x86_64, Rust/Cargo, C/C++ tools, Bash, Node.js and Python 3.
# The first launch builds from source.
git clone https://github.com/newjordan/angelX.git
cd angelX
./bin/angelX
Model setup · /commands · Feature evidence · Attributions · MIT

1 · Choose a model. /model lists every connected route with its thinking level; /think changes the level.

2 · Approve what it does. Commands and edits arrive as action capsules: approve one (y), approve the rest of the turn (a), or deny (n).

3 · Content-checked edits. Each edit is anchored to the file's current content and shows its byte change before it lands. /diff shows the result.
4 · Checked before done. The agent runs the checks and reports with receipts: every tool call, its time and the model route.

5 · Longer work. /goal sets a durable objective with acceptance criteria and a check. /loop runs it within time, iteration and token limits.
![]() Formations · model teams, from a lone coordinator to a full roster. | ![]() Context budget · where the window goes; compaction runs automatically. |
![]() Commands · /help and Tab completion for every command. | ![]() World view · Watch model behavior as an adventure in information. |
code_mode.136 repository-repair tasks (48 JS, 34 Python, 30 Rust, 24 C++) from the Aider polyglot set. One attempt per task, 600 s limit, graded by each task's tests. Wall: median agent time per attempt. Tokens: totals per cell.
| model | harness | solved | wall (median) | calls / task | input tokens | cache hit | output tokens |
|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash, thinking off | angelX | 133 / 136 | 9.9 s | 7.5 | 11.5 M | 88% | 257 k |
| OpenCode 1.18.31 | 132 / 136 | 8.6 s | 12.5 | 67.7 M | 97% | 349 k | |
| oh-my-pi 18.2.4 | 86 / 93 * | 15.3 s | 35.5 | 200.1 M | 99% | 866 k | |
| GLM-5.3-Flash, thinking low | angelX | 135 / 136 | 50.7 s | 8.9 | 10.9 M | 84% | 256 k |
| OpenCode 1.18.31 | 133 / 136 | 38.9 s | 7.4 | 10.1 M | 85% | 179 k | |
| oh-my-pi 18.2.4 | 133 / 136 | 37.7 s | 8.9 | 24.1 M | 90% | 203 k |
angelX runs a verification step before reporting a task done. That step accounts for its extra wall time on GLM; higher thinking levels add more.
Evaluator: Prime Intellect Verifiers v0.3.1 · angelX 98d7340 · OpenCode 1.18.31 · oh-my-pi 18.2.4 · DeepSeek V4.1 Flash, thinking off · GLM-5.3-Flash, thinking low (lowest available) · temperature 0 · 8,192-token output cap · 600 s per attempt · fresh environment per attempt · 2026-09-21
Research and public work that informed Angel:
Code and tooling credits include OpenAI Codex, Grok CLI, oh-my-pi, DeepSeek-Reasonix, Hermes Agent, Prime Agent, Dotmax, ureq, Ratatui, Crossterm, rusty_v8 and V8. Early inspiration: DotAgents and SpeakMCP by aj47. File-tool interface references include Claude Code, aider and OpenHands. Attributions records implementation links, authors and retained licenses; research inspiration and incorporated code are identified separately.
Field proven: 100+ records on public research leaderboards, with first-place results in kernel optimization, LLM inference and cryptography research, and 2nd place on the GPU MODE Cholesky leaderboard. Yukon field results
Yukon, the platform for open frontier research, where angelX set its leaderboard records.
115 commits
Rust
93.9%
JavaScript
3.1%
Python
2.8%
angelX is a terminal coding agent. It works inside your repository with the models you choose, asks before it runs a command or edits a file, checks its work before it reports done, and keeps what it learns about your project.
Field proven in kernel and cryptography work: GPU MODE Cholesky · 2nd · 100+ records on Yukon

# Requires Linux x86_64, Rust/Cargo, C/C++ tools, Bash, Node.js and Python 3.
# The first launch builds from source.
git clone https://github.com/newjordan/angelX.git
cd angelX
./bin/angelX
Model setup · /commands · Feature evidence · Attributions · MIT

1 · Choose a model. /model lists every connected route with its thinking level; /think changes the level.

2 · Approve what it does. Commands and edits arrive as action capsules: approve one (y), approve the rest of the turn (a), or deny (n).

3 · Content-checked edits. Each edit is anchored to the file's current content and shows its byte change before it lands. /diff shows the result.
4 · Checked before done. The agent runs the checks and reports with receipts: every tool call, its time and the model route.

5 · Longer work. /goal sets a durable objective with acceptance criteria and a check. /loop runs it within time, iteration and token limits.
![]() Formations · model teams, from a lone coordinator to a full roster. | ![]() Context budget · where the window goes; compaction runs automatically. |
![]() Commands · /help and Tab completion for every command. | ![]() World view · Watch model behavior as an adventure in information. |
code_mode.136 repository-repair tasks (48 JS, 34 Python, 30 Rust, 24 C++) from the Aider polyglot set. One attempt per task, 600 s limit, graded by each task's tests. Wall: median agent time per attempt. Tokens: totals per cell.
| model | harness | solved | wall (median) | calls / task | input tokens | cache hit | output tokens |
|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flash, thinking off | angelX | 133 / 136 | 9.9 s | 7.5 | 11.5 M | 88% | 257 k |
| OpenCode 1.18.31 | 132 / 136 | 8.6 s | 12.5 | 67.7 M | 97% | 349 k | |
| oh-my-pi 18.2.4 | 86 / 93 * | 15.3 s | 35.5 | 200.1 M | 99% | 866 k | |
| GLM-5.3-Flash, thinking low | angelX | 135 / 136 | 50.7 s | 8.9 | 10.9 M | 84% | 256 k |
| OpenCode 1.18.31 | 133 / 136 | 38.9 s | 7.4 | 10.1 M | 85% | 179 k | |
| oh-my-pi 18.2.4 | 133 / 136 | 37.7 s | 8.9 | 24.1 M | 90% | 203 k |
angelX runs a verification step before reporting a task done. That step accounts for its extra wall time on GLM; higher thinking levels add more.
Evaluator: Prime Intellect Verifiers v0.3.1 · angelX 98d7340 · OpenCode 1.18.31 · oh-my-pi 18.2.4 · DeepSeek V4.1 Flash, thinking off · GLM-5.3-Flash, thinking low (lowest available) · temperature 0 · 8,192-token output cap · 600 s per attempt · fresh environment per attempt · 2026-09-21
Research and public work that informed Angel:
Code and tooling credits include OpenAI Codex, Grok CLI, oh-my-pi, DeepSeek-Reasonix, Hermes Agent, Prime Agent, Dotmax, ureq, Ratatui, Crossterm, rusty_v8 and V8. Early inspiration: DotAgents and SpeakMCP by aj47. File-tool interface references include Claude Code, aider and OpenHands. Attributions records implementation links, authors and retained licenses; research inspiration and incorporated code are identified separately.
Field proven: 100+ records on public research leaderboards, with first-place results in kernel optimization, LLM inference and cryptography research, and 2nd place on the GPU MODE Cholesky leaderboard. Yukon field results
Yukon, the platform for open frontier research, where angelX set its leaderboard records.
115 commits
Rust
93.9%
JavaScript
3.1%
Python
2.8%