MIP-0017: Agentic tooling ideas from ai-job-search and September 2026's trending agent repos¶
| Status | Draft |
| Author | Claude Sonnet 5, for M. Hoffmann (request of 2026-09-06: survey ai-job-search's agentic tooling and this week's trending agent repos for ideas marola could use) |
| Created | 2026-09-06 |
| Phase | 0 — developer tooling; nothing a user of marola sees. No earlier-phase prerequisite |
| Related | MIP-0011 (Claude Code best practices — this MIP extends the same axis, dev-tooling skills/hooks, not marola's product), docs/3-Working-on-the-repo/AGENT-SKILLS.md §3 (skill candidates already on record), docs/3-Working-on-the-repo/DEV-FLOW.md §5 (the reviewer subagent pattern §1 below tightens), .claude/skills/mip-solve-perpetual/SKILL.md (the checkpointing contract §2 below extends) |
| Effort | S — each accepted idea below is a single skill/doc/script change, no new runtime dependency, no new module |
| Gain | infra/dev-loop (cheaper, more auditable agent sessions on this repo) |
| Effort vs Gain | cheap win for §5.1/§5.2 (build next); §5.3 is a nice-to-have housekeeping item with no urgency — do whenever MIP-0011 task 8 is picked up |
| Depends on | none blocking; Phase 0, developer tooling only. §1 touches the same reviewer-subagent text docs/3-Working-on-the-repo/DEV-FLOW.md §5 already describes; §2 is additive to mip-solve-perpetual's just-rewritten usage-guard section (this session's own work, still unmerged) |
| Risk | importing patterns from a single-tenant personal-automation repo (ai-job-search) or from repos optimized for managing fleets of agents (this week's trending crop) without checking they fit marola's one-repo, one-contributor, MIP-gated shape — several ideas from both sources are explicitly rejected below for exactly this reason |
| Cost so far | — (nothing from this MIP has merged yet) |
1. Summary¶
Two external sources were surveyed for agentic-workflow ideas transferable to marola's own
Claude Code setup (not marola-the-product): MadsLorentzen/ai-job-search
(a personal job-application automation framework built on Claude Code, ~13.8k★) and four repos
from GitHub's trending-agentic-tooling crop of the week of 2026-09-04. Four ideas transfer
concretely to marola's existing .claude/skills//docs/3-Working-on-the-repo/DEV-FLOW.md structure at low effort; three
more are named and explicitly rejected as not fitting marola's shape, so this MIP doesn't read as
"adopt everything trending."
2. Motivation¶
This session's own work exposed two concrete gaps ai-job-search's design happens to address:
- Rewriting
mip-solve-perpetual's usage-guard logic (this session, branchperf/mip-perpetual-skill) added a "checkpointing contract" in prose, but there is still no single file a human (or the next run) can read to see which tasks in an overnight run actually landed, in what order, at what cost without reconstructing it from git log andGH_POST_MORTEM.mdby hand.ai-job-search'sjob_search_tracker.csvis exactly this: one flat file, one row per unit of work, machine- and human-readable. docs/3-Working-on-the-repo/DEV-FLOW.md§5's reviewer-subagent invocation doesn't say whether the reviewer re-reads the diff itself or receives it inline.ai-job-search's drafter-reviewer pattern says explicitly: "The reviewer agent receives drafts inline rather than re-reading them, and the verification checklist runs once", a token-efficiency detail worth stating as a rule here too, especially given this session's own throttle rework was prompted by "Sonnet/medium wastes tokens on basic things."
3. User-visible change¶
None: Phase 0, developer tooling only. No CLI or Telegram-facing output changes.
4. Data sources and dependencies reviewed¶
4.1 MadsLorentzen/ai-job-search (fetched 2026-09-06)¶
What it is: a Claude Code-based framework that evaluates job postings, tailors CVs/cover letters,
and preps interviews, run entirely from a personal fork (github.com/MadsLorentzen/ai-job-search,
AGENTS.md and repo root fetched via WebFetch on 2026-09-06). Verified structure:
.claude/skills/job-application-assistant/holds seven numbered instruction files (01-candidate-profile.md…07-interview-prep.md) rather than one monolithicSKILL.md: ordered sub-documents within one skill..claude/commands/holds thirteen single-purpose slash commands (apply.md,scrape.md,rank.md,outcome.md,interview.md,gmail-sync.md, etc.), each a thin workflow trigger.- The
/applycommand runs a drafter-reviewer pattern: a drafter agent writes the CV/cover letter, a fresh reviewer agent researches the company and critiques the draft inline (without re-reading it from disk), the drafter revises, then a verification pass compiles and visually checks the PDF output once. - State lives in plain files, not a database:
job_search_tracker.csv(one row per application: status, outcome, funnel stage),gmail_sync/(processed message IDs + last-sync timestamp),upskill/(one markdown gap-analysis report per run),documents/applications/<company>_<role>/(archived submitted artifacts). AGENTS.mditself is a thin pointer: it states a "unified thin-pointer design" so multiple agent frameworks (Claude Code, others) read one canonical set of files rather than duplicating config; the same shape marola's ownCLAUDE.md→@AGENTS.mdimport already uses. This confirms marola's existing pattern rather than suggesting a new one.
What was not checked: the repo's actual commit history/quality, whether its LaTeX compile-and-inspect verification loop has a documented failure mode, or its Notion-sync implementation; none of those are candidates for marola, which has no CV/PDF or CRM surface.
4.2 GitHub trending agentic-tooling repos, week of 2026-09-04 (WebSearch, 2026-09-06)¶
Source: a devtools digest dated 2026-09-04 plus a general trending-agents search. Repos reviewed, each at the level of "what is this repo's one idea," not a deep audit:
| Repo | What it does | Relevant idea for marola? |
|---|---|---|
anthropics/skills |
Anthropic's own public agent-skills repo | Yes: a reference to audit marola's SKILL.md frontmatter/structure conventions against, §5.3 |
pacifio/atlas |
"source control layer for tracking and querying changes made by multiple coding agents" | Partially: the idea (a queryable ledger of agent-authored changes) is worth having; the tool is not (assumes a fleet-of-agents SaaS shape marola doesn't have); see §9 |
Graphify-Labs/graphify |
Converts a codebase into a queryable knowledge graph for agent use | No: new heavy dependency for a problem Explore/Grep already solve at marola's size; rejected, §9 |
eneskirca/nodeterm |
Terminal manager showing parallel agent sessions as a pan/zoomable node canvas | No: a UI tool for managing many concurrent agent sessions; marola's own convention is one session per feature (AGENTS.md), so the problem it solves doesn't exist here; rejected, §9 |
garrytan/gstack |
23 Claude Code tools spanning CEO/designer/engineering-manager/QA roles | Not adopted directly (role-per-tool is overkill for a one-repo project), but confirms docs/3-Working-on-the-repo/AGENT-SKILLS.md §3's existing direction (named, scoped subagents) is the right shape at marola's scale |
5. Design¶
Three concrete, additive changes, each a single file/skill edit, no new runtime dependency:
5.1 A flat-file run tracker for mip-solve-perpetual¶
Add docs/MIPs/MIP-NNNN.run-log.csv (one per MIP being worked overnight, gitignored like
GH_POST_MORTEM.md; it's an operator artifact, not a deliverable) with columns: timestamp,
task_k, branch, throttle_level, model, effort, outcome, cost_usd, pr_status. mip-solve-perpetual
appends one row per task attempt (including failed attempts, so a human sees the retry history,
not just the final state). This is strictly additive to the checkpointing contract this session
already wrote into the skill (perf/mip-perpetual-skill, unmerged): that contract defines when
the loop may stop; this tracker records what happened, in one grep-able/csv-parseable place,
the same shape as ai-job-search's job_search_tracker.csv. A future just cost-split run could
cross-check its numbers against this file's cost_usd column as a sanity check.
5.2 State the "reviewer receives the diff inline" rule explicitly¶
Add one sentence to docs/3-Working-on-the-repo/DEV-FLOW.md §5's reviewer-subagent bullet: the reviewer subagent's
prompt should paste git diff <BASE_SHA>..<HEAD_SHA> inline rather than instructing the subagent
to re-run that diff itself, and the verification checklist (build/test/quality) is not something
the reviewer re-derives; it's already reported by the author's own PR body. This is a one-line
efficiency rule, not a new mechanism: docs/3-Working-on-the-repo/AGENT-SKILLS.md §2.1's existing requesting-code-review
sequence already does most of this; the addition is naming the "inline, not re-read" detail so it
survives a future rewrite of that section.
5.3 Audit marola's skill frontmatter against anthropics/skills¶
A one-time housekeeping task, not a new skill: read anthropics/skills' published skill
frontmatter conventions (name, description, allowed-tools/disallowed-tools shape) and
diff against marola's six in-repo skills (.claude/skills/{mip,mip-tasks,mip-solve-perpetual,
site-frontend,voice-note-ingest,voice-to-feature}/SKILL.md) for drift: e.g. whether
disable-model-invocation is used consistently everywhere it should be (already flagged as
task 8 of MIP-0011), or whether a description could be tightened to Anthropic's own phrasing
convention for better trigger-matching. Output: a short paragraph in docs/3-Working-on-the-repo/AGENT-SKILLS.md
recording what was checked and what (if anything) changed, not a new file.
6. Scoring / safety impact¶
None: this MIP touches only Claude Code dev-tooling (skills, docs, an operator-facing CSV). No
change to Swimability, ranking, or any user-facing safety text.
7. Verification plan¶
- §5.1:
mip-solve-perpetual's next real overnight run (already scheduled once MIP-0011's remaining tasks resume) produces a run-log CSV with one row per task attempt, reviewable by hand against the actual PRs opened. - §5.2: the next reviewer-subagent invocation under
docs/3-Working-on-the-repo/DEV-FLOW.md§5 is checked to confirm the diff was pasted inline in the dispatch prompt, not re-fetched by the subagent. - §5.3: a diff of marola's six
SKILL.mdfrontmatter blocks againstanthropics/skills' convention, pasted into the implementing PR's description. - No new unit tests: nothing here is Scala/Python runtime code.
just build && just test && just qualitystill run perAGENTS.md's hard rule and are expected to be no-ops content-wise.
8. Risks, limitations, and honest caveats¶
- The
ai-job-searchstructure was read via WebFetch (an AI summarization pass over the page), not by cloning the repo and reading raw file contents; the described directory layout and drafter-reviewer mechanics are reported as WebFetch's model summarized them, not verified byte-for-byte against the source. Treat §4.1 as "confirmed shape, not confirmed exact wording." - The trending-repos list (§4.2) is a snapshot of one digest plus one search, not an exhaustive or reproducible ranking; GitHub's own trending page is unauthenticated and time-windowed differently than the digest queried; a different day would likely surface a different five.
- §5.1's CSV tracker only has value once
mip-solve-perpetualactually runs unattended for a full MIP; it can't be verified end-to-end until then, which is why this MIP doesn't propose building it inside this same change (per themipskill's own rule: design here, build in a separate PR).
9. Alternatives considered¶
- Do nothing (skip this survey's ideas entirely): loses two real efficiency/auditability ideas at zero-cost, and the counterfactual is another agent session re-doing this exact research because the "should marola pace/track its own dev-agent runs better" question resurfaces. Not chosen.
- Adopt
pacifio/atlas(a real dependency) instead of §5.1's plain-CSV tracker; rejected:atlasis built for orchestrating fleets of concurrent coding agents across a team; marola's own convention is one session per feature (AGENTS.md), so the coordination problem it solves (whose agent touched what, concurrently) doesn't exist here. A flat CSV a human can open in a spreadsheet is proportionate to the actual problem. - Adopt
Graphify-Labs/graphifyfor codebase context; rejected: marola is a four-module sbt project small enough thatExplore/Grep/readingAGENTS.md's own doc-index table already gives an agent working context in seconds; a knowledge-graph dependency solves a scale problem marola doesn't have yet. - Adopt
eneskirca/nodeterm-style parallel-session tooling; rejected: it manages many concurrent agent sessions visually; marola's workflow is explicitly one feature, one session (AGENTS.md"Attribution and cost accounting"), so there is nothing for it to visualize here. - Mirror
ai-job-search's numbered sub-file skill layout (01-…md,02-…mdinside one skill directory) for marola's own skills; considered formip's multi-step template, but rejected for now: marola'sSKILL.mdfiles are each already short enough (under ~150 lines) that splitting them would add navigation overhead without a real length problem to solve. Worth revisiting only if a future skill grows past that.
11. Open questions¶
- Should §5.1's run-log CSV live per-MIP (
docs/MIPs/MIP-NNNN.run-log.csv) or as one running file across all overnight runs? Per-MIP mirrors the existingMIP-NNNN.tasks.mdconvention and keeps it easy to gitignore-and-forget once a MIP's stack merges; a single running file would need its own rotation/archival policy. Leaning per-MIP; not decided here. - Is a CSV the right format, or would a JSON-lines file (one append-only write per row, no
quoting/escaping edge cases with commit messages containing commas) be more robust for a script
to append to reliably?
ai-job-searchuses CSV because its rows are simple; marola's rows would need to embed a PR title, which can contain commas. Leaning JSONL for that reason, but not decided; whichever implementation task builds §5.1 should pick and justify it. - §5.3 (skill-frontmatter audit) has no owner or scheduled time yet; it's a nice-to-have housekeeping task, not blocking anything; it may be worth folding into MIP-0011's still-open task 8 (skills hardening) rather than becoming its own task, since both touch the same files.
Appendix¶
Raw findings, for reference:
ai-job-searchAGENTS.md fetch (2026-09-06): described as a "thin-pointer design" workspace file, delegating toCLAUDE.md(candidate profile) and.claude/skills/job-application-assistant/(methodology), plus.agents/skills/for portable Agent-Skills-format portal tools. No explicit workflow-phase, memory, or cost-tracking mechanism documented in the file itself; those live in the skill/command files instead, per the pointer design.- Repo root fetch (2026-09-06) confirmed directory layout:
.agents/skills/(portal CLI tools),cv/,cover_letters/,documents/,templates/,job_scraper/,job_search_tracker.csv. - Trending-repos source:
startupcorners.com/digest/devtools-digest-2026-09-04plus a general WebSearch for "github trending agentic AI tools September 2026 claude code skills subagents" (2026-09-06). Also surfaced but not reviewed in depth (visual-builder platforms, out of scope for a CLI/Telegram project): Langflow, Dify, Flowise; multi-agent orchestration frameworks MetaGPT, LobeHub, CrewAI, AutoGen; none reviewed against marola's Scala/Kyo stack, noted only because they came up in the same search.