Diagnostic & Evidence Map
Start from three real scenarios and define Forge users, boundaries, failing inputs, and its first acceptance evidence.
4 hoursEighteen modules grow HeatStack Forge from a precise product contract into a recoverable, observable production agent. Software delivery is the required lab; office research and music creation prove the architecture can transfer.
No seven-day mastery. Build one piece of evidence today, then let the same system become stronger tomorrow.
Check only capabilities you can independently implement, explain, and verify. Results stay in this browser and travel with your exported local record.
Architecture answers how the parts work together; learning order answers what to study next. Both views use the same modules, URLs, and local progress.
Establish model inputs, context, and engineering foundations so the rest of the system starts from explicit, verifiable contracts.
Start from three real scenarios and define Forge users, boundaries, failing inputs, and its first acceptance evidence.
4 hoursCompare models and tools with benchmark tasks across quality, cost, latency, and context limits.
6 hoursBuild the Forge CLI, configuration, schemas, HTTP client, async flow, logging, and tests.
10 hoursUnderstand tokens, context, embeddings, and vision input while building a unified input adapter.
10 hoursUse instruction hierarchy, context selection, few-shot examples, and schemas to build a stable requirement extractor.
10 hoursSelect actions from goals and state, execute tools, observe and verify results, then stop or replan.
Implement tool selection, structured arguments, streaming, retries, idempotency, and stop conditions.
12 hoursExtend what an agent can do through capability packages, external evidence, persistent state, and standard protocols.
Package the requirement extractor as a Skill with a manifest, resource boundaries, and compatibility metadata.
9 hoursMove from corpus ingestion, chunking, retrieval, reranking, and citations to an evaluated knowledge system.
14 hoursDesign session state, checkpoints, short- and long-term memory, compression, and replay.
12 hoursImplement tools, resources, state, Streamable HTTP, authentication, and client integration.
16 hoursDecompose complex work into recoverable steps and coordinate execution units through explicit contracts.
Build task graphs, planning, execution, validation, replanning, budgets, and loop protection.
14 hoursUse message contracts, shared state, conflict handling, and cost evaluation to decide when roles should split.
14 hoursUse the runtime environment to control tools, permissions, side effects, budgets, human approval, recovery, and platform differences.
Build tool registration, dry runs, change plans, confirmation, rollback, and audit logs.
12 hoursCover identity, authorization, least privilege, human confirmation, injection attacks, dependency risk, and incident response.
14 hoursBuild capability, directory, and permission adapters for Codex, Claude Code, and WorkBuddy.
12 hoursUse evaluation, observability, deployment, and evidence to prove the system is reliable under real constraints.
Build eval sets, tracing, quality and cost metrics, deployment, rollback, and incident drills.
18 hoursIntegrate all three scenarios and complete the release, demo, architecture docs, and evidence index.
24 hoursUse Forge code, logs, evals, and failure evidence to practice project explanation and system design.
10 hoursTurn selected Skills into precise course stages, lab evidence, and interview claims.
The queue is empty. Choose one capability you genuinely want to practice from a Skill detail page.
Export diagnostic answers, the practice queue, progress, and lab evidence. Import and validation happen only in this browser; the file is never uploaded. Portfolio-ready means locally validated lab evidence is available.