最近更新
更新:2026-08-03 09:15:53 · 公开链接
External Agent Skills Design Patternsai/concepts
wiki/ai/concepts/External-Agent-Skills-Design-Patterns.md
Agent Graph — fact-routed work contracts for Agent Skillsai/sources
wiki/ai/sources/agent-graph-fact-routed-skill-workflows.md
AxisAgentic — runtime and trajectory collection framework for long-horizon agentsai/sources
wiki/ai/sources/axisagentic-runtime-trajectory-framework.md
OptMem — permanent append-only memory for AI agentsai/sources
wiki/ai/sources/optmem-permanent-agent-memory.md
SimpleEnglish — controlled-language agent skill for unambiguous documentationai/sources
wiki/ai/sources/simpleenglish-controlled-language-agent-skill.md
全部页面
按知识分区浏览
ai 263
Loop Engineering vs Harness Engineeringai/comparisons
wiki/ai/comparisons/Loop-Engineering-vs-Harness-Engineering.md
External Agent Skills Design Patternsai/concepts
wiki/ai/concepts/External-Agent-Skills-Design-Patterns.md
LLM Training From Scratch Education Stackai/concepts
wiki/ai/concepts/LLM-Training-From-Scratch-Education-Stack.md
Mediated Communication Can Steer Collective Opinion Generativeai/entities
wiki/ai/entities/Mediated_Communication_Can_Steer_Collective_Opinion_Generative.md
A Recipe for Training Neural Networksai/sources
wiki/ai/sources/A_Recipe_for_Training_Neural_Networks.md
ACQUIRE — QA-Driven Repository Knowledge Acquisitionai/sources
wiki/ai/sources/acquire-qa-driven-repository-knowledge.md
AI Dev Team — supervised multi-agent delivery harnessai/sources
wiki/ai/sources/ai-dev-team-supervised-delivery-harness.md
AI 生成代码 Review 调研 2026-07-01ai/sources
wiki/ai/sources/ai-code-review-agent-coding-research-2026-07-01.md
AI-Mediated Communication Can Steer Collective Opiai/sources
wiki/ai/sources/AI-Mediated_Communication_Can_Steer_Collective_Opi.md
AKM Eval — agentic knowledge management maturity indexai/sources
wiki/ai/sources/akm-eval-agentic-knowledge-management.md
Action-Graded Severity Scale for Tool-Using AI Agentsai/sources
wiki/ai/sources/action-graded-severity-scale-tool-agents.md
Agent Graph — fact-routed work contracts for Agent Skillsai/sources
wiki/ai/sources/agent-graph-fact-routed-skill-workflows.md
Agent Skills Hub and skills.sh High-Star Skillsai/sources
wiki/ai/sources/Agent-Skills-Hub-and-Skills-sh-High-Star-Skills.md
AgentBattler Bench — sealed harness benchmark for coding agentsai/sources
wiki/ai/sources/agentbattler-bench-sealed-harness-benchmark.md
AgentDoctor — coding-agent configuration auditai/sources
wiki/ai/sources/agentdoctor-coding-agent-config-audit.md
AgentKernelArena:GPU Kernel Agent 的 A/B 与 RL-ready 评测环境ai/sources
wiki/ai/sources/agentkernelarena-ab-rl-gpu-kernel-agents.md
AgentOps coding-agent verification membraneai/sources
wiki/ai/sources/agentops-coding-agent-verification.md
AgentOps — bounded coding-agent operating loop and skills bundleai/sources
wiki/ai/sources/agentops-bounded-operating-loop.md
AgentTether — graph-guided runtime repair for LLM agentsai/sources
wiki/ai/sources/agenttether-graph-guided-runtime-repair.md
Agentic coding and persistent returns to expertiseai/sources
wiki/ai/sources/claude-code-expertise.md
Approving — human-gated multi-agent workflowsai/sources
wiki/ai/sources/approving-human-gated-multi-agent-workflows.md
AutoDev Studio — repository knowledge base as multi-agent SDLC harnessai/sources
wiki/ai/sources/autodev-studio-repo-kb-sdlc-harness.md
AutoResearch 自动研究框架,通过 metric 搜索和 eval 三个关键动作,把赌变成ai/sources
wiki/ai/sources/AutoResearch_自动研究框架通过_metric_搜索和_eval_三个关键动作把赌变成.md
Autonomous Coding Agents 的 Security Debtai/sources
wiki/ai/sources/security-debt-autonomous-coding-agents.md
AxisAgentic — runtime and trajectory collection framework for long-horizon agentsai/sources
wiki/ai/sources/axisagentic-runtime-trajectory-framework.md
BOSS Console — agent operator console and governed workspaceai/sources
wiki/ai/sources/bossconsole-agent-operator-console.md
Better Harness — evidence-bounded coding workflow reviewai/sources
wiki/ai/sources/better-harness-evidence-bounded-workflow-review.md
Caliper — Reliability testing for agent skillsai/sources
wiki/ai/sources/caliper-skill-reliability-testing.md
Chancery — agent identity and writ-based controlai/sources
wiki/ai/sources/chancery-agent-identity-writ-control.md
CodeRail — convergent coding governance for agentic projectsai/sources
wiki/ai/sources/coderail-convergent-coding.md
ContextIQ — AST code graph for token-efficient agent contextai/sources
wiki/ai/sources/contextiq-ast-code-graph-agent-context.md
ContextNest / ContextNext — Verifiable Context Governanceai/sources
wiki/ai/sources/contextnest-verifiable-context-governance.md
CyVisGuard — MCP security control plane for AI agentsai/sources
wiki/ai/sources/cyvisguard-mcp-security-control-plane.md
Deep Neural Nets 33 Years Ago and 33 Years From Nowai/sources
wiki/ai/sources/Deep_Neural_Nets_33_Years_Ago_and_33_Years_From_Now.md
DeepSWE — Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasksai/sources
wiki/ai/sources/deepswe-long-horizon-coding-agent-benchmark.md
Deterministic Gates for Tool-Using Agent Policy Enforcementai/sources
wiki/ai/sources/deterministic-gates-tool-agent-policy.md
DexJoCo: A Benchmark and Toolkit for Task-Orientedai/sources
wiki/ai/sources/DexJoCo_A_Benchmark_and_Toolkit_for_Task-Oriented.md
Distributed Attacks in Persistent-State AI Controlai/sources
wiki/ai/sources/distributed-attacks-persistent-state-ai-control.md
Do Agent Optimizers Compound? — Terminal-Bench 2.0 持续学习评测ai/sources
wiki/ai/sources/agent-optimizers-compound-terminal-bench.md
Enola — deterministic codebase architecture graphai/sources
wiki/ai/sources/enola-deterministic-codebase-architecture-graph.md
EvoSOP — Iterative Tool Optimization for Self-Evolving LLM Agentsai/sources
wiki/ai/sources/evosop-iterative-tool-optimization.md
FastContext — read-only delegated repository exploration agentai/sources
wiki/ai/sources/fastcontext-read-only-repository-exploration.md
Flow-Next — repo-native agentic engineering workflowai/sources
wiki/ai/sources/flow-next-agentic-engineering-workflow.md
ForgeOS — skill intelligence and trust control planeai/sources
wiki/ai/sources/forgeos-skill-intelligence-control-plane.md
From Prompts to Contracts — Auditable Enterprise Harness Engineeringai/sources
wiki/ai/sources/from-prompts-to-contracts-harness-engineering.md
Graph Engineering — knowledge graph and task graph topology for agentsai/sources
wiki/ai/sources/graph-engineering-agent-topology.md
Greplica:持久仓库记忆的 coding-agent 规划评测ai/sources
wiki/ai/sources/greplica-persistent-coding-agent-memory.md
Halo Record — tamper-evident runtime records for AI agentsai/sources
wiki/ai/sources/halo-record-runtime-records.md
Harness Engineering Anthology — repository as agent context bundleai/sources
wiki/ai/sources/harness-engineering-anthology.md
Harness Handbook:让演化中的 Agent Harness 可读、可导航、可编辑ai/sources
wiki/ai/sources/harness-handbook-evolving-agent-harnesses.md
Harness Score — deterministic maturity scanner for coding-agent harnessesai/sources
wiki/ai/sources/harness-score-maturity-scanner.md
Hermes Field Kit — field-tested skill admission and validation contractai/sources
wiki/ai/sources/hermes-field-kit-skill-admission-contract.md
Karpathy 是 AI 领域知名研究者,前 Tesla AI 总监,前 OpenAI 联合创始人ai/sources
wiki/ai/sources/Karpathy_是_AI_领域知名研究者前_Tesla_AI_总监前_OpenAI_联合创始人.md
Kitbash — cross-host agent skill format and compilerai/sources
wiki/ai/sources/kitbash-cross-host-agent-skill-format.md
LLM Wiki 是 Karpathy 提出的概念,给个人 Wiki 加 LLM 加持。与 Wikiai/sources
wiki/ai/sources/LLM_Wiki_是_Karpathy_提出的概念给个人_Wiki_加_LLM_加持与_Wiki.md
Laborant Skills — evaluation design for AI systemsai/sources
wiki/ai/sources/laborant-evaluation-design-skills.md
Lilian Weng — Harness Engineering for Self-Improvementai/sources
wiki/ai/sources/lilian-weng-harness-engineering-self-improvement.md
Lilian Weng — Scaling Laws, Carefullyai/sources
wiki/ai/sources/lilian-weng-scaling-laws-carefully.md
Loop Engineering 实战:从日志扫描到预发部署的全自主闭环ai/sources
wiki/ai/sources/loop-engineering-autonomous-log-to-staging.md
Lovable — Scaling Agentic Coding with Token Spendai/sources
wiki/ai/sources/lovable-scaling-agentic-coding.md
Microsoft Agent Skills — context-driven development skill catalogai/sources
wiki/ai/sources/microsoft-agent-skills-context-driven-development.md
NVIDIA NeMo Relay — agent runtime control and trajectory layerai/sources
wiki/ai/sources/nemo-relay-agent-runtime-control.md
Nothing New Under The Sun — research-first scouting skillai/sources
wiki/ai/sources/nothing-new-under-the-sun-research-first-scouting.md
OKF Gem — local Open Knowledge Format toolkit for agentsai/sources
wiki/ai/sources/okf-gem-local-knowledge-bundles.md
OKFy — purpose-shaped knowledge bundles for agentsai/sources
wiki/ai/sources/okfy-purpose-shaped-knowledge-bundles.md
OpenBench — Harness Benchmarks for Coding Agentsai/sources
wiki/ai/sources/openbench-harness-benchmark.md
OpenSpec + Superpowers + gstack:一套让 AI 从「写代码」到「做项目」的组合拳ai/sources
wiki/ai/sources/OpenSpec+Superpowers+gstack-让AI从写代码到做项目.md
OptMem — permanent append-only memory for AI agentsai/sources
wiki/ai/sources/optmem-permanent-agent-memory.md
PERFOPT-Bench — Evaluating Coding Agents on Software Performance Optimizationai/sources
wiki/ai/sources/perfopt-bench-performance-optimization-agents.md
PM Manager — local project governance skill packai/sources
wiki/ai/sources/pm-manager-local-project-governance.md
PawBench:Model × Harness 共评测的 Agent Benchmarkai/sources
wiki/ai/sources/pawbench-model-harness-coevaluation.md
Portcullis — Claude Code security hooksai/sources
wiki/ai/sources/portcullis-claude-code-security-hooks.md
Preloop — Open-Source AI Agent Control Planeai/sources
wiki/ai/sources/preloop-agent-control-plane.md
Proctor — signed benchmark integrity bundlesai/sources
wiki/ai/sources/proctor-signed-benchmark-integrity-bundles.md
Proof-or-Stop:证据门控的 agent 生命周期控制ai/sources
wiki/ai/sources/proof-or-stop-evidence-gated-lifecycle-control.md
Prospective multi-pathogen disease forecasting usiai/sources
wiki/ai/sources/Prospective_multi-pathogen_disease_forecasting_usi.md
RimZ — terminal-native control room for coding agent fleetsai/sources
wiki/ai/sources/rimz-agent-fleet-control-room.md
SAIL Skill — Secure AI Lifecycle as an agent skillai/sources
wiki/ai/sources/sail-skill-secure-ai-lifecycle.md
SETA — Scaling Environments for Terminal Agentsai/sources
wiki/ai/sources/seta-scaling-environments-terminal-agents.md
SWE-Review — closing coding-agent loops with agentic code reviewai/sources
wiki/ai/sources/swe-review-agentic-code-review.md
Sequoia Ascent 2026 — Software 3.0, Agentic Engineering, and Jagged Intelligenceai/sources
wiki/ai/sources/karpathy-sequoia-ascent-2026.md
Set-shifting Behavioral Test for Harnessed Agentsai/sources
wiki/ai/sources/set-shifting-behavioral-test-harnessed-agents.md
SimpleEnglish — controlled-language agent skill for unambiguous documentationai/sources
wiki/ai/sources/simpleenglish-controlled-language-agent-skill.md
Sitegeist — visual diversity benchmark for coding agentsai/sources
wiki/ai/sources/sitegeist-visual-diversity-benchmark.md
Skills Are Not Islands — Agent Skill Supply Chainsai/sources
wiki/ai/sources/agent-skill-supply-chains.md
Stop Means Stop:agent framework 控制原语的执行缺口ai/sources
wiki/ai/sources/stop-means-stop-control-primitives.md
The Unreasonable Effectiveness of RNNsai/sources
wiki/ai/sources/The_Unreasonable_Effectiveness_of_RNNs.md
TraceProbe — coding agent trajectory diagnosticsai/sources
wiki/ai/sources/traceprobe-trajectory-structure-diagnostics.md
Vigiles — audit/lint/test/eval for agent harnessesai/sources
wiki/ai/sources/vigiles-agent-harness-audit.md
Vinv — runtime context bandits for coding agentsai/sources
wiki/ai/sources/vinv-runtime-context-bandits.md
What to Keep, What to Forget — A Rate-Distortion View of Memory Compaction in LLMs and Agentsai/sources
wiki/ai/sources/memory-compaction-rate-distortion-agents.md
Workflow as Knowledge — Semantic Persistence for LLM Workflowsai/sources
wiki/ai/sources/workflow-as-knowledge-semantic-persistence.md
agent-session-io — harness-neutral session substrateai/sources
wiki/ai/sources/agent-session-io-harness-neutral-session-substrate.md
alint — model-backed lint rules for agent-generated codeai/sources
wiki/ai/sources/alint-model-backed-code-analysis.md
coder_eval — evaluate AI coding agents & their skillsai/sources
wiki/ai/sources/coder-eval-skill-evaluation-ci.md
design-harness — evidence board for defensible designai/sources
wiki/ai/sources/design-harness-evidence-board.md
did-it claim-evidence reconciliationai/sources
wiki/ai/sources/did-it-claim-evidence-reconciliation.md
halu-core — claim-grounded execution and reporting honesty benchmark engineai/sources
wiki/ai/sources/halu-core-claim-grounded-agent-evaluation.md
hermes-skill-loop — closed skill learning loopai/sources
wiki/ai/sources/hermes-skill-loop-closed-skill-learning-loop.md
loop-board — autonomous PR task board with constraintsai/sources
wiki/ai/sources/loop-board-autonomous-pr-task-board.md
mcp-gauntlet:面向 MCP server 的 agentic 评估 harnessai/sources
wiki/ai/sources/mcp-gauntlet-agentic-mcp-server-evaluator.md
octopus-skill — host-agnostic long-horizon agent disciplineai/sources
wiki/ai/sources/octopus-skill-long-horizon-agent-discipline.md
quorum — MCP agent collaboration spineai/sources
wiki/ai/sources/quorum-mcp-agent-collaboration-spine.md
redstamp:确定性 agent tool-call firewallai/sources
wiki/ai/sources/redstamp-deterministic-agent-firewall.md
token-diet — always-on token-efficiency skill for coding agentsai/sources
wiki/ai/sources/token-diet-token-efficiency-skill.md
x-clipper 是一个浏览器扩展工具,通过 claude -p headless 加 tweetai/sources
wiki/ai/sources/x-clipper_是一个浏览器扩展工具通过_claude_-p_headless_加_tweet.md
通过累积行为规则让 Coding Agent 跨会话自我改进ai/sources
wiki/ai/sources/self-improving-coding-agents-behavioral-rules.md
X / Karpathy Radar Daily Ingest Templateai/templates
wiki/ai/templates/x-karpathy-radar-daily-ingest.md